Site Reliability Engineer - ClickHouse

Posthog · Remote (EMEA) · FullTime

Apply on company site

About PostHog

Product development used to mean manually writing code, running analysis, diagnosing bugs, and rolling out changes using dozens of tools.

PostHog is the only platform that acts like a co-pilot for you (and your AI agents) to do it all – autonomously.

We started with open-source product analytics, launched out of Y Combinator's W20 cohort. We've since shipped more than a dozen products, including:

We are:

  1. Product-led. More than 450,000 organizations have installed PostHog, mostly driven by word-of-mouth. We have intensely strong product-market fit.

  2. Default alive. Revenue is growing incredibly quickly, and we're very efficient. We raise money to push ambition and grow faster, not to keep the lights on.

  3. Well-funded. We've raised more than $180m from some of the world's top investors. We're set up for a long, ambitious journey.

We're focused on building an awesome product for end users, hiring exceptional teammates, shipping fast, and being as weird as possible.

 

Things we care about

Who we're looking for

We’re looking for people (EU/UK based) that like deep ownership of production systems, people that are not afraid of working with stateful infrastructure and love working in AWS, VMs, automation, and making messy systems reliable.

In general we seek SRE’s who are:

What you'll be doing

We run one of the largest self-managed ClickHouse installations on AWS, at petabyte scale, and we’re actively preparing it for the next 10–50× of growth. This role sits at the centre of that effort.

You won’t be in a typical “keep the lights on” SRE role. The work is about turning a fast-growing, stateful system into a predictable, well-automated platform. (provisioning, scaling, rebalancing, recovery)
That means reducing operational stress, designing safe automation for data-heavy workloads, and building the tooling and patterns that let the system scale without scaling human effort.

You’ll work on the kind of problems that only show up at large scale (petabytes of data, thousands of cores, constant ingestion).

You should join this team if you like deep ownership of production systems, and are not afraid of working with stateful infrastructure

 

Requirements

You don’t need to be a ClickHouse expert on day one. We’ll teach you the database internals, but you do need to enjoy owning complex infrastructure.

 

We are committed to ensuring a fair and accessible interview process. If you need any accommodations or adjustments, please let us know.

#LI-DNI

Job alert

Get new jobs by email

Save this search and get relevant new jobs when they appear.