ClickHouse Operations Engineer - PostHog

ClickHouse Operations Engineer

ClickHouse Team

Remote

GMT -3 to GMT -8

About PostHog

Product development used to mean manually writing code, running analysis, diagnosing bugs, and rolling out changes using dozens of tools.

PostHog makes products self-driving. It's the only platform that acts like a co-pilot for you (and your AI agents) to do it all – autonomously.

We started with open-source product analytics, launched out of Y Combinator's W20 cohort. We've since shipped more than a dozen products, including:

We are:

  1. Product-led. More than 450,000 organizations have installed PostHog, mostly driven by word-of-mouth. We have intensely strong product-market fit.
  2. Default alive. Revenue is growing incredibly quickly, and we're very efficient. We raise money to push ambition and grow faster, not to keep the lights on.
  3. Well-funded. We've raised more than $180m from some of the world's top investors. We're set up for a long, ambitious journey.

We're focused on building an awesome product for end users, hiring exceptional teammates, shipping fast, and being as weird as possible.

Things we care about

What you'll be doing

ClickHouse is the core piece of infrastructure at PostHog. Every product and customer relies on it to ingest, store, and query data.

We need someone to automate, manage, and maintain ClickHouse as we grow towards capturing trillions of events per year and having one of the world’s largest clusters.

This includes ClickHouse operations and scaling infrastructure, as well as node and instance-level performance optimization. We want to ensure that we have the right hardware deployed at the right time for each workload on ClickHouse.

You'll build systems and automations for the provisioning and scaling of our large ClickHouse clusters, handling over 100 PB's of data. You'll have the ability to investigate and experiment using the latest hardware that cloud providers have to offer in order to find the optimal setup for our solution. And yes, you'll have a budget to do this.

You'll be using Terraform, Ansible, and Kubernetes to automate the dynamic provisioning of instances and work on a bleeding-edge ClickHouse implementation, like open format backed tables, and not just maintenance.

We're also building a query optimizer for ClickHouse, which means you will work on query performance tooling.

You’ll fit right in if:

If this sounds like you, we should talk.

We are committed to ensuring a fair and accessible interview process. If you need any accommodations or adjustments, please let us know.

Your team's mission and objectives

ClickHouse at PostHog - Mission

The ClickHouse team's mission is to provide a central data store for PostHog that is reliable, fast, cost efficient, and secure.

We run PostHog's self-managed, multi-petabyte ClickHouse fleet.

We provide a platform that is:

Q3 2026 objectives

🧬 Cell-based architecture - (Driver: Rory Shanks)

Large customers can break smaller ones; the blast radius is huge, and changes are hard to test. Splitting things up and isolating them lets us manage all of this far better.

What we'll ship:

📦 JSON datatype in the app - (Driver: Rory Shanks)

We have a lot of duplicated data sitting in the cluster eating disk. Newer storage formats give us more efficient disk usage, and therefore better performance.

What we'll ship:

🔐 ClickHouse is up to date and secure - (Driver: Tommy Gilmore)

New ClickHouse versions bring big wins, so we should be able to upgrade faster. And secrets rotation shouldn't be a painful, high-risk process that can take down the app or silently break a team - secrets need to be swappable at any time with no real risk.

What we'll ship:

🚚 ClickHouse migration system - (Driver: Pawel Szczur)

Our ClickHouse migrations suck, and our topology is very complex. We should figure out a good way of doing this.

What we'll ship:

🗑️ Data lifecycle management - (Driver: Pawel Szczur)

We need to delete data all the time, and our systems for it need to be top-notch.

What we'll ship:

🤖 Self-driving operations - (Driver: Bryan Ciaraldi)

Interrupt-driven ops has the biggest impact on the team's ability to ship: operational toil, alert noise, and "fear of breaking ClickHouse" all track to it. We make routine operations self-driving so we're only interrupted when judgment is actually required.

What we'll ship:

Interview process

We do 2-3 short interviews, then pay you to do some real-life (or close to real-life) work.

  1. Application (You are here) Our talent team will review your application
    We're looking to see how your skills and experience align with our needs.
  2. Culture interview
    30-min video call
    Our goal is to explore your motivations to join our team, learn why you’d be a great fit, and answer questions about us.
  3. Technical interview
    45 minutes, varies by role
    You'll meet the hiring team who will evaluate skills needed to be successful in your role. No live coding.
  4. Culture & Motivation interview
    20 minutes, varies by role
    You have reached the final boss. It's time to chat with one of our Blitzscale team members.
  5. PostHog SuperDay
    Paid day of work
    You’ll meet a few more members of the team and work on an independent project. It's challenging, but most people say it's fun, and we'll pay you $1,000 for your efforts!
  6. Offer
    Pop the champagne (after you sign)
    If everyone is happy, we’ll make you an offer to join us - YAY!

Apply

(Now for the fun part...)

Just fill out this painless form and we'll get back to you within a few days. Thanks in advance!

Submit