ClickHouse is the core piece of infrastructure at PostHog. Every product and customer relies on it to ingest, store, and query data.
We need someone to automate, manage, and maintain ClickHouse as we grow towards capturing trillions of events per year and having one of the world’s largest clusters.
This includes ClickHouse operations and scaling infrastructure, as well as node and instance-level performance optimization. We want to ensure that we have the right hardware deployed at the right time for each workload on ClickHouse.
You’ll build systems and automations for the provisioning and scaling of our large ClickHouse clusters, handling over 100 PB’s of data. You’ll have the ability to investigate and experiment using the latest hardware that cloud providers have to offer in order to find the optimal setup for our solution. And yes, You’ll have a budget to do this.
You’ll be using Terraform, Ansible, and Kubernetes to automate the dynamic provisioning of instances and work on a bleeding edge ClickHouse implementation, like open format backed tables, and not just maintenance.
We’re also building a query optimizer for ClickHouse, which means you will work on query performance tooling.