The Infrastructure as a Service team is at the center of one of Algolia’s most consequential engineering transformations.
For years, Algolia has operated a production fleet of approximately 4,000 bare-metal servers to deliver the reliability, low latency, and scalability that our customers expect. We are now building the foundations of a unified cloud and Kubernetes platform designed to support Algolia’s growth for years to come.
This is not a lift-and-shift project. It is an opportunity to rethink how Algolia provisions, secures, operates, observes, upgrades, and scales production infrastructure and to build it as a platform that engineers can safely consume, rather than a queue of manual requests.
As a Senior Site Reliability Engineer in IaaS, you will help shape the next generation of Algolia’s production infrastructure.
You will lead major parts of the Cloud Baseline and the reliable lifecycle capabilities that enable teams to operate and migrate workloads safely on a cloud-native platform. You will work across cloud foundations, Kubernetes, platform engineering, automation, reliability, and large-scale production operations.
This role is for an engineer who enjoys solving infrastructure problems where the answer must work not once, but hundreds or thousands of times: creating repeatable cloud environments, enabling a growing fleet of production clusters, reducing manual operations, and maintaining the reliability and cost efficiency our customers expect throughout the transition.