We were founded in the US and have our home there, but our team is distributed across Europe and North America. We get our fix of in-person collaboration (and croissants) in Paris each month for 3 days, always Monday-Wednesday, with an open invitation to stay the whole week. We also do longer off-sites once a year.
Our team is a multidisciplinary blend of research, engineering, and business experts. What unites us is our deep care for what we build together. We’re in a race that requires hard work, intellectual curiosity, and obsession; to balance this intensity, we’ve assembled a team of low ego and kind-hearted individuals who have built the special culture Poolside has. By building collaboratively and with intention, we create a compounding effect that moves the entire company forward towards our mission: reaching AGI through intelligence systems built for software development.
You’ll be working in the compute team focusing on GPU workload scheduling and inference serving optimization. You would partner with the inference team to improve our inference throughput and latency for evals and reinforcement learning. You would collaborate with our scalability team to focus on stabilizing our large scale fault tolerant training. You would also be in close contact with the infra team to make sure our GPU nodes are all healthy and fully utilized.
We are one of the key teams to improve the research velocity. Any improvement on our systems has a wide impact on researchers and can contribute to the poolside mission on building a frontier model.
To optimize GPU utilization across the company and to deliver stable and scalable inference serving stack for Poolside’s researchers.