As a Member of Technical Staff in the Safety for Agents team, you will make a meaningful impact on the development of better, fairer, more trustworthy, and more secure Large Language Models (LLMs). Your primary focus will be on data generation, post-training algorithms, and evaluation methods to ensure Safety in the next generation of models that can access external resources and take actions in the world. You will work closely with other cross-functional machine learning teams and data annotation teams, and will also collaborate with product and policy teams.
This role combines expertise in machine learning, ethical and responsible AI, experimental design, and data generation and management.
It will require curiosity to tackle totally new scientific problems, engineering skills to implement the pieces we need to test solutions to these, and a desire to dive into messy data and results. You will be on a small team with a lot of autonomy and decision-making power, responsible for making the next generation of LLMs better for society as a whole.
Please Note: The existing team work in offices in London, Edinburgh, Paris, Toronto, and New York, but we also embrace being remote-friendly! For this role you need to have ~50% working day overlap with UK/EU timezone (e.g. US East is fine) but there are otherwise no restrictions on where you can be located for this role.