About Red-Teaming at FAR.AI
FAR.AI’s red team is building toward a simple outcome: materially raising the bar for safety and security of the most widely deployed and capable AI systems in the world. We intend to be the tip of the spear in AI safety: the team that consistently finds the failures others miss, resulting in real mitigations, and setting the standard that labs and governments converge on. We also leverage our in-depth understanding of weaknesses in frontier models to advise frontier developers on mitigations, to guide our own research and grantmaking for improving model security, and to inform the public of key AI risks.
We are already one of the leading independent red-teaming organizations. Our work has helped most Western frontier model developers improve safeguards through pre- and post-deployment testing (e.g., we have directly influenced safeguards at major frontier developers like OpenAI and Anthropic), and are increasingly embedded in high-leverage government efforts (e.g., leading a consortium building CBRN evaluations for the European Commission/EU AI Office, and collaborating with the UK AI Security Institute).
“FAR.AI’s pre-deployment testing of GPT-5 series models identified failure modes and mitigations, improving the security of our model releases.” – Senior Technical Program Manager, OpenAI
“FAR.AI have been a trusted and thoughtful collaborator for us, and they have progressed the state of frontier red-teaming through research like STACK. We expect this to be a high impact role and are excited to explore collaborations with the successful candidate.” – Xander Davies, Technical Lead, Red Team at UK AISI
In 2026, we are scaling from a strong team with standout wins into a new level of impact for any AI red team globally:
•
Red-teaming all major frontier model releases (closed and open-weight) within days/weeks of release;
•
Expanding strategic engagements with governments and conducting pre-deployment testing with most frontier labs;
•
Deepening our testing of key risk areas like CBRN, cyber, and agents, and exploring new ones like AI control and alignment;
•
Building tools, agents, and insights that raise the global standard for red-teaming.
As engineering manager on the FAR.AI Red Team, you will be the senior technical owner of our engineering, reporting to Kellin Pelrine with a dotted line to Edward Yee. You will build the engineering team and the systems to test critical safeguards of the next generation of AI models. Success looks like vulnerabilities fixed, safeguards strengthened, and global standards shifted – working with leading frontier labs and governments globally to make this happen. Systems your team builds will enable our high-stakes engagements, accelerate cutting-edge AI security research, and create the tools/products/services that let us and our partners red-team the frontier continuously and comprehensively. The current red team will expand in size significantly, and your engineering team will lay the scalable foundation for a relentless and high velocitydivision to make advanced AI systems safer.
We expect this role to start with a small team and include both substantial hands-on engineering and substantial teambuilding and management, e.g., 40% hands-on 60% management. The balance will increasingly evolve towards the management side as we scale the team. In practice, this role spans: