Our mission is to underwrite superintelligence. We red-team, certify, and insure agents, e.g. ElevenLabs, UiPath, Intercom. Soon, agents will take on critical work in hospitals, banks and governments. This could create abundance - or go terribly wrong. History teaches us that standards and insurance are capitalism’s best tools for making risk legible and shaping incentives. We’ve built the gold standard for agent security & reliability, AIUC-1, with 100+ CISOs from the Fortune 1000. Our team comes from Anthropic, McKinsey, METR, Weights & Biases, Perplexity, Thiel Fellowship, etc. We’ve raised $55m from investors like Ribbit, First Harmonic and Nat Friedman. Read more here, here, and here.
The role
You’ll lead the red-teaming of the world’s most used AI agents, working directly with our customers’ agents end-to-end from scoping to certification.
•
Refine the core engine that runs all our red-teaming evals