An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Patronus AI, a frontier lab advancing simulation research for human-aligned AGI, seeks a Strategic Projects Lead in San Francisco. You will own end-to-end delivery of high-quality simulations that drive training, evaluation, and deployment of frontier models.
You will lead a team, manage customer alignment, and shape reward design, QA, and tooling to ensure robust environments. Bases in SF, in-office 5 days a week.
Patronus AI is a frontier lab developing simulation research and infrastructure to accelerate progress toward human-aligned AGI. We are on a mission to simulate all of the world’s intelligence.
We are the team behind some of the earliest and most influential research in AI evaluation like FinanceBench, Lynx, SimpleSafetyTests, CopyrightCatcher, Humanity’s Last Exam, and more. We are formerly AI researchers and engineers from companies like Meta AI, Amazon AGI, and Google. Our customers include foundation model labs and Fortune 500 enterprises like Adobe. We are backed by top-tier investors like Lightspeed Venture Partners, Notable Capital, Stanford University, Noam Brown, Gokul Rajaram, and more.
As a Strategic Projects Lead at Patronus AI, you will lead the delivery of high-quality simulations that define how AI systems are trained, evaluated, and improved. You will work at the intersection of reinforcement learning, scalable oversight, and real-world workflow simulation, building environments and simulation data that directly influence how frontier models are developed, stress-tested, and deployed.
This is a highly autonomous role. You will lead a team in building simulations of impactful real-world workflows, owning project execution, quality standards, and customer alignment from requirements through delivery. You will work across reward design, tool simulations, behavior analysis, QA processes, and automated review tooling, helping set the standard for robust, high-quality environments.
Your work will inform how frontier labs design, train, and improve the next generation of agents for long-horizon tasks, progressing our path toward safe, human-aligned general intelligence.
In this role, you will:
“The number one qualification to succeed in this machine learning course is gumption” - John Lafferty, CS Professor at Yale
Above all, we look for a proactive mindset, willingness to learn, relentless drive, and passion for engineering and product. You are a great fit if you have a background in the following:
To support close collaboration, this role is based in our San Francisco headquarters and requires in-office attendance 5 days a week.