Strategic Projects Lead

Patronus AI

San Francisco (CA)

On-site

USD 150,000 - 210,000

Full time

7 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Patronus AI, a frontier lab advancing simulation research for human-aligned AGI, seeks a Strategic Projects Lead in San Francisco. You will own end-to-end delivery of high-quality simulations that drive training, evaluation, and deployment of frontier models.

You will lead a team, manage customer alignment, and shape reward design, QA, and tooling to ensure robust environments. Bases in SF, in-office 5 days a week.

Qualifications

  • BS, MS, or equivalent in CS/ML/Engineering
  • Experience in human data space working on RL envs and task design
  • Strong technical fluency: coding, data analysis, and AI tool usage
  • Excellent organization, execution, and cross-functional coordination
  • Attention to edge cases, inconsistencies, and failure modes
  • Clear written and verbal communication, translating customer needs to tech requirements
  • Integrity and respect for others

Responsibilities

  • Lead end-to-end delivery of high-quality simulations and environments for real-world workflows.
  • Maintain delivery timelines, identify gaps, risks, and blockers, and communicate progress weekly.
  • Serve as primary delivery interface with customers, aligning on tasks and incorporating feedback.
  • Collaborate with technical, research, and QA leads to prioritize what to build or review.
  • Define and uphold objective quality standards for simulations and environments; build QA tooling.
  • Analyze model behavior, task quality, and failure modes to improve reward design and QA processes.
  • Turn strategic project learnings into scalable systems, processes, and tooling.

Job description


Patronus AI is a frontier lab developing simulation research and infrastructure to accelerate progress toward human-aligned AGI. We are on a mission to simulate all of the world’s intelligence.


We are the team behind some of the earliest and most influential research in AI evaluation like FinanceBench, Lynx, SimpleSafetyTests, CopyrightCatcher, Humanity’s Last Exam, and more. We are formerly AI researchers and engineers from companies like Meta AI, Amazon AGI, and Google. Our customers include foundation model labs and Fortune 500 enterprises like Adobe. We are backed by top-tier investors like Lightspeed Venture Partners, Notable Capital, Stanford University, Noam Brown, Gokul Rajaram, and more.


Responsibilities

As a Strategic Projects Lead at Patronus AI, you will lead the delivery of high-quality simulations that define how AI systems are trained, evaluated, and improved. You will work at the intersection of reinforcement learning, scalable oversight, and real-world workflow simulation, building environments and simulation data that directly influence how frontier models are developed, stress-tested, and deployed.


This is a highly autonomous role. You will lead a team in building simulations of impactful real-world workflows, owning project execution, quality standards, and customer alignment from requirements through delivery. You will work across reward design, tool simulations, behavior analysis, QA processes, and automated review tooling, helping set the standard for robust, high-quality environments.


Your work will inform how frontier labs design, train, and improve the next generation of agents for long-horizon tasks, progressing our path toward safe, human-aligned general intelligence.


In this role, you will:



  • Lead end-to-end delivery of high-quality simulations and environments for impactful real-world workflows, from translating customer requirements into concrete build targets through final delivery.

  • Stay close to the details and maintain a clear view of open tasks, delivery gaps, quality risks, technical blockers, and timeline confidence.

  • Serve as the primary delivery interface with customers, aligning on sample tasks, incorporating feedback, communicating progress, and managing expectations week to week.

  • Partner with technical, research, and QA leads to prioritize what to build, improve, or review next based on delivery needs, model behavior, quality gaps, and customer requirements.

  • Define and uphold objective quality standards for simulations, tasks, environments, and deliverables, ensuring each shipment meets requirements for correctness, difficulty, diversity, volume, and readiness. Build and maintain automated QA and review tooling that upholds these standards.

  • Analyze model behavior, task quality, and failure modes to understand what separates strong environments from weak ones, then translate those insights into better reward design, task generation, QA processes, and review workflows.

  • Turn learnings from strategic projects into scalable systems, processes, and tooling that improve how Patronus builds, evaluates, and delivers simulation environments for frontier AI systems.


Qualifications

“The number one qualification to succeed in this machine learning course is gumption” - John Lafferty, CS Professor at Yale


Above all, we look for a proactive mindset, willingness to learn, relentless drive, and passion for engineering and product. You are a great fit if you have a background in the following:



  • BS, MS, or equivalent experience in Computer Science, Machine Learning, Engineering, Mathematics, or another technical / quantitative field.

  • >1 year experience in the human data space working on RL envs, task design at a top tier company.

  • Strong technical fluency, including comfort using AI tools, writing or reviewing code, and analyzing data or model outputs.

  • Excellent organization and execution skills, with the ability to manage tasks, timelines, quality reviews, customer requirements, and cross-functional stakeholders.

  • Strong eye for quality and detail, with a bias toward catching edge cases, inconsistencies, and subtle failure modes.

  • Clear written and verbal communication skills, including the ability to translate customer needs into concrete technical requirements.

  • Good character, integrity, and respect for others!


To support close collaboration, this role is based in our San Francisco headquarters and requires in-office attendance 5 days a week.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Research
Member of Technical Staff - Research

Patronus AI • San Francisco (CA)

On-site
USD 180,000 - 260,000
Technical Program Manager
Technical Program Manager

Patronus AI • San Francisco (CA)

On-site
USD 125,000 - 250,000
Competitive salary and equity packages
15 days of paid vacation per annum
Parental & sick leave
+9
Member of Technical Staff - Engineering
Member of Technical Staff - Engineering

Engg • San Francisco (CA)

On-site
USD 190,000 - 280,000
Expert Operations Manager
Expert Operations Manager

Patronus AI • San Francisco (CA)

On-site
USD 150,000 - 190,000
Lead, Strategic AI Simulations & Delivery
Lead, Strategic AI Simulations & Delivery

Patronus AI • San Francisco (CA)

On-site
USD 150,000 - 210,000
Technical Recruiter
Technical Recruiter

Patronus AI • San Francisco (CA)

On-site
USD 120,000 - 180,000
Staff Engineer, AI Simulation & Infrastructure
Staff Engineer, AI Simulation & Infrastructure

Engg • San Francisco (CA)

On-site
USD 190,000 - 280,000
AI Engineer
AI Engineer

Valsoft Corporation • United States

On-site
USD 140,000 - 230,000
AI Engineer
AI Engineer

Valsoft Corporation • Northern (KY)

Hybrid
USD 120,000 - 180,000
Forward Deployed AI Strategy Lead
Forward Deployed AI Strategy Lead

Prime Intellect • San Francisco (CA)

On-site
USD 180,000 - 260,000
Competitive cash compensation
Meaningful equity
Visa sponsorship and relocation
+3