Research Manager

Center for Ai Safety

San Francisco (CA)

On-site

USD 170,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401K plan + 4% matching
Unlimited PTO
Lunch and dinner at the office
Annual Professional Development Stiped

Job summary

Center for AI Safety in San Francisco, CA seeks a Research Manager to unblock hard technical problems, manage a growing research team, and translate strategic direction into day‑to‑day execution. The role centers on empirical deep learning research with LLMs and multimodal models.

You will lead efforts, mentor researchers, and contribute to papers while shaping a strong team culture and research agenda. This is a senior, impact-driven leadership position.

Qualifications

  • Can get up to speed quickly on literature and track latest developments.
  • Can propose new, promising experiments drawing on philosophy and AI safety context.
  • Can reason precisely about subtle conceptual distinctions like AI deception or AI wellbeing.
  • Has a track record of strong published research.

Responsibilities

  • Solve hard problems and unblock the team; diagnose failures and propose a path forward.
  • Manage research performance; provide honest feedback and coach researchers.
  • Set and maintain priorities; translate the Director's vision into concrete team priorities.
  • Drive planning and accountability; own project management with clear owners and deadlines.
  • Build a strong team culture; shape norms, coordinate meetings, and ensure information flow.
  • Write and contribute to papers, and help hire researchers and interns.

Skills

Research taste
Research execution
People management
Project management
Communication

Job description

Research Manager
About the role

The Center for AI Safety (CAIS) is a leading research and advocacy organization focused on mitigating societal-scale risks from AI. We address the toughest challenges in AI safety through technical research, field-building initiatives, and policy engagement, along with our sister organization, Center for AI Safety Action Fund.

What distinguishes us is what we choose to work on. Our work is aimed at reducing real-world risks from advanced AI systems. We deliberately pursue research directions that the field is not yet paying attention to, and we move on once the rest of the field catches up. Our focus is on problems that are both highly important and highly neglected—and our track record is built on getting to them first.

  • In 2022–2023, we focused on AI honesty, robustness, transparency, and trojan/backdoor behaviors.
  • In 2023–2024, we turned to malicious use and weaponization capabilities, introducing the first state-of-the-art benchmarks for measuring it.
  • More recently, we've been working on AI value systems and the functional well-being of AI systems.

This is a research philosophy as opposed to a fixed agenda: we go where the important, unworked problems are. Because our work tends to be timely and to open up territory rather than crowd into it, our papers have repeatedly gone on to become widely cited and to set the standard for underexplored research areas. Our work is regularly used by AI safety institutes and frontier AI labs, and they have shaped real policy outcomes, including being presented directly to senators and policymakers.

About the role

You will act as a senior researcher who can unblock hard technical problems, a people manager who can develop and hold a team accountable, and a translator who bridges the Director's strategic direction and the team's day-to-day execution. The best person for this role combines research taste with the operational discipline to make a team reliably deliver.

Our work centers on empirical deep learning research with large language models and/or multimodal models.

What you'll do
  • Solve hard problems and unblock the team. Step in when researchers get stuck. Diagnose what's going wrong in an experiment or a research direction, propose a path forward, and get people moving again. You've been senior long enough to have seen these failure modes before.
  • Manage research performance. Continuously evaluate the quality and pace of the team's work. Give honest, well-calibrated feedback, coach researchers toward higher standards, and surface who is delivering and who needs support.
  • Set and maintain priorities. Translate the Research Director's vision into concrete priorities for the team, and keep the team's work aligned to them. Push back on low-value side quests and scattered directions; make sure effort flows to what matters most.
  • Drive planning and accountability. Own project management for the team. Make sure tasks have clear owners and deadlines, that work is actually completed at the quality bar, and that slippage and dropped work are caught and corrected rather than quietly tolerated.
  • Build a strong team culture. Shape team norms, resolve conflict, support morale and engagement, run an effective meeting cadence, and keep information flowing well across the team.
  • Write and contribute to papers, and help hire researchers and interns as the team grows.
What we're looking for
  • Research taste. You can tell good research directions from bad ones, detect when an argument or result doesn't hold up, and know what to prioritize. Concretely, that means you:
  • Can get up to speed quickly on a literature and track the latest developments.
  • Can propose new, promising experiments and subdirections (often drawing on philosophy, concepts from other disciplines, and AI safety context).
  • Can reason precisely about subtle conceptual distinctions—the kind of verbal and philosophical clarity needed to handle concepts like AI deception or AI wellbeing rigorously.
  • Have a track record of strong published research.
  • Research execution. You can validate experiments, reproduce or pressure-test results, and find the flaws in more junior researchers' work.
  • People management. Ability to manage performance, coach people to improve, and build healthy team culture and collaboration norms.
  • Project management. Organized, and able to keep work on track and people accountable. Prior project-management experience helps but isn't essential if you're naturally orderly and conscientious.
  • The basics. Sharp, conscientious, and a strong communicator who can translate effectively between leadership and the team. A genuine collaborator—low ego and motivated by the mission of making AI safe.

Not sure you're a "manager"? Apply anyway. If you're a PhD student or someone with a strong research record who hasn't formally managed a team but is conscientious, organized, and has mentored or guided people before, this is a profile we're genuinely excited about. We weight research ability most heavily, and we'd rather find a great researcher we can grow into this role than someone with a management title and weaker research instincts.

Nice to have
  • Significant AI safety context and familiarity with the field's open problems.
  • Prior experience collaborating with our Director or others on the team.
Compensation:

$170,000 - $260,000 a year

Benefits:
  • Health insurance for you and your dependents
  • 401K plan + 4% matching
  • Unlimited PTO
  • Lunch and dinner at the office
  • Annual Professional Development Stipend
  • Access to some of the top talent working on technical and conceptual research in AI safety

The Center for AI Safety is an Equal Opportunity Employer. We consider all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, ancestry, age, disability, medical condition, marital status, military or veteran status, or any other protected status in accordance with applicable federal, state, and local laws. In alignment with the San Francisco Fair Chance Ordinance, we will consider qualified applicants with arrest and conviction records for employment.

If you require a reasonable accommodation during the application or interview process, please contact contact@safe.ai.

We value diversity and encourage individuals from all backgrounds to apply.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Manager
Research Manager

Aisafety • San Francisco (CA)

On-site
USD 170,000 - 260,000
Health insurance
401K plan + 4% matching
Unlimited PTO
+2
Research Engineer/Scientist
Research Engineer/Scientist

Center for Ai Safety • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health insurance for you and your dep.
401K plan + 4% matching
Unlimited PTO
+2
Research Engineer/Scientist
Research Engineer/Scientist

AI Safety, Inc • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health insurance
401K plan + 4% matching
Unlimited PTO
+2
Principal, Special Projects
Principal, Special Projects

Center for Ai Safety • San Francisco (CA)

On-site
USD 150,000 - 250,000
Health insurance
401K plan + 4% matching
Unlimited PTO
+2
Research Manager, AI Safety
Research Manager, AI Safety

SwiftCruit • Cambridge (MA)

On-site
USD 100,000 - 145,000
5% 403(b) match contribution
Comprehensive health insurance
Generous PTO policy
+3
Researcher, Safety Oversight
Researcher, Safety Oversight

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Research Engineer Intern
Research Engineer Intern

Center for Ai Safety • San Francisco (CA)

On-site
USD 9,000 - 19,000
Stipend
AI Safety Research Manager | Fellowship Programs
AI Safety Research Manager | Fellowship Programs

cbai • Cambridge (MA)

On-site
Researcher, Robustness & Safety Training
Researcher, Robustness & Safety Training

OpenAI • Los Angeles (CA)

On-site
USD 310,000 - 460,000
Equity offers
Competitive salary
Diverse work culture
Principal, Special Projects
Principal, Special Projects

Aisafety • San Francisco (CA)

On-site
USD 120,000 - 210,000