Lead, Product Safety & Wellbeing

Flagship Pioneering, Inc.

Cambridge (MA)

On-site

USD 138,000 - 182,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Flagship Pioneering's FL105 is seeking an experienced leader to own and continually improve product safety performance across AI tools and workflows. You will define safety measurements, build evaluation frameworks, and partner across psychology, product, engineering, and ML teams to translate insights into safer, more effective user experiences.

You will balance safety with speed, lead human and automated evaluations, and shape launch readiness for new features, helping scale wellbeing-focused

Qualifications

  • PhD, PsyD, LCSW, or equivalent experience in psychology, behavioral science, or a related field.
  • Deep expertise in psychological safety, crisis response, and risk assessment.
  • Experience developing evaluation frameworks or QA processes.
  • Strong communication, analytical, and problem-solving skills for AI systems.
  • Ability to collaborate with engineers, researchers, designers, and product managers.
  • Experience with LLMs, conversational AI, or generative AI products.
  • Up to date on current clinical AI best practices and standards.

Responsibilities

  • Serve as the subject matter expert on safety systems, owning the safety roadmap and continuously improving guardrails as the product evolves.
  • Develop a deep, data-driven understanding of product performance across safety risks; identify strengths, limitations, failure modes, emerging risks, and opportunities for improvement.
  • Develop and evolve safety frameworks, benchmarks, evaluation methodologies, rubrics, quality standards, and meaningful safety metrics.
  • Lead human and automated evaluation programs, including annotation programs, annotator and AI judge training and calibration, and improvements to evaluation reliability and scalability.
  • Conduct qualitative and quantitative reviews from both expert and user perspectives, translating findings into clear, actionable product recommendations.
  • Partner with Engineering, ML, Product, and Data Science to improve risk detection, prompting strategies, response generation, routing and escalation logic, automated evaluation pipelines, and the overall user experience.
  • Prototype and test new safety approaches, product features, and user experiences to improve safety and broader product quality.
  • Own ongoing safety monitoring and assessment, including dashboards, key metrics, trend analysis, emerging risk identification, safety reviews, and refinement of safety protocols.
  • Evolve and own safety and launch-readiness criteria for new features and capabilities, identifying risks and ensuring appropriate mitigations are in place.

Skills

PhD/ PsyD / LCSW equivalent
Psychology/behavioral science
Safety risk assessment
LLMs / Generative AI
Communication & collaboration
Data-driven evaluation

Education

PhD, PsyD, LCSW or equivalent

Tools

Annotation programs
Quality assurance processes

Job description

FL105 is a privately held, early-stage company pioneering the use of artificial intelligence to transform how people navigate life’s challenges and wellbeing. We are creating a platform that empowers people to build a life they are proud of and fulfilled by. We bring together psychology, product, design, engineering, and machine learning to build digital experiences that are engaging, trustworthy, and genuinely helpful.

FL105 is backed by Flagship Pioneering, an innovation enterprise that conceives, creates, resources, and builds companies that invent breakthrough technologies that transform the world. Flagship has created over 100 groundbreaking companies since 2000, including Moderna.

The Role

We’re looking for an experienced and creative leader to own and continually improve product performance. This person will lead on how we define and measure safety, and partner across disciplines to strengthen our evaluations, safety systems, and user experience. This role sits at the intersection of psychology, AI, product development, and applied research. Candidates should have a strong background in clinical science with the ability to lead projects and communicate effectively across teams to translate insights into actionable and meaningful improvements across AI tools and workflows. The role offers a unique opportunity to help drive the strategy of an early stage AI-focused company.

The ideal candidate is a self-starter who knows how to balance safety, quality, and rigor with speed and experimentation. They are passionate about creating user experiences that help people feel supported, understood, and empowered in their everyday lives. This dynamic role offers significant opportunities for impact and professional growth, helping shape how AI can responsibly support wellbeing at scale.

Key Responsibilities
  • Serve as the subject matter expert on safety systems, owning the safety roadmap and continuously improving guardrails as the product evolves.
  • Develop a deep, data-driven understanding of product performance across safety risks; identify strengths, limitations, failure modes, emerging risks, and opportunities for improvement.
  • Develop and evolve safety frameworks, benchmarks, evaluation methodologies, rubrics, quality standards, and meaningful safety metrics.
  • Lead human and automated evaluation programs, including annotation programs, annotator and AI judge training and calibration, and improvements to evaluation reliability and scalability.
  • Conduct qualitative and quantitative reviews from both expert and user perspectives, translating findings into clear, actionable product recommendations.
  • Partner with Engineering, ML, Product, and Data Science to improve risk detection, prompting strategies, response generation, routing and escalation logic, automated evaluation pipelines, and the overall user experience.
  • Prototype and test new safety approaches, product features, and user experiences to improve safety and broader product quality.
  • Own ongoing safety monitoring and assessment, including dashboards, key metrics, trend analysis, emerging risk identification, safety reviews, and refinement of safety protocols.
  • Evolve and own safety and launch-readiness criteria for new features and capabilities, identifying risks and ensuring appropriate mitigations are in place.
Qualifications
  • Advanced training (PhD, PsyD, LCSW, or equivalent experience) in psychology, behavioral science, mental health, or a closely related field
  • Deep expertise in psychological safety, crisis response, and risk assessment; strong judgment around conversations involving emotional distress, self-harm, suicide, abuse, or other sensitive situations
  • Experience developing evaluation frameworks or quality assurance processes
  • Excellent communication, analytical, and problem-solving skills with the ability to translate expertise into clear operational and product guidance for AI systems.
  • Ability to collaborate effectively with engineers, researchers, designers, and product managers
  • Ability to make and communicate tradeoffs and clear decisions in ambiguous environments
  • Experience building with LLMs, conversational AI, or generative AI products
  • Up to date on current clinical AI best practices and emerging guidelines and standards.
Preferred Qualifications
  • Experience leading annotation or human evaluation programs
  • Experience developing AI evaluation benchmarks or automated evaluation systems
  • Experience working closely with product or engineering teams
About Flagship

Flagship Pioneering is a bioplatform innovation company that invents and builds platform companies, each with the potential for multiple products that transform human health or sustainability. Since its launch in 2000, Flagship has originated and fostered more than 100 scientific ventures, resulting in more than $90 billion in aggregate value. Many of the companies Flagship has founded have addressed humanity’s most urgent challenges: vaccinating billions of people against COVID-19, curing intractable diseases, improving human health, preempting illness, and feeding the world by improving the resiliency and sustainability of agriculture. Flagship has been recognized twice on FORTUNE’s "Change the World" list, an annual ranking of companies that have made a positive social and environmental impact through activities that are part of their core business strategies, and has been twice named to Fast Company’s annual list of the World’s Most Innovative Companies. Learn more about Flagship at www.flagshippioneering.com.

Flagship Pioneering and our ecosystem companies are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status.

At Flagship, we recognize there is no perfect candidate. If you have some of the experience listed above but not all, please apply anyway. Experience comes in many forms, skills are transferable, and passion goes a long way. We are dedicated to building diverse and inclusive teams and look forward to learning more about your unique background.

Recruitment & Staffing Agencies

Flagship Pioneering and its affiliated Flagship Lab companies (collectively, "FSP") do not accept unsolicited resumes from any source other than candidates. The submission of unsolicited resumes by recruitment or staffing agencies to FSP or its employees is strictly prohibited unless contacted directly by Flagship Pioneering’s internal Talent Acquisition team. Any resume submitted by an agency in the absence of a signed agreement will automatically become the property of FSP, and FSP will not owe any referral or other fees with respect thereto.

Privacy Notice for Applicants

When you apply for a role at Flagship Pioneering or one of its portfolio companies, we collect and use personal information you provide (such as your name, contact details, work history, and application materials) to evaluate your application, communicate with you, and comply with legal obligations. Your application data is processed through Greenhouse, our applicant tracking system, and may also be reviewed using AI-assisted screening tools. We do not sell your personal information. California residents have rights under the CCPA/CPRA including to know, delete, and opt out of the sharing of their personal information. If you are located in the EU or UK, we process your data under GDPR and you have rights to access, rectify, and erase your data. To exercise your rights or for questions, contact privacy@flagshippioneering.com.

The salary range for this role is $138,000 - $181,500. Compensation for the role will depend on a number of factors, including a candidate’s qualifications, skills, competencies, and experience. FL105 currently offers healthcare coverage, annual incentive program, retirement benefits and a broad range of other benefits. Compensation and benefits information is based on FL105's good faith estimate as of the date of publication and may be modified in the future.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal / Senior Software Engineer
Principal / Senior Software Engineer

Flagship Pioneering, Inc. • Cambridge (MA)

On-site
USD 108,000 - 209,000
Healthcare coverage
Annual incentive program
Retirement benefits
+1
Machine Learning Research Engineer
Machine Learning Research Engineer

Flagship Pioneering, Inc. • Cambridge (MA)

On-site
USD 120,000 - 193,000
FL105 | Cambridge, MA Machine Learning Research Engineer
FL105 | Cambridge, MA Machine Learning Research Engineer

Flagship Pioneering • Cambridge (MA)

On-site
USD 120,000 - 192,500
Healthcare coverage
Annual incentive program
Retirement benefits
FL105 | Cambridge, MA Principal / Senior Software Engineer
FL105 | Cambridge, MA Principal / Senior Software Engineer

Flagship Pioneering • Cambridge (MA)

On-site
USD 108,000 - 209,000
Healthcare coverage
Annual incentive program
Retirement benefits
Vice President, AI Research and Real-World Evidence
Vice President, AI Research and Real-World Evidence

Flagship Pioneering Co-Op Program • Cambridge (MA)

On-site
USD 263,000 - 347,000
Healthcare coverage
Annual incentive program
Retirement benefits
Vice President, AI Research and Real-World Evidence
Vice President, AI Research and Real-World Evidence

Flagship Pioneering • Cambridge (MA)

On-site
USD 263,000 - 347,000
Healthcare coverage
Annual incentive program
Retirement benefits
+1
Senior AI-Driven Cloud Security and AppSec Architect
Senior AI-Driven Cloud Security and AppSec Architect

Flagship Pioneering • Cambridge (MA)

On-site
USD 148,000 - 203,500
Healthcare coverage
Annual incentive program
Retirement benefits
FL113 | Cambridge, MA Vice President, AI Research and Real-World Evidence
FL113 | Cambridge, MA Vice President, AI Research and Real-World Evidence

Flagship Pioneering • Cambridge (MA), Northern (KY)

On-site
USD 263,000 - 347,000
Healthcare coverage
Annual incentive programme
Retirement benefits
+1
FL119 | Cambridge, MA (Senior) Scientist, Computational Neuroscience
FL119 | Cambridge, MA (Senior) Scientist, Computational Neuroscience

Flagship Pioneering • Cambridge (MA)

On-site
USD 132,000 - 259,000
Healthcare benefits
Annual incentive program
Retirement benefits
Flagship Pioneering | Cambridge, MA Manager, Visual Design
Flagship Pioneering | Cambridge, MA Manager, Visual Design

Flagship Pioneering • Cambridge (MA), Northern (KY)

On-site
USD 108,000 - 149,000