AI Safety Expert (Seattle or Boston)

mpathic

Seattle (WA)

On-site

USD 70,000 - 110,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

mpathic.ai in Seattle is seeking AI Safety Experts for a temporary project to evaluate and improve the safety, reliability, and real-world behavior of frontier AI systems. This role suits professionals with expertise in human behavior, policy, education, healthcare, or trust & safety who can exercise sound judgment.

You will evaluate conversations, rate outputs with structured rubrics, annotate data for training and benchmarking, and identify risks and improvement opportunities while maintaining

Qualifications

  • Professional experience in psychology, behavioral science, trust & safety, or related research.
  • Strong written communication and attention to detail.
  • Able to learn structured evaluation frameworks and apply them consistently; strong critical thinking.
  • High ethical standards and sound judgment with sensitive content.
  • Willingness to sign NDAs and work on confidential projects.
  • Availability to work on-site in Seattle, San Francisco, or Boston.

Responsibilities

  • Evaluating AI-generated conversations, responses, and reasoning for quality, safety, and usefulness
  • Rating model outputs using structured evaluation rubrics and project guidelines
  • Annotating conversational data to support AI training and benchmarking
  • Identifying emerging risks, behavioral patterns, and opportunities for model improvement
  • Providing written feedback that helps researchers and engineers improve model performance
  • Maintaining strict confidentiality while working with proprietary AI systems and sensitive content
  • Participating in calibration sessions and quality reviews to ensure consistent evaluations

Skills

Psychology
Behavioral science
Trust & Safety
Research
Critical thinking
Written communication

Tools

Google Workspace
Slack

Job description

About mpathic.ai

mpathic is keeping humans safe in the AI era through automated tools and expert datasets that are rooted in psychology and powered by clinicians. We are a series A start-up backed by Tier 1 investors including Foundry.vc and Next Frontier Capital.

About the role

mpathic is seeking AI Safety Experts for a temporary project to support confidential projects evaluating and improving the safety, reliability, and real-world behavior of frontier AI systems.

This role is ideal for professionals with expertise in human behavior, communication, policy, education, healthcare, technology, trust & safety, or other domains where judgment, critical thinking, and nuanced decision-making matter. You'll help identify model strengths and weaknesses, uncover failure modes, and provide the high-quality human feedback that makes AI systems safer and more useful.

What You'll Be Doing
  • Evaluating AI-generated conversations, responses, and reasoning for quality, safety, and usefulness
  • Rating model outputs using structured evaluation rubrics and project guidelines
  • Annotating conversational data to support AI training and benchmarking
  • Identifying emerging risks, behavioral patterns, and opportunities for model improvement
  • Providing written feedback that helps researchers and engineers improve model performance
  • Maintaining strict confidentiality while working with proprietary AI systems and sensitive content
  • Participating in calibration sessions and quality reviews to ensure consistent evaluations
What We're Looking For

Successful candidates are curious, analytical, thoughtful communicators who enjoy solving complex problems and exercising sound judgment. They are comfortable evaluating nuanced situations, following detailed guidelines, and contributing to the development of trustworthy AI.

Basic Qualifications
  • Professional experience or subject matter expertise in a relevant field such as psychology, behavioral science, social work, trust & safety, research, or a related discipline
  • Strong written communication skills with excellent attention to detail
  • Comfortable learning structured evaluation frameworks and applying them consistentlyStrong critical thinking and problem-solving skills
  • High ethical standards and sound judgment when working with sensitive or ambiguous content
  • Comfortable using AI tools, Google Workspace, Slack, and other web-based collaboration platforms
  • Willingness to sign NDAs and work on confidential projects
  • Availability to work on-site in Seattle, San Francisco, or Boston
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safety Evaluator for Real-World AI Systems
AI Safety Evaluator for Real-World AI Systems

mpathic • Seattle (WA)

On-site
USD 70,000 - 110,000
Cybersecurity for AI Safety (Contract)
Cybersecurity for AI Safety (Contract)

Empathy • Northern (KY)

Hybrid
USD 82,656 - 165,312
AI Safety Specialist - Fully Remote | Upto $70/hr
AI Safety Specialist - Fully Remote | Upto $70/hr

Obsidian • San Francisco (CA)

Remote
USD 120,000 - 180,000
Financial Risk & Safety Specialist (AI Systems)
Financial Risk & Safety Specialist (AI Systems)

Next Frontier Capital • Seattle (WA)

On-site
USD 41,328 - 275,520
Cybersecurity for AI Safety (Contract)
Cybersecurity for AI Safety (Contract)

Next Frontier Capital • United States

On-site
USD 103,000 - 165,000
Financial Risk & Safety Specialist (AI Systems)
Financial Risk & Safety Specialist (AI Systems)

Portland Seed Fund • Seattle (WA)

Hybrid
USD 41,000 - 276,000
Remote AI Safety & Mental Health Expert (Contract)
Remote AI Safety & Mental Health Expert (Contract)

mpathic • Seattle (WA)

On-site
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Obsidian • San Francisco (CA)

On-site
USD 140,000 - 210,000
QA Reviewer & Project Coordinator (On-Site - Seattle or Boston)
QA Reviewer & Project Coordinator (On-Site - Seattle or Boston)

mpathic • Seattle (WA)

On-site
USD 90,000 - 140,000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • San Francisco (CA)

On-site
USD 150,000 - 190,000