Remote AI Safety ML Engineer — Experiments & Evaluation

10a Labs

United States

On-site

USD 130,000 - 200,000

Full time

13 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Fully remote (US-based)
Health benefits
Generous PTO
Professional Development
Conference support

Job summary

10a Labs is seeking a Machine Learning Engineer to design, build, and evaluate advanced ML systems focused on AI safety and model evaluation. The role emphasizes experiments, scalable tooling, and rigorous analysis.

You'll work with engineers, analysts, red teamers, and subject-matter experts to push frontier AI capabilities while ensuring robust, secure deployments in a fully remote US-based environment. Candidates should pair strong ML fundamentals with the ability to turn ambiguous questions

Qualifications

  • 3–5 years of experience in machine learning, research engineering, or a related technical field.
  • Strong Python skills and experience with ML frameworks such as PyTorch or JAX.
  • Hands-on experience training, fine-tuning, or evaluating modern ML models.
  • Strong understanding of experimental design, model evaluation, and quantitative analysis.
  • Familiarity with agentic AI fundamentals, including common harnesses and security risks to AI agents.
  • Experience in one or more of the following: reinforcement learning, NLP/LLMs, computer vision, or multimodal ML.

Responsibilities

  • Design and run ML experiments to evaluate AI systems' capabilities and robustness.
  • Develop and evaluate models across RL, NLP/LLMs, CV, and multimodal ML.
  • Build evaluation pipelines, benchmarks, datasets, and metrics.
  • Train, fine-tune, and evaluate models for safety and security.
  • Develop scalable tooling and infrastructure for ML experiments.
  • Analyze results and translate findings into new experiments.

Skills

Python
PyTorch/JAX
ML Experimentation
Experimental Design
Agentic AI
Reinforcement Learning
NLP/LLMs

Tools

PyTorch
JAX

Job description

10a Labs is seeking a Machine Learning Engineer to design, build, and evaluate advanced ML systems focused on AI safety and model evaluation. The role emphasizes experiments, scalable tooling, and rigorous analysis.

You'll work with engineers, analysts, red teamers, and subject-matter experts to push frontier AI capabilities while ensuring robust, secure deployments in a fully remote US-based environment. Candidates should pair strong ML fundamentals with the ability to turn ambiguous questions

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote ML Engineer - AI Safety & Evaluation Expert
Remote ML Engineer - AI Safety & Evaluation Expert

10a Labs • Seattle (WA)

On-site
USD 130,000 - 200,000
Comprehensive health, dental, and vision coverage
Performance-based annual bonus
Support for professional development and conferences
+1
Senior ML Engineer - Safety & AI Evaluation (Remote)
Senior ML Engineer - Safety & AI Evaluation (Remote)

10a Labs • Los Angeles (CA)

On-site
USD 130,000 - 200,000
Performance-based annual bonus
Comprehensive health, dental, and vision coverage
Support for conferences and continuing education
Remote ML Engineer - Safety & AI Evaluation
Remote ML Engineer - Safety & AI Evaluation

Aisafety • Washington

On-site
USD 130,000 - 200,000
Performance-based annual bonus
Support for professional development
Generous PTO and paid holiday schedule
Remote AI Security Engineer: Build & Test Resilience
Remote AI Security Engineer: Build & Test Resilience

10a Labs • United States

Remote
USD 135,000 - 160,000
Performance-based bonus
Professional development support
Fully remote, US-based
Remote AI Infrastructure & Platform Engineer
Remote AI Infrastructure & Platform Engineer

10a Labs • United States

Remote
USD 110,000 - 160,000
Fully remote, U.S.-based
Performance-based annual bonus
Professional development support (con-
+1
Remote Engineering Manager, Safety & AI Defense
Remote Engineering Manager, Safety & AI Defense

EngineersOfAI • San Francisco (CA), Los Angeles (CA), New York (NY), Chicago (IL)

Hybrid
USD 150,000 - 200,000
Comprehensive Healthcare Benefits
401k with Employer Match
Global Benefit programs
+2
Senior AI Safety Research Engineer (Remote)
Senior AI Safety Research Engineer (Remote)

Safetytalent • San Francisco (CA)

On-site
USD 150,000 - 250,000
Senior Safety ML Engineer - Remote - LLMs & NLP
Senior Safety ML Engineer - Remote - LLMs & NLP

EngineersOfAI • San Francisco (CA), Los Angeles (CA), New York (NY), Chicago (IL)

Hybrid
USD 180,000 - 240,000
Equity in RSUs
Commission (depending on role)
Benefits
Senior Member of Technical Staff - Model Safety
Senior Member of Technical Staff - Model Safety

Xcede • San Francisco (CA)

On-site
USD 130,000 - 160,000
Research Scientist - AI Safety & Scalable ML
Research Scientist - AI Safety & Scalable ML

Center for AI Safety (CAIS) • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health insurance for you and dependets
401K plan + 4% matching
Unlimited PTO
+2