Senior | Staff Software Engineer - AI / ML Software

Front Door Defense

San Francisco, Northern (CA, KY)

Hybrid

USD 208,000 - 315,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Snorkel AI in San Francisco, CA, is hiring a Senior | Staff Software Engineer for AI/ML to build scalable evaluation pipelines and accelerate frontier-grade data generation at scale. Hybrid work in SF. You will explore how frontier-grade data is produced and evaluated, shape standards, and scale the team’s impact.

The role focuses on efficient evals, routing, and production-ready ML infrastructure with strong emphasis on data-driven quality and cost control.

Qualifications

  • 5+ years building production ML or software systems with end-to-end ownership
  • Hands-on experience running LLM/ML workloads in production
  • Statistics and experimentation: experiment design, hypothesis testing, sampling, confidence intervals
  • Strong Python and software engineering fundamentals
  • Experience designing evaluations and interpreting results rigorously
  • Clear communication with researchers, engineers, and business partners

Responsibilities

  • Efficient agentic evals with adaptive sampling and early stopping
  • AI model routing to cheapest model that meets quality bar
  • Fine-tune and serve open-weight models where they match frontier quality
  • Predictive difficulty estimates for frontier tasks
  • Build golden datasets and measure calibration of LLM-as-judge systems
  • Turn research prototypes into reusable production components

Skills

5+ years building production ML apps
Production ML workloads
Statistics & experimentation
Python & software engineering
Evaluations design & interpretation
Communication with researchers/enginee

Education

MS or PhD in CS/ML/Statistics

Tools

LoRA
Model routing
Open-weight models

Job description

Senior | Staff Software Engineer - AI / ML

Build scalable AI evaluation pipelines to accelerate frontier-grade data generation and validation at scale effectively

Location: San Francisco, California

Compensation: $208,000 - 315,000 USD / year

About The Role
Senior | Staff Software Engineer - AI / ML

San Francisco, CA (Hybrid)

At Snorkel, we believe meaningful AI doesn't start with the model, it starts with the data.

We're on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world's largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. We work with some of the world's largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before.

In September 2026 we raised $350 million Series E at a $3.5 billion valuation, and we are scaling our engineering and research teams to meet demand.

The Role

Frontier AI data is expensive to make and hard to measure. Every task we deliver is tested against the strongest models, often through many long-running agent rollouts. Your job is to make that process faster, cheaper, and more rigorous with ML and AI

You will be one of the early members of ML & Research Engineering at Snorkel. You will study how frontier-grade data is generated and evaluated, form hypotheses, validate them against real production data, and ship the winners at scale. You will shape the discipline's direction, its standards, and the team that grows around it.

What You'll Work On
  • Efficient agentic evals. Cut the cost of long-horizon agent evaluation with adaptive sampling, statistically grounded early stopping, model cascades, caching, and cheap-first gating.
  • AI model routing. Route every eval and judge call to the cheapest model that clears the quality bar, with fallback, monitoring, and cost attribution.
  • Fine-tuned small models. Fine-tune and serve open-weight models (LoRA and other parameter-efficient methods) where they match frontier quality, and know when they don't.
  • Predictive difficulty. Build models that estimate how hard a task is for frontier systems before running a single rollout.
  • Measurement for AI data. Build golden datasets, quantify the accuracy and calibration of LLM-as-judge systems, and make quality reproducible across projects.
  • Research to production. Turn research prototypes into reusable, configurable components that forward deployed engineers and researchers use on every project.
What You'll Bring
  • 5+ years building production ML or software systems, with end-to-end ownership from prototype to production
  • Hands-on experience running LLM or ML workloads in production, and comfort reasoning about non-deterministic systems
  • Deep grounding in statistics and experimentation: experiment design, hypothesis testing, sampling, and confidence intervals
  • Strong Python and software engineering fundamentals, including testing, code review, and system design
  • Experience designing evaluations and interpreting results rigorously
  • A habit of finding high-impact problems before they are assigned, and clear communication with researchers, engineers, and business partners
Nice To Have
  • Fine-tuning and serving open-weight models, and judging when a smaller model meets the quality bar
  • Building LLM evaluation or experimentation platforms, model gateways, or routing systems
  • Experience with agentic workloads, benchmarks, or RL environments
  • A record of taking research into production: publications, open-source work, or shipped research-driven features
  • MS or PhD in Computer Science, Machine Learning, Statistics, or a related field
Why Join Now
  • Frontier problems. Measure and shape the tasks designed to challenge the strongest models in the world.
  • All AI, no plumbing. Every problem on this team is an open ML or LLM problem.
  • Founding impact. Help define what ML engineering means at Snorkel and grow the team that carries it forward.
  • Visible results. Your work shows up directly in the speed, quality, and cost of the data frontier AI is built on.

Actual compensation will be determined based on factors including skills, qualifications, experience, and geographic location.

Salary range(s) for this role

$208,000 - $315,000 USD

Be Your Best at Snorkel

Joining Snorkel AI means becoming part of a company that has market proven solutions, robust funding, and is scaling rapidly—offering a unique combination of stability and the excitement of high growth. As a member of our team, you'll have meaningful opportunities to shape priorities and initiatives, influence key strategic decisions, and directly impact our ongoing success. Whether you're looking to deepen your technical expertise, explore leadership opportunities, or learn new skills across multiple functions, you're fully supported in building your career in an environment designed for growth, learning, and shared success.

Snorkel AI is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel AI embraces diversity and provides equal employment opportunities to all employees and applicants for employment. Snorkel AI prohibits discrimination and harassment of any type on the basis of race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local law. All employment is decided on the basis of qualifications, performance, merit, and business need.

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior | Staff Software Engineer - AI / ML New San Francisco, CA (Hybrid)
Senior | Staff Software Engineer - AI / ML New San Francisco, CA (Hybrid)

Snorkel • San Francisco (CA), Northern (KY)

Hybrid
USD 208,000 - 315,000
Senior Software Engineer - AI / ML
Senior Software Engineer - AI / ML

Snorkel AI • San Francisco (CA)

Hybrid
USD 190,000 - 240,000
Senior Manager, Forward Deployed Research
Senior Manager, Forward Deployed Research

Socket.dev • New York (NY)

On-site
USD 185,000 - 321,600
Senior / Staff AI Engineer
Senior / Staff AI Engineer

Snorkel AI • New York (NY), San Francisco (CA)

On-site
USD 150,000 - 190,000
Senior Software Engineer — AI Infrastructure
Senior Software Engineer — AI Infrastructure

Snorkel AI • San Francisco (CA)

On-site
USD 220,000 - 260,000
Director, Research - Evaluation & Training
Director, Research - Evaluation & Training

Snorkel AI • San Francisco (CA)

On-site
USD 275,000 - 425,000
Research Scientist - Frontier Benchmarks
Research Scientist - Frontier Benchmarks

Snorkel AI • New York (NY)

On-site
USD 200,000 - 350,000
Director, Research - Evaluation & Training
Director, Research - Evaluation & Training

Snorkel AI • United States

On-site
USD 275,000 - 425,000
Research Scientist - Frontier Benchmarks
Research Scientist - Frontier Benchmarks

Snorkel AI • San Francisco (CA)

On-site
USD 200,000 - 325,000
Staff Product Manager – AI/ML
Staff Product Manager – AI/ML

jobs.frontdoordefense.com - Jobboard • San Francisco (CA)

On-site
USD 240,000 - 300,000