Staff Researcher - Post-Training LLMs for Production

Mixpeek

San Francisco (CA)

On-site

USD 150,000 - 230,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Modal is building a platform that covers the full life cycle of LLMs—from training to deployment and observability. We already run multi-node training and elastic inference, and we seek a research‑leaning post‑training specialist to join our team in San Francisco or NYC.

You will own hands‑on post‑training research with the research lead, partner with customers and Forward Deployed Engineers, and explore questions like distillation from production traffic, while showing how the platform enables

Qualifications

  • A research‑leaning background in post‑training LLMs, with work you can point to.
  • Ability to tell which frontier techniques matter to users and which stay academic.
  • A track record of shipping research that others can build on.
  • The drive to take a research bet from idea to result with minimal hand‑holding.
  • Ability to work in person, in our NYC or San Francisco office.

Responsibilities

  • Own end-to-end post-training research bets: RL, distillation, long-context models, and routing.
  • Collaborate with customers alongside engineers to train models and translate findings into research.
  • Expand collaborations with external labs, e.g., ZLab on DFlash.
  • Work with engineering to turn research into product capabilities and frameworks.
  • Help shape the research agenda to guide Modal's future.

Job description

Modal is building a platform that covers the full life cycle of LLMs—from training to deployment and observability. We already run multi-node training and elastic inference, and we seek a research‑leaning post‑training specialist to join our team in San Francisco or NYC.

You will own hands‑on post‑training research with the research lead, partner with customers and Forward Deployed Engineers, and explore questions like distillation from production traffic, while showing how the platform enables

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Researcher - Post-Training LLMs & Production
Staff Researcher - Post-Training LLMs & Production

Mixpeek • New York (NY)

On-site
USD 180,000 - 230,000
Diversity of problem spaces
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Staff Research Engineer - LLM Inference & Serving
Staff Research Engineer - LLM Inference & Serving

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Staff Research Engineer, LLM Inference & Systems
Staff Research Engineer, LLM Inference & Systems

modal • New York (NY)

On-site
USD 180,000 - 240,000
Member of Technical Staff - Research, Post-Training
Member of Technical Staff - Research, Post-Training

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Staff LLM Inference Research Engineer
Staff LLM Inference Research Engineer

Mixpeek • New York (NY)

On-site
USD 180,000 - 260,000
Member of Technical Staff - Post-Training Research
Member of Technical Staff - Post-Training Research

Mixpeek • New York (NY)

On-site
USD 180,000 - 230,000
Diversity of problem spaces
Member of Technical Staff - Research, Post-Training
Member of Technical Staff - Research, Post-Training

Mixpeek • San Francisco (CA)

On-site
USD 150,000 - 230,000
Staff ML Engineer: LLM Post-Training & Deployment
Staff ML Engineer: LLM Post-Training & Deployment

Sanas • Palo Alto (CA)

On-site
USD 180,000 - 240,000
LLM Post-Training Research Scientist (SFT/RLHF)
LLM Post-Training Research Scientist (SFT/RLHF)

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2