Member of Technical Staff - Research, Post-Training

Modal Labs

New York (NY)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Modal Labs is building an AI infrastructure platform and seeks a researcher focused on post-training LLMs. You will work with the research lead to pick high-impact bets and own them from idea to production, alongside our engineering and product teams.

You will partner with customers and Forward Deployed Engineers to train models, distill insights, and extend collaborations with external labs. This role requires in-person work in NYC or San Francisco and a strong track record of shipping research.

Qualifications

  • A research-leaning background in post-training LLMs.

Responsibilities

  • Own end-to-end post-training research bets: async and agentic RL, on-policy distillation, long-context RL.
  • Collaborate with customers and Forward Deployed Engineers to train models.
  • Carry collaborations with outside research labs (e.g., ZLab on DFlash).
  • Turn frontier post-training techniques into products with engineering teams.
  • Help shape the research agenda to guide future work.

Skills

Post-training LLMs
Product sense
Open research

Job description

About Us:

AI needs a new infrastructure layer. We're building it at Modal.

Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now.

Our customers include category-defining companies like Lovable, Ramp, Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub‑second container starts, and native storage, so it's simple to serve low‑latency inference, fine‑tune models, and access production‑ready sandboxes at scale.

We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September.

Our team includes creators of popular open‑source projects (e.g.,Seaborn,Luigi), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience.

The Role:

We're building a platform that covers the whole life of an LLM: training it, deploying it, and observing it in production. We already run multi‑node training, elastic inference, sandboxes, and distributed volumes, and we control the infrastructure underneath. We’re looking for research depth in post‑training to sit alongside our systems and product work.

You will do hands‑on post‑training research at Modal, working with the research lead to pick high‑impact bets and owning them end to end. The work that pays off fastest is tied to production workloads -- we're already experts at training speculators for deployed models, and there are open research questions like distilling a target model from its own production traffic. There is also room to prove what the platform makes possible, where training AI scientists or kernel engineers is a natural fit given our GPU sandboxes.

What you'll do:
  • Own end‑to‑end post‑training research bets: async and agentic RL, on‑policy distillation, long‑context RL, small routing models, and whatever else the research agenda calls for.

  • Work directly with customers alongside our Forward Deployed Engineers to train models and bring what you learn back into the research.

  • Carry and expand collaborations with outside research labs. For example, our work with ZLab on DFlash, a speculator design built on KV injection and blockwise parallel drafting.

  • Work with engineering to turn frontier post‑training techniques into products: an opinionated post‑training framework, distributed‑training approaches (DiLoCo, evolutionary strategies), online training for deployed models, and more.

  • Help shape the research agenda. None of the above is prescriptive; your work will help guide our future.

Requirements:
  • A research‑leaning background in post‑training LLMs, with work you can point to.

  • Enough product sense to tell which frontier techniques matter to users and which stay academic.

  • A record of shipping research that other people build on, whether in a lab or in industry.

  • The drive to take a research bet from idea to result without much hand‑holding, working in the open with the rest of the team.

  • Ability to work in‑person, in our NYC or San Francisco office.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Research, Post-Training
Member of Technical Staff - Research, Post-Training

Mixpeek • San Francisco (CA)

On-site
USD 150,000 - 230,000
Member of Technical Staff - Post-Training Research
Member of Technical Staff - Post-Training Research

Mixpeek • New York (NY)

On-site
USD 180,000 - 230,000
Diversity of problem spaces
Member of Technical Staff - Research, Inference
Member of Technical Staff - Research, Inference

modal • New York (NY)

On-site
USD 180,000 - 240,000
Member of Technical Staff - Research, Inference
Member of Technical Staff - Research, Inference

Mixpeek • San Francisco (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff - Inference Research
Member of Technical Staff - Inference Research

Mixpeek • New York (NY)

On-site
USD 180,000 - 260,000
ML Research Intern
ML Research Intern

Neura Market • New York (NY)

On-site
ML Research Intern
ML Research Intern

Mixpeek • San Francisco (CA)

On-site
USD 60,000 - 90,000
Mentorship by senior researchers
Research publication opportunities
PhD internship stipend
ML Research Intern
ML Research Intern

Triwill Group • New York (NY)

On-site
ML Research Intern
ML Research Intern

Modal Labs • New York (NY)

On-site
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000