ML Systems Engineer: Post-Training Pipelines & RL

ThirdLayer, Inc.

San Francisco (CA)

On-site

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ThirdLayer, Inc. is seeking a Research Engineer to build end-to-end training pipelines and infrastructure that support continuous post-training runs. You will bridge research and production, turning experiments into scalable, reliable systems that run with production data.

You will work with researchers to co-design methods and systems and partner with product engineers to ship impactful improvements across our RL and data platforms.

Qualifications

  • Strong software engineering fundamentals with ML depth.
  • Experience with distributed training or large-scale experiments.
  • Proficiency in Python and ML frameworks (PyTorch, JAX).
  • Ability to turn prototyping into production-ready systems.
  • Experience with RL environments or data quality systems is a plus.

Responsibilities

  • Build and own pipelines for post-training runs from data ingestion to deployment.
  • Turn research prototypes into reliable systems that run on production data.
  • Create infrastructure to generate, scale, and version training environments and evals.
  • Develop tooling to extract signals from traces and curate data for learning loops.
  • Build instrumentation to inspect, debug, and understand training runs.
  • Collaborate with researchers and product engineers to ship results.

Skills

Distributed training
ML systems
Python
PyTorch
JAX
Data pipelines
System design

Job description

ThirdLayer, Inc. is seeking a Research Engineer to build end-to-end training pipelines and infrastructure that support continuous post-training runs. You will bridge research and production, turning experiments into scalable, reliable systems that run with production data.

You will work with researchers to co-design methods and systems and partner with product engineers to ship impactful improvements across our RL and data platforms.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Infrastructure Engineer - Fast, Reliable RL Pipelines
ML Infrastructure Engineer - Fast, Reliable RL Pipelines

Thinking Machines Lab • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health benefits
Unlimited PTO
Parental leave
+1
Staff Engineer - RL Training Infrastructure
Staff Engineer - RL Training Infrastructure

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Research Engineer
Research Engineer

ThirdLayer, Inc. • San Francisco (CA)

On-site
USD 140,000 - 210,000
ML Systems Engineer: RL Training & AI Safety
ML Systems Engineer: RL Training & AI Safety

Anthropic • Seattle (WA)

Hybrid
USD 520,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
ML Systems Engineer — RL Training & Finetuning
ML Systems Engineer — RL Training & Finetuning

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
LLM Post-Training Engineer - RL, SFT & Data Pipelines
LLM Post-Training Engineer - RL, SFT & Data Pipelines

GoTo Meeting • Mountain View (CA)

On-site
USD 150,000 - 230,000
Health, dental, and vision care for you and your family
Top-tier 401(K) plan with company matching
Paid time off and paid holidays
+2
RL Infrastructure Engineer - Scale Distributed Training
RL Infrastructure Engineer - Scale Distributed Training

Elorian • Palo Alto (CA)

On-site
USD 200,000 - 400,000
Health, dental, vision benefits
Unlimited PTO
Parental leave
+1
Research Engineer: RL & Post-Training LLM Systems
Research Engineer: RL & Post-Training LLM Systems

Preference Model • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and equity compensation (>90th percentile)
Health, vision, dental benefits
401K match
+2
Member of Technical Staff, RL Infra
Member of Technical Staff, RL Infra

Inception • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior ML Infra Engineer: Scalable AI Training Systems
Senior ML Infra Engineer: Scalable AI Training Systems

Preference Model • Seattle (WA)

On-site
USD 180,000 - 300,000
Health insurance
Vision insurance
Dental insurance
+3