Staff Research Software Engineer: Open ML Infra & RL Training

Reflection AI Ltd

New York (NY)

On-site

USD 180,000 - 260,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Top-tier compensation & equity
Stock options
Health & wellness benefits
Meals provided
Generous parental leave
Unlimited vacation (US) / 30 days (UK)
Visa sponsorship
Team off-sites

Job summary

Reflection AI Ltd. is seeking an engineer to architect and optimize the core training infrastructure powering our frontier AI models.

You will work with researchers to turn new ideas into reliable, scalable training systems capable of handling RL training loops, distributed GPUs, and vast data pipelines. You will collaborate across teams to implement techniques, ensure numerical stability, and reduce bottlenecks, delivering production-ready training stacks that support rapid iteration at massive

Qualifications

  • Strong software engineer with ML literacy.
  • Ability to translate research into production-grade infrastructure.
  • Experience in distributed training or data infrastructure is highly relevant.

Responsibilities

  • Design and optimize large-scale training loops and data pipelines.
  • Implement state-of-the-art techniques ensuring numerical stability and efficiency.
  • Build tooling to launch, monitor, and reproduce complex experiments.
  • Diagnose bottlenecks across the training stack (GPU memory, comms, dataloader).
  • Translate research prototypes into production-grade infrastructure.

Skills

Strong software engineering
ML concepts & research-to-prod
Distributed systems mindset

Tools

PyTorch
JAX
Megatron-style stacks
Triton
Ray
Kubernetes
Slurm
NCCL
RDMA

Job description

Reflection AI Ltd. is seeking an engineer to architect and optimize the core training infrastructure powering our frontier AI models.

You will work with researchers to turn new ideas into reliable, scalable training systems capable of handling RL training loops, distributed GPUs, and vast data pipelines. You will collaborate across teams to implement techniques, ensure numerical stability, and reduce bottlenecks, delivering production-ready training stacks that support rapid iteration at massive

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff ML Engineer — Open Models & RL Research
Staff ML Engineer — Open Models & RL Research

Reflection AI Ltd • New York (NY)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness
+5
Staff ML Engineer - Open-Models & RL Systems
Staff ML Engineer - Open-Models & RL Systems

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 300,000
Top-tier compensation
Stock options
Health & wellness
+5
Senior ML Infrastructure Engineer — Frontier RL & LLM Training
Senior ML Infrastructure Engineer — Frontier RL & LLM Training

Preference Model • San Francisco (CA)

On-site
USD 200,000 - 350,000
Health, vision, dental benefits
401K match
Lunch provided onsite
+2
Staff Engineer - Large-Scale GPU Inference & RL Infra
Staff Engineer - Large-Scale GPU Inference & RL Infra

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 320,000
Top-tier compensation
Stock options
Health & wellness
+5
Research Infra Engineer for AI & RL Systems
Research Infra Engineer for AI & RL Systems

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 300,000 - 475,000
Health, dental, vision benefits
Unlimited PTO
Paid parental leave
+1
Staff ML Researcher — Open-Foundations & Scaling
Staff ML Researcher — Open-Foundations & Scaling

Reflection AI Ltd • New York (NY)

On-site
USD 140,000 - 220,000
Top-tier pay
Stock options
Health insurance
+5
Strategic Research Programs Lead in AI Infrastructure
Strategic Research Programs Lead in AI Infrastructure

Reflection AI Ltd • New York (NY)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness
+3
Research Engineer, RL & LLM Post-Training — Scale ML Infra
Research Engineer, RL & LLM Post-Training — Scale ML Infra

Preference Model • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive cash and equity (>90th pct
Ownership and autonomy
Lunch onsite
+4
Research Software Engineer — Scalable RL & Distributed Training
Research Software Engineer — Scalable RL & Distributed Training

Reflection AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health insurance
Fully paid parental leave
+2
Senior RL Post-Training Systems Engineer
Senior RL Post-Training Systems Engineer

NVIDIA • Massachusetts

On-site
USD 224,000 - 357,000