Lead Large-Scale RL & Mid-Training Infra Engineer

Engg

Palo Alto (CA)

On-site

USD 180,000 - 240,000

Full time

10 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Peano AI is seeking a Founding Large-Scale Mid-Training/RL Infrastructure Engineer to design, run, and scale our in-house training stack. You will work with thousands of GPUs and TPU accelerators to optimize pretraining and RL-based methods for large language models, impacting core products and future roadmap.

This role is deeply technical, with responsibilities spanning data pipelines, sharding, checkpointing, and accelerator-aware scheduling.

Qualifications

  • Experience with large-scale deep learning model training and distributed systems.
  • Proven track record with pretraining and RL-based fine-tuning of large models.
  • Hands-on with Megatron, Transformer-Engine, verl, slime.
  • Experience in data pipelines, model/data/optimizer sharding, and checkpointing.
  • Strong accelerator/memory optimization skills at cluster scale.
  • Excellent debugging and performance-tuning in distributed environments.
  • Comfort working in deeply technical, early-stage startup settings.

Responsibilities

  • Build, optimize, and scale distributed training infrastructure for foundation models.
  • Own throughput, efficiency, scalability, cost, and reliability of pretraining and RL pipelines.
  • Architect data loading, sharding, checkpointing, batching for multi-node training.
  • Integrate and optimize RL components: rollout, reward modeling, environment orchestration.
  • Work with Megatron, Transformer-Engine, verl, slime and other libraries.
  • Tune memory, mixed-precision, and runtime performance for massive models.
  • Debug and profile training bottlenecks across code, compute, networking, and infra.
  • Collaborate with research, infra, and applications teams to deliver performant models.

Skills

Large-scale model training
Distributed systems
RL training
Performance tuning
Debugging
Startup mindset

Tools

Megatron
Transformer-Engine
verl
slime

Job description

Peano AI is seeking a Founding Large-Scale Mid-Training/RL Infrastructure Engineer to design, run, and scale our in-house training stack. You will work with thousands of GPUs and TPU accelerators to optimize pretraining and RL-based methods for large language models, impacting core products and future roadmap.

This role is deeply technical, with responsibilities spanning data pipelines, sharding, checkpointing, and accelerator-aware scheduling.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Infra Engineer for Large-Scale Mid-Training & RL
Senior ML Infra Engineer for Large-Scale Mid-Training & RL

Peano AI • Palo Alto (CA)

On-site
USD 200,000 - 260,000
Founding Mid-Training/RL Infrastructure Engineer
Founding Mid-Training/RL Infrastructure Engineer

Engg • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Founding Mid-Training/RL Infrastructure Engineer
Founding Mid-Training/RL Infrastructure Engineer

Peano AI • Palo Alto (CA)

On-site
USD 200,000 - 260,000
Senior RL Post-Training Systems Engineer (Equity)
Senior RL Post-Training Systems Engineer (Equity)

NVIDIA • Washington

On-site
USD 184,000 - 357,000
Equity
Comprehensive benefits
Member of Technical Staff, RL Infra
Member of Technical Staff, RL Infra

Inception • San Francisco (CA)

On-site
USD 180,000 - 240,000
RL Post-Training Systems Architect (Equity Eligible)
RL Post-Training Systems Architect (Equity Eligible)

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
RL Infrastructure Engineer: Scale End-to-End ML
RL Infrastructure Engineer: Scale End-to-End ML

Lever, Inc. • Sunnyvale (CA)

On-site
USD 150,000 - 450,000
Medical benefits
Dental benefits
Vision benefits
+7
Staff Engineer, Scalable RL Infrastructure
Staff Engineer, Scalable RL Infrastructure

Inception • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior RL Post-Training Frameworks Architect
Senior RL Post-Training Frameworks Architect

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
RL Post-Training Infrastructure Architect
RL Post-Training Infrastructure Architect

NVIDIA • Washington

On-site
USD 184,000 - 357,000