Frontier RL Research Engineer: Scale & Systems

United States Digital Space LLC

San Francisco (CA)

Hybrid

USD 500,000 - 850,000

Full time

5 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

United States Digital Space LLC is seeking a senior RL researcher to scale reinforcement learning systems from small results to frontier-scale runs. You will design and implement next-generation architectures and RL algorithms, diagnose scale-related behavior, and own end-to-end performance from research code to hardware.

The role emphasizes building fast, reproducible experimentation infrastructure, evaluating throughput and cost implications, and collaborating across research and engineering

Qualifications

  • Experience with modern transformer language models, training dynamics, and large-scale optimization.
  • Hands-on experience training large models in a distributed setting with data, tensor, and pipeline parallelism.
  • Track record of original ML research or production impact (methods, architectures, optimizations).
  • Ability to design rigorous, scalable experiments with baselines, ablations, and robust statistics.

Responsibilities

  • Study how RL training and sampling scale with model size, context length, and compute; identify changes to keep scaling efficient.
  • Develop next-generation model architectures and RL algorithms for frontier-scale runs.
  • Take small-scale results to frontier-scale, diagnosing behavior and root causes (numerical, algorithmic, or systemic).
  • Build experimental infrastructure for fast, reproducible comparisons at meaningful scale.
  • Own end-to-end performance of the largest RL runs from research code to hardware.
  • Build models to estimate throughput and cost of architecture changes.

Skills

Transformer models
Distributed training
Python
JAX or PyTorch
C++ or Rust
Experiment design
Cost & scaling

Education

Bachelor's degree or equivalent

Job description

United States Digital Space LLC is seeking a senior RL researcher to scale reinforcement learning systems from small results to frontier-scale runs. You will design and implement next-generation architectures and RL algorithms, diagnose scale-related behavior, and own end-to-end performance from research code to hardware.

The role emphasizes building fast, reproducible experimentation infrastructure, evaluating throughput and cost implications, and collaborating across research and engineering

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

RL Frontiers Engineer: Scale-Driven Research & Systems
RL Frontiers Engineer: Scale-Driven Research & Systems

Alex Loftus • New York (NY)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Flexible hours
Vacation and parental leave
+1
RL Frontier Engineer: Scale-Savvy AI Architect
RL Frontier Engineer: Scale-Savvy AI Architect

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Generous vacation
Flexible working hours
RL Scaling Research Scientist for Frontier Models
RL Scaling Research Scientist for Frontier Models

Periodic Labs • San Francisco (CA)

On-site
USD 225,000 - 350,000
RL Systems Engineer - Scale, Reliability & Observability
RL Systems Engineer - Scale, Reliability & Observability

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
RL Scaling Research Scientist — Frontier Models
RL Scaling Research Scientist — Frontier Models

Speedrun Talent Network • Menlo Park (CA)

On-site
USD 225,000 - 350,000
Senior Research Engineer — Frontier AI & RL Systems
Senior Research Engineer — Frontier AI & RL Systems

Turing Enterprises, Inc. • San Francisco (CA)

On-site
USD 250,000 - 350,000
Scale-Out RL Systems Engineer
Scale-Out RL Systems Engineer

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching (optional)
Generous vacation and parental leave
+1
Remote RL Engineer - Scale & Deploy AI
Remote RL Engineer - Scale & Deploy AI

Bright Vision Technologies • Naperville (IL)

Remote
USD 96,000 - 120,000
Senior Research Engineer — Large-Agent Scaling
Senior Research Engineer — Large-Agent Scaling

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Remote Reinforcement Learning Engineer — Scale & Deploy
Remote Reinforcement Learning Engineer — Scale & Deploy

Bright Vision Technologies • Reston (VA)

On-site
USD 100,000 - 150,000