Research Engineer - Midtraining

Periodic Labs

Menlo Park (CA)

On-site

USD 250,000 - 350,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Doist is seeking a Midtraining Research Engineer to advance scientific reasoning in frontier models. You will curate data, generate synthetic datasets, and build robust evals while enabling scalable, GPU-driven training experiments.

The role supports pre-training groundwork and requires hands-on distillation techniques and scalable tooling. You will collaborate with RL researchers, scientists, and engineers to push data-driven breakthroughs and improve model intelligence in science-focused

Qualifications

  • Bachelor's degree or equivalent experience required.
  • Experience training LLMs at scale is essential.
  • Ability to design and evaluate scientific reasoning data.

Responsibilities

  • Identify, process, and curate novel sources of scientific data for large-scale model training.
  • Generate high-quality synthetic data to fill gaps in scientific knowledge.
  • Build evaluations that correlate with downstream scientific task performance.
  • Develop and apply techniques such as self-distillation and on-policy distillation.
  • Design and run large-scale training experiments, partnering with supercompute engineers to scale across thousands of GPUs.
  • Build tools to study how data choices shape model intelligence.

Skills

LLM training
Data curation
Large-scale training
Eval design
Distributed training

Education

Bachelor's degree or equivalent

Tools

PyTorch
TensorFlow

Job description

We're an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across materials, energy, and beyond. Backed by world-class investors and growing rapidly, we operate at the pace the frontier requires. Our team brings deep expertise, genuine ownership, and a drive to push the boundaries of what's scientifically possible.

About the Role

We're training frontier models to develop deep scientific knowledge and reasoning for scientific discovery. As a Midtraining Research Engineer, you'll take base models and improve their scientific reasoning: curating and generating data, building evals, and running large-scale training experiments. Your work will also lay the groundwork for our pre‑training efforts down the line.

What You'll Do
  • Identify, process, and curate novel sources of scientific data for large-scale model training.
  • Generate high-quality synthetic data to fill gaps in scientific knowledge and reasoning.
  • Build evaluations that correlate with downstream scientific task performance, working closely with RL researchers, physicists, and chemists.
  • Develop and apply techniques such as self-distillation and on-policy distillation to improve model capability.
  • Design and run large-scale training experiments, partnering with supercompute engineers to scale efficiently across thousands of GPUs.
  • Build tools for yourself and the team to investigate how data choices shape model intelligence.
You Will Thrive in This Role If You Have
  • Experience training LLMs on curated mixes of trillions of tokens.
  • Experience with mid-training or pre-training at scale — big‑lab experience is a strong plus.
  • Experience on a dedicated evals team supporting a large production training run.
  • Hands‑on use of self‑distillation, on‑policy distillation, or similar methods in a real training pipeline.
  • The ability to calculate scaling laws and compute‑optimal hyperparameters.
  • Comfort working across data, evals, and training infrastructure.
  • Especially Strong Candidates May Also HaveExperience optimizing throughput and reliability for large-scale distributed training runs.
  • A background in AI for science or training on specialized domain data (e.g., protein, materials, or other scientific datasets).
  • Experience on a big training run tracking evals and driving interventions while the run was live, not just as a peripheral contributor.
Mechanics

Minimum education: Bachelor's degree or similar experience

  • Location: Menlo Park, CA (Soon: San Francisco, too)
  • Compensation: $250,000–$350,000 + equity
  • Visa sponsorship: Yes, we sponsor visas and will do everything we can to assist in this process.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Post-Training, RL
Member of Technical Staff, Post-Training, RL

Mirendil • United States

On-site
USD 350,000 - 500,000
Equity grant
Competitive benefits
Research Scientist, Data
Research Scientist, Data

Periodic Labs • Menlo Park (CA)

On-site
USD 250,000 - 350,000
Visa sponsorship
Research Scientist, Post-Training
Research Scientist, Post-Training

David Joseph & Company • San Francisco (CA)

On-site
USD 150,000 - 450,000
Research Engineer, Post-Training
Research Engineer, Post-Training

cognition • San Francisco (CA)

On-site
USD 150,000 - 210,000
Research Engineer, Applied AI
Research Engineer, Applied AI

HeyMilo AI • San Francisco (CA)

On-site
USD 150,000 - 210,000
Research Scientist / Research Engineer
Research Scientist / Research Engineer

Clera • San Francisco (CA)

On-site
Lead Research Engineer, Data Quality
Lead Research Engineer, Data Quality

Clera • San Francisco (CA)

On-site
USD 150,000 - 250,000
Equity participation
Visa sponsorship
Relocation assistance
Post-Training Research Scientist: Data-Driven AI Experiments
Post-Training Research Scientist: Data-Driven AI Experiments

David Joseph & Company • San Francisco (CA)

On-site
USD 150,000 - 450,000
Member of Technical Staff - Research & Post-training
Member of Technical Staff - Research & Post-training

Preference Model • Seattle (WA)

On-site
USD 200,000 - 350,000
Competitive cash and equity compensation (>90th percentile)
Ownership and autonomy
Health, vision, dental benefits
+4
Research Engineer Infrastructure Training Systems
Research Engineer Infrastructure Training Systems

Thinking Machines Lab • San Francisco (CA)

On-site
USD 350,000 - 475,000
Generous health benefits
Unlimited PTO
Paid parental leave
+1