Member of Technical Staff - Post-Training

reflectionai

San Francisco, New York (CA, NY)

On-site

USD 180,000 - 300,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Top-tier compensation
Stock options
Health & wellness
Meals provided in office
Parental leave
Unlimited vacation (US)
Visa sponsorship where applicable
Team building

Job summary

Reflection AI in San Francisco is building open foundational models and scalable systems to advance intelligent agents. We are seeking a seasoned AI/ML engineer to transform pre-trained models into aligned, general agents and to push the frontier of post-training with data curation and large-scale optimization.

You will collaborate across research and infrastructure, own ambitious agendas, and help shape the future of open intelligence through concrete experiments and robust implementation.

Qualifications

  • Deep understanding of machine learning fundamentals and practical experience with large-scale LLM training.
  • Strong engineering skills with complex ML codebases and distributed systems.
  • Experience improving model behavior through data, reward modeling, or RL techniques.
  • Evidence of owning ambitious research or engineering agendas leading to measurable model improvements.
  • Thrive in a fast-paced startup environment with action and clarity of execution.

Responsibilities

  • Build systems that transform powerful pre-trained models into aligned and general agents.
  • Drive research and engineering initiatives from data curation to large-scale optimization.
  • Develop data generation pipelines, reward models, reinforcement learning algorithms, and inference-time scaling techniques.
  • Collaborate across pre-training and post-training teams to deliver step-function gains in model capability.
  • Contribute to understanding how large models learn to reason, follow instructions, and improve through reinforcement learning.

Skills

Machine learning fundamentals
LLM training
Distributed systems
Data generation / reward modeling
Research leadership

Job description

Our Mission

Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on. We build open models that let anyone control their intelligence and help shape the future of AI. Our mission: make intelligence open and accessible to all.

About the Role
  • Build systems that transform powerful pre-trained models into aligned and general agents.

  • Drive research and engineering initiatives that push the frontier of post-training, from data curation to large-scale optimization.

  • Develop data generation pipelines, reward models, reinforcement learning algorithms, and inference-time scaling techniques.

  • Collaborate across pre-training and post-training teams to deliver step-function gains in model capability.

  • Contribute to shaping our understanding of how large models learn to reason, follow instructions, and improve through reinforcement learning.

About You
  • Deep understanding of machine learning fundamentals and practical experience with large-scale LLM training.

  • Strong engineering skills, comfortable diving into complex ML codebases and distributed systems.

  • Experience improving model behavior through data, reward modeling, or RL techniques.

  • Evidence of owning ambitious research or engineering agendas that led to measurable model improvements.

  • Thrive in a fast-paced, high-agency startup environment; bias toward action and clarity of execution.

  • Able to work fluidly across research and infra boundaries

  • Strong communication capabilities and comfort working collaboratively

  • Passionate about advancing the frontier of intelligence.

What We Offer:

We believe that to make intelligence open and accessible to all, you need to start at the foundation. Joining Reflection means building from the ground up as part of a talent-dense team. You will help define our future as a company, and help define the future of open foundational models.

We want you to do the most impactful work of your career with the confidence that you and the people you care about most are supported.

  • Top-tier compensation: Salary and equity structured to recognize and retain our talent globally.

  • Stock options: Everyone who joins and contributes to Reflection's success gets to share in the upside through stock options.

  • Health & wellness: Comprehensive medical, dental, vision, and life, with an annual wellness allowance.

  • Meals: Lunch and dinner are provided in the office daily.

  • Life & family: 22 weeks paid parental leave for all new birthing and non-birthing parents, including adoptive and surrogate journeys.

  • Vacation days: Unlimited paid time off in the U.S. and 30 days in the U.K.

  • Sponsorship support: We sponsor visas to help exceptional talent join our team and support long-term immigration pathways where applicable.

  • Team building: We have regular off-sites, happy hours, and team celebrations.

Export Control Notice: This position may require access to technology or source code subject to the U.S. Export Administration Regulations. Any offer of employment for this role may be conditioned on the Company's ability to provide the candidate with access to such technology or source code in compliance with applicable U.S. export control laws, which may require the Company to seek government authorization.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Pre-Training
Member of Technical Staff - Pre-Training

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 280,000
Top-tier compensation
Stock options
Health & wellness
+5
Member of Technical Staff - Research Software Engineer
Member of Technical Staff - Research Software Engineer

reflectionai • New York (NY)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness benefits
+4
Member of Technical Staff - Pre-Training Infra
Member of Technical Staff - Pre-Training Infra

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 280,000
Stock options
Health insurance
Meals provided
+4
Forward Deployed Engineer, Lead - LLM Post-training
Forward Deployed Engineer, Lead - LLM Post-training

reflectionai • New York (NY), California (MO)

On-site
USD 180,000 - 280,000
Top-tier compensation
Stock options
Health & wellness
+2
Forward Deployed Engineer, Lead - LLM Post-training
Forward Deployed Engineer, Lead - LLM Post-training

Reflection AI • New York (NY)

On-site
USD 180,000 - 240,000
Stock options
Health insurance
Meals provided
+4
Member of Technical Staff - Evaluations
Member of Technical Staff - Evaluations

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 150,000 - 230,000
Top-tier compensation
Stock options
Health & wellness
+5
Member of Technical Staff - Data Quality Engineer (Post-training)
Member of Technical Staff - Data Quality Engineer (Post-training)

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 140,000 - 190,000
Top-tier compensation
Stock options
Health & wellness
+5
Research Program Manager - Model Development
Research Program Manager - Model Development

reflectionai • New York (NY)

On-site
USD 140,000 - 170,000
Top-tier compensation
Stock options
Health & wellness coverage
+1
Forward Deployed Engineer - LLM Post-training
Forward Deployed Engineer - LLM Post-training

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 150,000 - 230,000
Top-tier compensation
Stock options
Health & wellness coverage
+2
Member of Technical Staff - Engineering Lead, Data Ingestion
Member of Technical Staff - Engineering Lead, Data Ingestion

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 240,000
Stock options
Healthcare
Meals provided
+4