Member of Technical Staff - Post-Training

Reflection AI Ltd

New York (NY)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Top-tier compensation
Stock options
Health & wellness
Meals provided
Life & family
Unlimited vacation
Visa sponsorship
Team building

Job summary

Reflection AI Ltd. is hiring to build systems that transform powerful pre-trained models into aligned, general agents. You will drive research and engineering initiatives across data curation, large-scale optimization, and reinforcement learning to push the frontiers of AI.

The role demands deep ML fundamentals, strong software engineering, and a track record of improving model behavior. Collaboration across research and infra, with a bias for action and clear execution, is essential.

Qualifications

  • Deep understanding of machine learning fundamentals.
  • Strong engineering skills and ability to work in large ML codebases.
  • Experience improving model behavior via data, reward modeling, or RL.
  • Proven record of ambitious research or engineering agendas with measurable impact.
  • Ability to operate across research and infrastructure boundaries.
  • Excellent communication and collaboration.
  • Passion for advancing the frontier of intelligence.

Responsibilities

  • Build systems that transform pre-trained models into aligned, general agents.
  • Drive research and engineering initiatives from data curation to large-scale optimization.
  • Develop data generation pipelines, reward models, RL algorithms, and inference-time scaling techniques.
  • Collaborate across pre-training and post-training teams to deliver improvements in model capability.
  • Contribute to understanding how large models learn to reason and follow instructions.

Skills

ML fundamentals
Distributed systems
LLM training
Data generation
RL techniques
Execution clarity

Job description

Our Mission

Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on. We build open models that let anyone control their intelligence and help shape the future of AI. Our mission: make intelligence open and accessible to all.

About the Role
  • Build systems that transform powerful pre-trained models into aligned and general agents.

  • Drive research and engineering initiatives that push the frontier of post-training, from data curation to large-scale optimization.

  • Develop data generation pipelines, reward models, reinforcement learning algorithms, and inference-time scaling techniques.

  • Collaborate across pre-training and post-training teams to deliver step-function gains in model capability.

  • Contribute to shaping our understanding of how large models learn to reason, follow instructions, and improve through reinforcement learning.

About You
  • Deep understanding of machine learning fundamentals and practical experience with large-scale LLM training.

  • Strong engineering skills, comfortable diving into complex ML codebases and distributed systems.

  • Experience improving model behavior through data, reward modeling, or RL techniques.

  • Evidence of owning ambitious research or engineering agendas that led to measurable model improvements.

  • Thrive in a fast-paced, high-agency startup environment; bias toward action and clarity of execution.

  • Able to work fluidly across research and infra boundaries

  • Strong communication capabilities and comfort working collaboratively

  • Passionate about advancing the frontier of intelligence.

What We Offer:

We believe that to make intelligence open and accessible to all, you need to start at the ground up as part of a talent‑dense team. You will help define our future as a company, and help define the future of open foundational models.

We want you to do the most impactful work of your career with the confidence that you and the people you care about most are supported.

  • Top-tier compensation: Salary and equity structured to recognize and retain our talent globally.

  • Stock options: Everyone who joins and contributes to Reflection's success gets to share in the upside through stock options.

  • Health & wellness: Comprehensive medical, dental, vision, and life, with an annual wellness allowance.

  • Meals: Lunch and dinner are provided in the office daily.

  • Life & family: 22 weeks paid parental leave for all new birthing and non-birthing parents, including adoptive and surrogate journeys.

  • Vacation days: Unlimited paid time off in the U.S. and 30 days in the U.K.

  • Sponsorship support: We sponsor visas to help exceptional talent join our team and support long-term immigration pathways where applicable.

  • Team building: We have regular off-sites, happy hours, and team celebrations.

Export Control Notice: This position may require access to technology or source code subject to the U.S. Export Administration Regulations. Any offer of employment for this role may be conditioned on the Company's ability to provide the candidate with access to such technology or source code in compliance with applicable U.S. export control laws, which may require the Company to seek government authorization.

Candidate and Employee Privacy Notice: We collect and process personal data about applicants for the purposes described in our Candidate and Employee Privacy Notice

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Post-Training
Member of Technical Staff - Post-Training

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 300,000
Top-tier compensation
Stock options
Health & wellness
+5
Member of Technical Staff - Pre-Training
Member of Technical Staff - Pre-Training

Reflection AI Ltd • New York (NY)

On-site
USD 140,000 - 220,000
Top-tier pay
Stock options
Health insurance
+5
Forward Deployed Engineer, Lead - LLM Post-training
Forward Deployed Engineer, Lead - LLM Post-training

reflectionai • New York (NY), California (MO)

On-site
USD 180,000 - 280,000
Top-tier compensation
Stock options
Health & wellness
+2
Member of Technical Staff - Pre-Training Infra
Member of Technical Staff - Pre-Training Infra

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 280,000
Stock options
Health insurance
Meals provided
+4
Member of Technical Staff - Pre-Training
Member of Technical Staff - Pre-Training

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 180,000 - 280,000
Top-tier compensation
Stock options
Health & wellness
+5
Member of Technical Staff - Evaluations
Member of Technical Staff - Evaluations

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 150,000 - 230,000
Top-tier compensation
Stock options
Health & wellness
+5
Forward Deployed Engineer - LLM Post-training
Forward Deployed Engineer - LLM Post-training

reflectionai • San Francisco (CA), New York (NY)

On-site
USD 150,000 - 230,000
Top-tier compensation
Stock options
Health & wellness coverage
+2
Member of Technical Staff - Research Software Engineer
Member of Technical Staff - Research Software Engineer

Reflection AI Ltd • New York (NY)

On-site
USD 180,000 - 260,000
Top-tier compensation & equity
Stock options
Health & wellness benefits
+5
Member of Technical Staff - Pre-Training Infra
Member of Technical Staff - Pre-Training Infra

Reflection AI Ltd • New York (NY)

On-site
USD 180,000 - 270,000
Top-tier compensation
Stock options
Health & wellness
+5
Member of Technical Staff, Data Flywheel
Member of Technical Staff, Data Flywheel

Sierra Ventures • San Francisco (CA)

On-site
USD 180,000 - 280,000
Stock options
Health & wellness
Meals provided in office