Principal Research Engineer, Model Training & Post-Training

Inflection AI, Inc.

Palo Alto, Northern (CA, KY)

Hybrid

USD 400,000 - 550,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Robust medical, dental and vision with
401k matching
Flexible time off
Holidays & leave
Equity participation
Visa support for Bay Area

Job summary

Inflection AI, Inc. in the San Francisco Bay Area is seeking a Principal Research Engineer to own the model-improvement loop from data and training through evals, post-training, release criteria and production feedback.

This hands-on technical leader will drive training strategies, architecture decisions, and large-scale distributed training across GPUs to ship models that are measurably better for users.

Qualifications

  • Experience leading large-scale LLM or foundation-model training or post-training programs.
  • Strong experience with transformer-based models and distributed training systems.
  • Experience operating large-scale training infrastructure, e.g., GPU clusters.

Responsibilities

  • Own the model-improvement roadmap across capability, reliability, and enterprise readiness.
  • Lead training and post-training strategy, including supervised fine-tuning and reward modeling.
  • Drive architecture and optimization decisions for training and inference.
  • Coordinate large-scale training on distributed GPU clusters (thousands of GPUs).
  • Define data strategy and evaluation data for production use and feedback loops.
  • Build and improve evaluation and release-quality systems and model-readiness reviews.
  • Partner with infrastructure and research teams to improve reliability and cost-performance.
  • Debug and improve model behavior across data, training, and production.

Skills

Transformer models
Distributed training
Post-training methods
RLHF
SFT
DPO
GRPO
RLAIF
Model evaluation
Budgeting cost and throughput

Education

PhD in Computer Science / ML / AI or related field

Tools

PyTorch
TensorFlow

Job description

Principal Research Engineer, Model Training & Post-Training
About Inflection AI

Inflection AI is a Public Benefit Corporation empowering people with human-centered, emotionally intelligent AI. We’re shaping the future of AI by combining emotional intelligence (EQ) and raw intelligence (IQ) to elevate people’s potential.
Inflection AI created Pi, the world’s first emotionally intelligent AI, to help people work through decisions, emotions, and challenges. Pi is a personal AI agent powered by Inflection AI’s foundation model, proving that AI can be personal, empathetic, and contextually aware.

About the Role

Inflection’s models are central to our product and platform strategy, and we are looking for a hands-on technical leader to own the model-improvement loop from data and training through evals, post-training, release criteria, and production feedback. This person will sit at the intersection of research, production engineering, and model release, with a mandate to ship models that are measurably better for users. The ideal candidate has led serious model training or post-training work before, can make principled tradeoffs across data, compute, architecture, and quality, around a clear technical roadmap.

What You’ll Do

  • Own the model-improvement roadmap across capability, reliability, emotional intelligence, tool-use, safety, latency, cost, and enterprise readiness.
  • Lead training and post-training strategy, including supervised fine-tuning, RLHF, DPO, GRPO, RLAIF, reward modeling, preference optimization, tool-use fine-tuning, distillation, synthetic data, and related methods.
  • Drive model architecture and optimization decisions across modern transformer-based and hybrid architectures, including both training-time and inference-time performance.
  • Lead large-scale training efforts on distributed GPU clusters, including systems operating at the scale of 1,000+ GPUs.
  • Define and execute data strategy across data curation, mixture design, deduplication, decontamination, human-in-the-loop pipelines, preference data, evaluation data, synthetic data, and production feedback loops.
  • Build and improve evaluation and release-quality systems, including model evals, quality gates, regression detection, release criteria, model-readiness reviews, and post-release monitoring.
  • Partner closely with infrastructure and research engineering teams to improve distributed training reliability, checkpointing, fault tolerance, observability, reproducibility, and cost-performance tradeoffs.
  • Debug and improve model behavior across the full stack: data, training, post-training, evaluation, infrastructure, product integration, and production feedback.

What We’re Looking For

  • Experience leading, or serving as a principal contributor to, large-scale LLM, multimodal, or foundation-model training or post-training programs.
  • Deep experience with transformer-based models, hybrid architectures, modern deep-learning frameworks, and distributed training systems.
  • Strong practical experience with post-training and alignment methods such as SFT, RLHF, DPO, GRPO, RLAIF, reward modeling, preference optimization, tool-use fine-tuning, or related approaches.
  • Experience operating or partnering on large-scale training infrastructure, ideally including GPU clusters at the scale of 1,000+ GPUs.
  • Strong systems instincts around throughput, cost, reliability, observability, debugging, checkpointing, reproducibility, and fault tolerance.
  • Excellent judgment around data quality, evaluation design, model regressions, release readiness, and production model behavior.
  • Ability to balance research ambition with product pragmatism, user impact, and operational discipline.
  • Experience leading senior technical teams while continuing to contribute directly to technical decisions and implementation.
  • PhD in Computer Science, Machine Learning, Artificial Intelligence, or a related field, or equivalent practical experience.
Employee Pay Disclosures

At Inflection AI, we aim to attract and retain the best employees and compensate them in a way that appropriately and fairly values their individual contributions to the company. For this role, Inflection AI estimates a starting annual base salary to fall within the range of $400,000 to $550,000,depending on a candidate’s qualifications and level of experience. This role also includes a meaningful equity component, allowing employees to share in the long-term success of the company.

Inflection AI values and supports our team’s mental, emotional, financial and physical health. We are focused on building a positive, safe, inclusive and inspiring place to work. Our benefits include:

  • Robust medical, dental and vision options with employer contributions for HSA, FSA and DFSA
  • 401k matching program
  • Flexible Time Off, 10 paid holidays, 5 days sick leave
  • Parental, Medical and Family care leave
  • Generous cell-phone, wellness and office set up stipends
  • Support of country-specific visa needs for international employees living in the Bay Area
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Research Engineer, Model Training & Post-Training
Principal Research Engineer, Model Training & Post-Training

Inflection AI • Palo Alto (CA)

On-site
USD 400,000 - 550,000
Diverse medical, dental and vision options
401k matching program
Unlimited paid time off
+2
Member of Technical Staff – Model Training
Member of Technical Staff – Model Training

Inflection AI • Palo Alto (CA)

Hybrid
USD 175,000 - 350,000
Competitive stock options
Diverse medical, dental, and vision options
401k matching program
+3
Principal Engineer, Agentic AI Systems
Principal Engineer, Agentic AI Systems

Inflection AI • Palo Alto (CA)

On-site
USD 400,000 - 550,000
Diverse medical, dental and vision options
401k matching program
Unlimited paid time off
+1
Principal Research & Engineering, Realtime Voice AI
Principal Research & Engineering, Realtime Voice AI

Inflection AI, Inc. • Palo Alto (CA), Northern (KY)

Hybrid
USD 400,000 - 550,000
Robust medical, dental and vision with
401k matching
Flexible Time Off
+3
Principal Research & Engineering, Realtime Voice AI
Principal Research & Engineering, Realtime Voice AI

Inflection AI • Palo Alto (CA)

On-site
USD 400,000 - 550,000
Diverse medical, dental, and vision options
401k matching program
Unlimited paid time off
+1
Senior Data Analyst (Growth)
Senior Data Analyst (Growth)

Inflection AI • Palo Alto (CA)

On-site
USD 235,000 - 300,000
Diverse medical, dental and vision options
401k matching program
Unlimited paid time off
+1
Principal AI Model Training & Post-Training Lead
Principal AI Model Training & Post-Training Lead

Inflection AI • Palo Alto (CA)

On-site
USD 400,000 - 550,000
Diverse medical, dental and vision options
401k matching program
Unlimited paid time off
+2
Senior AI Forward Deployed Engineer
Senior AI Forward Deployed Engineer

Handshake • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Equity in a fast-growing company
401(k) match
Paid parental leave
+2
Member of Technical Staff – Backend (Agentic)
Member of Technical Staff – Backend (Agentic)

Inflection AI • Palo Alto (CA)

On-site
USD 175,000 - 350,000
Diverse medical, dental and vision options
401k matching program
Unlimited paid time off
+2
Senior AI Forward Deployed Engineer
Senior AI Forward Deployed Engineer

Apply • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Equity in a fast-growing company
401(k) match
Paid parental leave
+2