Senior LLM Researcher - Post-Training & Alignment (Remote)

techire ai

San Francisco (CA)

On-site

USD 350,000 - 500,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Stock options
Remote work worldwide
Competitive compensation

Job summary

techire ai is a Series A company building the next generation of conversational AI, with an ex-NVIDIA & Meta research leader and a co-creator behind open-source models. The team powers hundreds of millions of conversations monthly, offering work on research that scales to real-world deployments.

They're seeking researchers who have owned meaningful post-training improvements and can tackle data, modelling, evaluation and RL infrastructure challenges to raise model quality and alignment at scale.

Qualifications

  • Experience leading post-training projects delivering tangible capability improvements.
  • Understanding practical challenges behind post-training techniques, not just theory.
  • Experience across data, modelling, evaluation and RL infrastructure, improving model quality.

Responsibilities

  • Design and run post-training experiments using techniques like DPO, GRPO, SFT and rejection sampling.
  • Build and scale reinforcement learning and post-training infrastructure.
  • Develop reward models, evaluation frameworks and assessment rubrics to improve model quality and behavior.
  • Collaborate with vendors and data partners to curate high-quality post-training datasets and feedback pipelines.
  • Own end-to-end projects spanning data collection, modelling, evaluation and infrastructure, driving improvements in reasoning and alignment.

Skills

Post-training projects
RL infrastructure
Evaluation frameworks
Data partnerships
End-to-end project ownership

Job description

techire ai is a Series A company building the next generation of conversational AI, with an ex-NVIDIA & Meta research leader and a co-creator behind open-source models. The team powers hundreds of millions of conversations monthly, offering work on research that scales to real-world deployments.

They're seeking researchers who have owned meaningful post-training improvements and can tackle data, modelling, evaluation and RL infrastructure challenges to raise model quality and alignment at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Research Scientist LLM
Senior Research Scientist LLM

techire ai • San Francisco (CA)

On-site
USD 350,000 - 500,000
Stock options
Remote work worldwide
Competitive compensation
Senior LLM Training & Post-Training Engineer (Hybrid/Remote)
Senior LLM Training & Post-Training Engineer (Hybrid/Remote)

Lightning AI • San Francisco (CA), Seattle (WA), New York (NY)

Hybrid
USD 165,000 - 310,000
Comprehensive Health Coverage
Meaningful Equity
401(k) matching & pension
+2
LLM Post-Training Research Scientist (SFT & RLHF)
LLM Post-Training Research Scientist (SFT & RLHF)

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Alignment & Post-Training Research Scientist
LLM Alignment & Post-Training Research Scientist

Scale AI • New York (NY)

On-site
USD 150,000 - 200,000
Health & wellbeing benefits
Learning & development stipend
Community & ERG events
+1
Senior AI Engineer (LLM Training & RLHF) - Remote
Senior AI Engineer (LLM Training & RLHF) - Remote

Prolific • Virginia Beach (VA)

On-site
USD 100,000 - 140,000
Competitive pay rates
Flexible hours
Ability to work from home
Remote LLM Engineer - Research & Deployment
Remote LLM Engineer - Research & Deployment

Fastino Labs • San Francisco (CA)

Hybrid
GBP 65,000 - 90,000
Remote AI/ML Research Engineer — LLM Training & Evaluation
Remote AI/ML Research Engineer — LLM Training & Evaluation

Rex.zone • United States

Remote
USD 80,000 - 100,000
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Senior AI Engineer (LLM Training & RLHF) - Remote
Senior AI Engineer (LLM Training & RLHF) - Remote

Prolific • Mesa (AZ)

On-site
USD 140,000 - 230,000
LLM Post-Training Scientist - SFT & RLHF
LLM Post-Training Scientist - SFT & RLHF

Scale AI, Inc. • San Francisco (CA)

On-site
USD 181,000 - 226,000
Health benefits
Retirement plan
Learning stipend
+2