LLM Post-Training & Agents Research Engineer

Kaon (prev. FlowGPT)

San Francisco (CA)

On-site

USD 200,000 - 500,000

Full time

45 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Kaon, based in the San Francisco Bay Area, is seeking a Research Engineer to advance the models behind our interactive storytelling and character experiences. You will focus on post-training methods to boost model behavior and deploy capable agents with memory, tools, and long-horizon interactions.

In this role, you will build training pipelines (SFT, DPO, GRPO), craft datasets, design reward models, and evaluate improvements through offline tests and live experiments, collaborating with

Qualifications

  • Experience training or post-training large language models, including RL or preference optimization.
  • Ability to turn open-ended research problems into clear hypotheses, experiments, and measurable results.
  • Strong Python and software engineering fundamentals; C/C++ is a plus.

Responsibilities

  • Build and refine LLM post-training pipelines (SFT, DPO, GRPO).
  • Develop training and preference datasets with held-out evaluations.
  • Design reward models for narrative quality, memory, and personalization.
  • Create agent training environments and evaluations for tool use and memory.
  • Collaborate with engineering and product teams to bring improvements into production.
  • Conduct offline evaluations and online A/B tests.

Skills

Python
C/C++
Post-training LLMs
RL/Pref optimization
LLM-based agents
Hypothesis testing
Ownership

Tools

PyTorch
TensorFlow
Linux/UNIX

Job description

Kaon, based in the San Francisco Bay Area, is seeking a Research Engineer to advance the models behind our interactive storytelling and character experiences. You will focus on post-training methods to boost model behavior and deploy capable agents with memory, tools, and long-horizon interactions.

In this role, you will build training pipelines (SFT, DPO, GRPO), craft datasets, design reward models, and evaluate improvements through offline tests and live experiments, collaborating with

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer - LLM Post-Training & Agents
Research Engineer - LLM Post-Training & Agents

Kaon (prev. FlowGPT) • San Francisco (CA)

On-site
USD 200,000 - 500,000
Agent Post-Training Researcher
Agent Post-Training Researcher

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
AI Research Scientist - Post-Training LLMs & Agent Systems
AI Research Scientist - Post-Training LLMs & Agent Systems

Perplexity • San Francisco (CA)

On-site
USD 220,000 - 485,000
Staff Researcher - Post-Training LLMs for Production
Staff Researcher - Post-Training LLMs for Production

Mixpeek • San Francisco (CA)

On-site
USD 150,000 - 230,000
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
LLM MaaS Engineer – Agents & Evaluation
LLM MaaS Engineer – Agents & Evaluation

ByteDance • San Jose (CA)

On-site
USD 128,000 - 256,000
Senior Research Scientist LLM
Senior Research Scientist LLM

techire ai • San Francisco (CA)

On-site
USD 350,000 - 500,000
Stock options
Remote work worldwide
Competitive compensation
LLM Post-Training Research Scientist (SFT & RLHF)
LLM Post-Training Research Scientist (SFT & RLHF)

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Agent Engineer — Production & Adoption
LLM Agent Engineer — Production & Adoption

Obsidian • San Francisco (CA)

On-site
USD 170,000 - 250,000
LLM Post-Training Engineer - RL, SFT & Data Pipelines
LLM Post-Training Engineer - RL, SFT & Data Pipelines

GoTo Meeting • Mountain View (CA)

On-site
USD 150,000 - 230,000
Health, dental, and vision care for you and your family
Top-tier 401(K) plan with company matching
Paid time off and paid holidays
+2