Research Engineer — LLM Fine-Tuning & Model Routing

Run BiOS

San Francisco, Northern (CA, KY)

Hybrid

USD 140,000 - 200,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Run BiOS runs serverless inference across the major open model families and fine-tunes custom models on dedicated GPUs. As a Research Engineer you improve the models themselves — from bios-adaptive routing across the model pool, to the fine-tuning methods and alignment recipes customers run.

You turn research into behavior that serves production traffic, and collaborate with product engineering to make research outcomes visible and controllable in the product.

Qualifications

  • At least 5 years in machine learning engineering or research, with meaningful time on large language models.
  • Hands-on experience fine-tuning open-weight models: PyTorch, distributed training, parameter-efficient methods.
  • Evaluation rigor: when you say a model got better, you can show how you know.
  • Engineering fundamentals strong enough that someone else can reproduce your experiments.

Responsibilities

  • Develop and evaluate routing strategies for bios-adaptive across the open-model pool, and measure what they do to quality, latency, and cost.
  • Design fine-tuning and alignment recipes — SFT, LoRA/QLoRA, DPO-family objectives, continued pre-training — that behave predictably on dedicated-GPU infrastructure.
  • Build evaluation harnesses that catch regressions before customers do.
  • Read the literature, reproduce what matters, discard what does not, and ship what wins.
  • Work with product engineering to make research outcomes visible and controllable in the product.
  • Write down what you learn — internal notes, docs, and occasionally public posts.

Skills

ML engineering
Large language models
PyTorch
Distributed training
Parameter-efficient methods
Reproducible research

Job description

Run BiOS runs serverless inference across the major open model families and fine-tunes custom models on dedicated GPUs. As a Research Engineer you improve the models themselves — from bios-adaptive routing across the model pool, to the fine-tuning methods and alignment recipes customers run.

You turn research into behavior that serves production traffic, and collaborate with product engineering to make research outcomes visible and controllable in the product.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer Turn routing, fine-tuning, and alignment research into production behavior. Remote · San Francisco, CA · Bangalore, India
Research Engineer Turn routing, fine-tuning, and alignment research into production behavior. Remote · San Francisco, CA · Bangalore, India

Run BiOS • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Staff Research Engineer, LLM Inference & Systems
Staff Research Engineer, LLM Inference & Systems

modal • New York (NY)

On-site
USD 180,000 - 240,000
Research Scientist - LLM Evaluation & Routing
Research Scientist - LLM Evaluation & Routing

OpenRouter • United States

On-site
USD 150,000 - 230,000
Staff Research Engineer - LLM Inference & Serving
Staff Research Engineer - LLM Inference & Serving

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Production-Grade LLM Inference Runtime Engineer
Production-Grade LLM Inference Runtime Engineer

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 360,000
Research Scientist: LLM Evaluation & Routing Innovator
Research Scientist: LLM Evaluation & Routing Innovator

OpenRouter • New York (NY)

On-site
USD 170,000 - 250,000
Remote LLM Engineer - Fine-Tuning & Pipelines
Remote LLM Engineer - Fine-Tuning & Pipelines

Bright Vision Technologies • Glastonbury (CT)

On-site
USD 100,000 - 150,000
Mid-Training Research Engineer — Scientific LLMs
Mid-Training Research Engineer — Scientific LLMs

Doist • Menlo Park (CA)

On-site
USD 250,000 - 350,000
Research Scientist: Post-Training
Research Scientist: Post-Training

Generalist • Somerville (MA), San Mateo (CA)

On-site
USD 100,000 - 130,000
Remote Research Engineer: Real-Time AI Inference
Remote Research Engineer: Real-Time AI Inference

ElevenLabs • Maine

Hybrid
USD 140,000 - 190,000
Annual discretionary stipend
Annual company offsite
Co-working stipend