Research Scientist - LLM Evaluation & Routing

OpenRouter

United States

On-site

USD 150,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenRouter seeks a Research Scientist to conduct original research advancing how the world understands, evaluates, and routes large language models. You will analyze billions of LLM generations across providers and use cases.

You will own a research agenda, design experiments and evaluation frameworks, and produce work that influences model comparison, selection, and deployment. Collaboration with product and engineering is key, emphasizing depth and rigor.

Qualifications

  • Advanced degree in ML, CS, statistics, or mathematics.
  • Track record of original research with publications or notable impact.
  • Strong statistical, experimental design, and causal inference skills.
  • Experience with Python and large-scale experiments.

Responsibilities

  • Own and pursue a research agenda focused on LLM evaluation, model quality, routing optimization, and AI usage patterns.
  • Design novel evaluation frameworks and benchmarks using real-world generation data.
  • Conduct large-scale empirical studies on LLM behavior across providers and time.
  • Develop the statistical and mathematical foundations behind routing systems and model selection.
  • Identify opportunities to apply research findings to OpenRouter's product and platform.
  • Collaborate with external researchers and model providers to advance LLM capabilities.
  • Translate research insights into platform improvements with product/engineering teams.

Skills

Python
SQL
Statistics
Experiment design
Research publications
LLM techniques

Education

MS or PhD in a quantitative field

Tools

ClickHouse
BigQuery

Job description

OpenRouter seeks a Research Scientist to conduct original research advancing how the world understands, evaluates, and routes large language models. You will analyze billions of LLM generations across providers and use cases.

You will own a research agenda, design experiments and evaluation frameworks, and produce work that influences model comparison, selection, and deployment. Collaboration with product and engineering is key, emphasizing depth and rigor.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist: LLM Evaluation & Routing Innovator
Research Scientist: LLM Evaluation & Routing Innovator

OpenRouter • New York (NY)

On-site
USD 170,000 - 250,000
Research Scientist
Research Scientist

OpenRouter • New York (NY)

On-site
USD 170,000 - 250,000
Research Scientist
Research Scientist

OpenRouter • United States

On-site
USD 150,000 - 230,000
Research Scientist
Research Scientist

OpenRouter, LLC • Northern (KY)

Hybrid
USD 140,000 - 230,000
Remote LLM Evaluation Scientist: Benchmarking Models
Remote LLM Evaluation Scientist: Benchmarking Models

Anyone AI • United States

Remote
MXN 2,626,000 - 3,678,000
Research Engineer — LLM Fine-Tuning & Model Routing
Research Engineer — LLM Fine-Tuning & Model Routing

Run BiOS • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Generative AI Research Scientist: LLM Post-Training
Generative AI Research Scientist: LLM Post-Training

Scale AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
Machine Learning Researcher – LLM
Machine Learning Researcher – LLM

Susquehanna International Group • Bala Cynwyd (PA)

On-site
USD 180,000 - 280,000
Research Engineer
Research Engineer

Nace.AI • Palo Alto (CA)

On-site
USD 120,000 - 160,000
RL Post-Training Scientist for LLMs & Tooling
RL Post-Training Scientist for LLMs & Tooling

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 167,000 - 226,000
Health insurance
401(k)