Applied Scientist: Agent Evaluation & Adaptive Routing

Bitdeer

Austin, Northern (TX, KY)

Hybrid

USD 125,000 - 170,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Mentoring program
Training opportunities
Competitive benefits

Job summary

Bitdeer AI Lab in Austin is seeking a senior researcher to build evaluation pipelines for agentic inference and to extend our LLM and agent evaluation frameworks. You will design task-level and trajectory-level metrics, prototype routing strategies, and drive production-ready research artifacts.

The role requires deep expertise in Python, PyTorch, and evaluating multi-turn, tool-using agents, with a track record of robust experimental design and open-source contributions.

Qualifications

  • Hands-on experience in LLM evaluation and adaptive inference.
  • Strong Python programming and PyTorch experience; able to build reliable experimental pipelines.
  • Experience designing evaluation pipelines for LLMs or agents, including task-level and trajectory-level metrics.
  • Depth in model selection and routing, uncertainty estimation/calibration, cascading/escalation, or stage-aware inference.

Responsibilities

  • Build evaluation and decision systems for agentic inference; own LLM/agent evaluation pipeline.
  • Research and prototype adaptive model-routing strategies and escalation policies.
  • Collaborate with MaaS and platform teams on production integration.

Skills

Python programming
LLM evaluation
Agentic systems
Uncertainty estimation

Education

Bachelor’s/Master’s/PhD in CS/ML/Statistics/EE

Tools

PyTorch
vLLM
SGLang

Job description

Bitdeer AI Lab in Austin is seeking a senior researcher to build evaluation pipelines for agentic inference and to extend our LLM and agent evaluation frameworks. You will design task-level and trajectory-level metrics, prototype routing strategies, and drive production-ready research artifacts.

The role requires deep expertise in Python, PyTorch, and evaluating multi-turn, tool-using agents, with a track record of robust experimental design and open-source contributions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Applied Scientist, Agent Evaluation & Adaptive Model Routing
Applied Scientist, Agent Evaluation & Adaptive Model Routing

Bitdeer • Austin (TX), Northern (KY)

Hybrid
USD 125,000 - 170,000
Mentoring program
Training opportunities
Competitive benefits
Research Scientist: Efficient AI Inference
Research Scientist: Efficient AI Inference

Bitdeer (NASDAQ: BTDR) • Austin (TX)

On-site
USD 140,000 - 220,000
RESEARCHER, AGENTS FOR AUTOMATED DISCOVERY
RESEARCHER, AGENTS FOR AUTOMATED DISCOVERY

MakerMaker.AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Research Engineer: AGI Eval & Agent Quality
Research Engineer: AGI Eval & Agent Quality

Pantera Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive cash compensation
Equity opportunities
Relocation support
+1
Senior AI Engineer: Production LLM Agents & Inference
Senior AI Engineer: Production LLM Agents & Inference

Bitus Labs • Irvine (CA)

Hybrid
USD 140,000 - 190,000
Senior AI Engineer — Inference & Agent Systems
Senior AI Engineer — Inference & Agent Systems

Arcana Analytics • United States

On-site
USD 120,000 - 160,000
Senior AI Engineer - Agent Team
Senior AI Engineer - Agent Team

FurtherAI Inc • San Francisco (CA)

On-site
USD 120,000 - 150,000
Fully covered health, dental, and vision benefits
Competitive Compensation and stock options
Unlimited PTO
+4
Senior AI Infrastructure Engineer: LLM Agents
Senior AI Infrastructure Engineer: LLM Agents

NVIDIA • Austin (TX)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior AI Agent Simulation & Evaluation Engineer
Senior AI Agent Simulation & Evaluation Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Research Scientist, Agentic Data & Benchmarking
Research Scientist, Agentic Data & Benchmarking

Institute of Foundation Models • Sunnyvale (CA)

On-site
USD 180,000 - 250,000