Applied ML Research Intern - Post-Training & RL Experiments

NextToken AI Inc.

San Francisco, Northern (CA, KY)

Ibrido

USD 41.000 - 83.000

Tempo pieno

14 giorni+
Generatore di candidature

Non inviare un curriculum generico — genera un curriculum e una lettera di presentazione personalizzati per questo specifico impiego.

Supera i filtri ATS

Descrizione del lavoro

NextToken AI Inc. seeks a motivated student or early graduate in CS/ML to contribute to evaluation-heavy ML research. You will design experiments using real production traces, build verifier and reward models, and run end-to-end post-training workflows with SFT, RL, and ablations.

The role emphasizes rigorous evaluation, bias-aware benchmarking, and a bias-to-ship mindset, aiming to translate research into deployed models or concrete results.

Competenze

  • Strong empirical ML background with hands-on experiments on LLMs or related models.
  • Ability to design and justify evaluation benchmarks that avoid biases.
  • Comfort working with real production traces, not curated datasets.

Mansioni

  • Build evaluation suites from production traces and define metrics.
  • Develop verifiers and reward models to guide training.
  • Run end-to-end post-training experiments (SFT/DPO/RL), ablations, and analysis.

Conoscenze

Empirical ML
RL experiments
Evaluation benchmarks
Production data handling
Diagnostic reasoning

Formazione

BS/MS/PhD in CS/ML

Strumenti

RL environments
Verifier models

Descrizione del lavoro

NextToken AI Inc. seeks a motivated student or early graduate in CS/ML to contribute to evaluation-heavy ML research. You will design experiments using real production traces, build verifier and reward models, and run end-to-end post-training workflows with SFT, RL, and ablations.

The role emphasizes rigorous evaluation, bias-aware benchmarking, and a bias-to-ship mindset, aiming to translate research into deployed models or concrete results.

Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

AI Research Intern: RL & LLM Post-Training
AI Research Intern: RL & LLM Post-Training

AMD • Santa Clara (CA)

Ibrido
USD 36.000 - 47.000
AI Research Intern: RL & LLM Post-Training
AI Research Intern: RL & LLM Post-Training

Advanced Micro Devices, Inc. • Santa Clara (CA)

Ibrido
USD 34.000 - 55.000
PhD AI Research Intern: RL & LLM Post-Training
PhD AI Research Intern: RL & LLM Post-Training

AMD • Santa Clara (CA)

Ibrido
USD 13.000 - 20.000
AMD Benefits at a glance
PhD AI Research Infra Intern — RL Post-Training
PhD AI Research Infra Intern — RL Post-Training

Advanced Micro Devices • Santa Clara (CA)

Ibrido
USD 50.000 - 67.000
AI Research Intern — RL & LLM Post-Training (PhD)
AI Research Intern — RL & LLM Post-Training (PhD)

Advanced Micro Devices • Santa Clara (CA)

In loco
USD 55.000 - 83.000
PhD AI Research Infra Intern: RL Post-Training
PhD AI Research Infra Intern: RL Post-Training

AMD • Santa Clara (CA)

Ibrido
USD 55.000 - 90.000
RESEARCHER, POST-TRAINING
RESEARCHER, POST-TRAINING

MakerMaker.AI • San Francisco (CA)

In loco
USD 120.000 - 160.000
Remote Applied Scientist Intern: AI & ML Research
Remote Applied Scientist Intern: AI & ML Research

OhioX • Northern (KY)

Ibrido
USD 141.000 - 150.000
Flexible remote work
Annual equity grants
401(k) match
+2
Applied ML Scientist Intern - End-to-End AI
Applied ML Scientist Intern - End-to-End AI

Ramp Corp. • New York (NY)

In loco
USD 11.000 - 14.000
Apple MacBook
Catered lunches in NYC office
Weekly coffee stipend
Member of Technical Staff - Research Engineer, Post-training
Member of Technical Staff - Research Engineer, Post-training

Preference Model • San Francisco (CA)

In loco
USD 180.000 - 240.000
Competitive cash and equity (>90th pct
Ownership and autonomy
Lunch onsite
+4