Research Intern, Agent RL Training

NewsBreak

Mountain View (CA)

In loco

USD 48.216 - 68.880

Tempo pieno

14 giorni+
Generatore di candidature

Ricevi una risposta da questo datore di lavoro — un curriculum e una lettera di presentazione personalizzati, che corrispondono esattamente a ciò che sta cercando.

Supera i filtri ATS

Descrizione del lavoro

NewsBreak is seeking a Research Intern to join their Agent RL Training team in Mountain View, CA. In this hands-on role, you will work closely with a mentor to apply large language models to impactful projects, driving experiments and contributing to publications.

The ideal candidate is passionate about research, independent, and skilled in Python and PyTorch. Compensation ranges from $35 to $50 per hour based on experience.

Competenze

  • Highly motivated and committed, willing to put in extra hours when needed.
  • Genuine passion for research and model experimentation.
  • Capable of end-to-end model SFT with understanding of RL-based methods.
  • Strong reasoning about effective model behavior.

Mansioni

  • Collaborate with a mentor on research directions for applying LLMs.
  • Run end-to-end SFT experiments on LLM-based agents.
  • Curate high-quality training datasets.
  • Contribute to public publications.

Conoscenze

Strong Python
PyTorch skills
Research passion
Independent problem-solving

Descrizione del lavoro

Mountain View, California, United States

About NewsBreak

Founded in 2015, NewsBreak is the Content Intelligence platform shaping the future content economy. With over 40 million monthly active users, our flagship platform delivers highly personalized local news and information powered by advanced AI, recommendation systems, and adtech.

Recognized by Fast Company as #32 on the Top Workplaces for Innovators, we're proud to be Great Place to Work® certified and home to a dynamic team of technologists, product innovators, and business leaders who are passionate about solving meaningful challenges at scale.

Together, we reached unicorn status in 2021, and we remain committed to continuing this high-growth trajectory with the right team to fulfill our mission: building the infrastructure layer for content intelligence.

About the Role

We are looking for a Research Intern to join our Agent RL Training team. You will be paired with a full-time employee as your mentor, working together to explore how to apply large language models to NewsBreak’s core business, including content understanding, recommendation, agentic web browsing, and autonomous multi‑step task completion.

This is a hands‑on research role. You are expected to independently drive experiments, propose novel ideas, and iterate quickly. We value self‑starters with deep intellectual curiosity and the drive to push boundaries in LLM post‑training and agent capabilities.

Location: Onsite in Mountain View, CA office

What You’ll Work On
  • Collaborate with your full‑time mentor to identify high‑impact research directions for applying LLMs to NewsBreak’s products
  • Independently run end‑to‑end SFT experiments on LLM‑based agents, and assist with RL‑related exploration such as reward design and training iteration
  • Curate and build high‑quality training datasets: instruction‑following, preference pairs, agent trajectories, and synthetic data
  • Contribute to public publications; we encourage and support top‑venue submissions during your internship
What We’re Looking For
Requirements
  • Highly motivated and committed: willing to put in extra hours when needed to push projects across the finish line
  • Genuine passion for research: you read papers for fun, tinker with models on weekends, and care deeply about advancing the field
  • Independently capable of end‑to‑end model SFT: with basic understanding of RL‑based post‑training methods (RLHF, DPO, PPO, GRPO, etc.)
  • Excellent taste in model behavior: able to reason about what “good” looks like across user‑facing domains and articulate why
  • Strong Python and PyTorch skills
Preferred Qualifications
  • Publication at a top‑tier venue (NeurIPS, ICML, ICLR, ACL, EMNLP, or equivalent)
  • Experience with multi‑node distributed training (FSDP, DeepSpeed, Megatron‑LM)
  • Proficiency in writing custom GPU kernels with Triton or CUDA
  • Experience building synthetic data pipelines for agent training
  • Familiarity with open‑source RL frameworks: TRL, OpenRLHF, veRL/vLLM

Hourly Pay: $35–$50
The US base salary range for this full‑time position is listed below. Pay may vary based on a number of factors including job‑related skills, level, experience, geographic location and relevant education or training. At NewsBreak, we design our overall rewards package to attract top talents. Depending on the position, the role may also be eligible for discretionary bonus and options. Your recruiter can share more details during the hiring process.

Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Research Intern — LLM Agents & RL Training
Research Intern — LLM Agents & RL Training

NewsBreak • Mountain View (CA)

In loco
USD 48.216 - 68.880
Machine Learning Engineer, LLM Post-Training
Machine Learning Engineer, LLM Post-Training

NewsBreak • Mountain View (CA)

In loco
USD 130.000 - 160.000
Health, dental, and vision care
401(k) plan with company matching
Paid time off and holidays
Nearby AI Internship Program - Engineering Track
Nearby AI Internship Program - Engineering Track

News Break • Mountain View (CA)

In loco
USD 60.000 - 140.000
Research Intern
Research Intern

Quadrillion Labs • New York (NY)

In loco
USD 300.000 - 500.000
Medical, dental, and vision insurance
Lunch and dinner covered
Other varied stipends
AI Engineer, Agent Platform
AI Engineer, Agent Platform

NewsBreak • Mountain View (CA)

In loco
USD 120.000 - 220.000
Health coverage for you and family
Top-tier 401(k) plan with companymatch
Paid time off and holidays
+2
AI Engineer, Agent Platform
AI Engineer, Agent Platform

News Break • Mountain View (CA)

In loco
USD 120.000 - 220.000
Health, dental, and vision care forYou
401(K) matching program
Paid time off and holidays
+1
Machine Learning Engineer, LLM Post-Training
Machine Learning Engineer, LLM Post-Training

News Break • Mountain View (CA)

In loco
USD 150.000 - 230.000
Health, dental, and vision care for you and your family
Top-tier 401(K) plan with company matching
Paid time off and paid holidays
+2
Software Engineer, ML Infra (Junior & New Grad)
Software Engineer, ML Infra (Junior & New Grad)

NewsBreak • Mountain View (CA)

In loco
USD 125.000 - 175.000
Machine Learning Researcher
Machine Learning Researcher

Multicoin • San Francisco (CA)

In loco
USD 250.000 - 350.000
Competitive compensation
Equity in a high-growth startup
Comprehensive benefits
Machine Learning Researcher
Machine Learning Researcher

SOLANA FOUNDATION • San Francisco (CA)

In loco
USD 250.000 - 350.000
Equity in a high-growth startup
Comprehensive benefits