Lead RL & Agentic AI Research—On-Device & Tools

Socket.dev

Cupertino (CA)

On-site

USD 220,000 - 320,000

Full time

6 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Socket.dev is seeking a hands-on research lead to guide reinforcement learning and post-training initiatives for agentic AI, while managing a small team of senior researchers across RL, tool-calling, synthetic environments, and multimodal action models.

You will shape the infrastructure, training, runtime, and evaluation strategies for interactive agents, and remain actively involved in experiments, coding, and publishing.

Qualifications

  • PhD in machine learning or related field, or equivalent research experience.
  • 7-10+ years of research experience beyond PhD in industry or academia.
  • Strong track record in RL and/or post-training of large models, evidenced by publications, open-source contributions, or shipped systems.
  • Leadership experience coordinating researchers and engineers over multiple years.

Responsibilities

  • Lead RL and post-training research for agentic capabilities, including reward, preference optimization, and verifier design.
  • Build and own synthetic data and task-generation pipelines and the environments and benchmarks that accompany them.
  • Drive codebases and infrastructure for the core RL research effort and coordinate with partner teams.
  • Manage and mentor a small team (3–4) of senior researchers and research engineers, shaping a shared direction while preserving bottom-up work.
  • Stay hands-on: run experiments, write code, and contribute directly to the team's core technical problems.
  • Bridge post-training research with deployment considerations for on-device and hybrid compute environments.
  • Collaborate across the organization on environment-agent co-optimization, self-improvement, and long-context agents.
  • Publish in top venues and engage with the broader research community.

Skills

RL research
Leadership
ML infrastructure
Hands-on engineering

Education

PhD or equivalent

Tools

PyTorch
RL frameworks

Job description

Socket.dev is seeking a hands-on research lead to guide reinforcement learning and post-training initiatives for agentic AI, while managing a small team of senior researchers across RL, tool-calling, synthetic environments, and multimodal action models.

You will shape the infrastructure, training, runtime, and evaluation strategies for interactive agents, and remain actively involved in experiments, coding, and publishing.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead RL Researcher for Agentic AI & Environments
Lead RL Researcher for Agentic AI & Environments

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 216,000 - 394,000
Apple stock programs
Medical coverage
Retirement benefits
+1
AIML - Machine Learning Research Lead, RL Agents, MLR
AIML - Machine Learning Research Lead, RL Agents, MLR

Socket.dev • Cupertino (CA)

On-site
USD 220,000 - 320,000
Post-Training AI Research Engineer – RL & Agentic Infra
Post-Training AI Research Engineer – RL & Agentic Infra

Storm3 • San Francisco (CA)

On-site
USD 140,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Senior Staff Scientist, Agentic AI & RL — Lead AI Systems
Senior Staff Scientist, Agentic AI & RL — Lead AI Systems

Centific Global Solutions, Inc. • Palo Alto (CA)

On-site
USD 150,000 - 200,000
Remote RL Research Intern — Agentic AI & LLMs
Remote RL Research Intern — Agentic AI & LLMs

Centific Global Solutions, Inc. • United States

On-site
Competitive stipend
Mentorship from researchers
Access to modern GPU infrastructure
Remote AI Research Engineer — Agentic Post-Training
Remote AI Research Engineer — Agentic Post-Training

Tether.io • Town of Italy (NY)

On-site
Staff Software Engineer - RL Frameworks & Tooling
Staff Software Engineer - RL Frameworks & Tooling

Anthropic Limited • New York (NY)

Hybrid
USD 405,000 - 625,000
Staff AI Engineer - Post-Training & Agentic RL
Staff AI Engineer - Post-Training & Agentic RL

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
Remote AI Research Engineer: Agentic Models & Tooling
Remote AI Research Engineer: Agentic Models & Tooling

Visa Hunt • United States

Remote
USD 120,000 - 190,000
RL & Agentic AI Researcher — Benchmarking Data Quality
RL & Agentic AI Researcher — Benchmarking Data Quality

Protege • United States

On-site
USD 110,000 - 140,000