Research Scientist - Long-Horizon Multi-Agent Systems

techire ai

San Francisco (CA)

On-site

USD 400,000 - 450,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

techire ai is seeking a researcher to push the frontiers of long-horizon reasoning and RL, building systems that outperform current baselines. You will tackle problems without standard solutions and shape memory and reasoning components for real-world impact.

Ideal candidates hold a PhD with top-conference publications and have hands-on post-training RL experience, including RLHF or reward modelling, plus a track record in open-ended research.

Qualifications

  • PhD with publications at top conferences in long-horizon reasoning or RL.
  • Post-training experience (RLHF, DPO, reward modelling) preferred.
  • Experience working on open-ended research projects.

Responsibilities

  • Advance open-ended RL research into deployable systems and real-world impact.
  • Design environments and feedback loops for long-horizon tasks.
  • Collaborate across teams to translate ideas into working prototypes and systems.

Skills

PhD publications in long-horizon RL
RLHF / post-training experience
Open-ended research experience

Education

PhD in AI/ML or related

Job description

Rip up the playbook and step into uncharted territory.

If you've been building long-horizon multi-agent systems and pushing the boundaries of AI research, this is the kind of role where curiosity and ambition meet real execution, exploring truly novel problems at the frontier of what's currently possible.

You will work on systems designed to outperform the current state of the art, tackling problems that don't yet have standardised solutions across RL, long-horizon reasoning, LLM post-training for non-myopic objectives, environment and feedback design.

Whether you're early-career PhD or highly experienced, what matters most is your ability to push novel ideas into working systems, execute your knowledge across reasoning, RL and memory to make real-world impact.

This is a small, ambitious team operating where few others are, building and executing quickly in areas such as computational R&D science. This is your opportunity to shape the systems that generate and validate new discovery in environment primed for success.

Skills & experience
  • PhD and/or publications at top conferences across long-horizon reasoning, RL, or similar
  • Post-training experience (RLHF, DPO, reward modelling)
  • Experience working on open-ended research
Location

San Francisco

Salary

$400k base 0.5–1%+ equity Negotiable DOE

All applicants will receive a response.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer - Post training & RL
Research Engineer - Post training & RL

techire ai • California (MO)

Hybrid
USD 180,000 - 300,000
Equity
401k
Unlimited PTO
+1
Senior Research Scientist, Long-Horizon AI & RL (Equity)
Senior Research Scientist, Long-Horizon AI & RL (Equity)

techire ai • San Francisco (CA)

On-site
USD 400,000 - 450,000
Research Scientist - Multimodal Agent, Consumer Devices
Research Scientist - Multimodal Agent, Consumer Devices

OpenAI • United States

Hybrid
USD 140,000 - 230,000
Relocation assistance
Hybrid work model (3 days in office)
Research Scientist - Multimodal Agent, Consumer Devices
Research Scientist - Multimodal Agent, Consumer Devices

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 445,000
Hybrid work model
Relocation assistance
RESEARCH SCIENTIST
RESEARCH SCIENTIST

Good Start Labs • New York (NY)

Hybrid
USD 100,000 - 150,000
Direct collaboration with top universities
Visa sponsorship for exceptional candidates
$3.6M raised, two years funded
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 445,000
RESEARCHER (GENERAL)
RESEARCHER (GENERAL)

MakerMaker.AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Research Scientist
Research Scientist

techire ai • San Francisco (CA)

On-site
USD 250,000 - 400,000
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

SupportFinity™ • San Francisco (CA)

On-site
USD 100,000 - 180,000
RESEARCHER, AGENTS FOR AUTOMATED DISCOVERY
RESEARCHER, AGENTS FOR AUTOMATED DISCOVERY

MakerMaker.AI • San Francisco (CA)

On-site
USD 120,000 - 160,000