Research Engineer, Agent Systems — Frontier AI Lab

Aionia Group

San Francisco (CA)

On-site

USD 300,000 - 600,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Meaningful equity
Top-of-market compensation
Collaborative environment with researchers

Job summary

Aionia Group in San Francisco is looking for a Research Engineer, Agent Systems. This role involves developing foundational systems that ensure agent reliability and safety in real-world applications. You will work directly with top researchers in a mission-driven environment.

Offering a competitive compensation range of $300K to $600K plus equity, the position emphasizes engineering quality and contributes directly to advancements in AI safety.

Qualifications

  • Experience building and operating complex systems in production.
  • Ability to debug complex systems and identify root causes.
  • Strong backend engineering foundation.

Responsibilities

  • Build and evolve agent execution frameworks.
  • Develop infrastructure for reliability and capability measurement.
  • Design systems for routing and planning with correctness guarantees.

Skills

Debugging complex systems
Building complex production systems
Agile environments
Data-driven product improvement

Tools

Python
Distributed Systems
ML Pipelines

Job description

Research Engineer, Agent Systems

One of the most mission-driven organizations in AI is building the infrastructure that makes intelligent agents safe and reliable in the real world. This is not an application layer role. You’ll work directly alongside researchers to build the execution layer that determines how agents reason, act, fail, recover, and improve in production.

$300K – $600K+ Total Comp + Equity San Francisco · On-Site Frontier AI Lab · Confidential Highly Selective · 1 Engineer No Visa Sponsorship

Where the Frontier Is Actually Being Built

We’re partnering with a mission-driven frontier AI lab focused on building safe, reliable intelligent systems — not shipping products, not chasing growth metrics. The work is foundational. The team is small. The impact is real.

The organization operates with a founding team at the forefront of AI safety and superintelligence research, a deliberate research-first culture where engineering quality is non-negotiable, and a small team where every engineer shapes the direction of the system.

Founding team from the top of the AI research world

Small

Every hire shapes system architecture — no passengers

Mission

Research-first culture — engineering quality is non-negotiable

$300K+

Top-of-market comp with meaningful equity for the right person

“This lab exists to get superintelligence right. If that motivates you, you’ll find this environment unlike anywhere else.”

The Work

What You’ll Build

This is a research engineering role at the frontier. You’ll translate model insights into production-grade systems — sitting at the boundary between what researchers discover and what actually runs reliably in the world.

  • Build and evolve agent execution frameworks used directly in research and production
  • Develop evaluation infrastructure that measures reliability, capability, and safety together
  • Design control-plane systems for routing, planning, and tool use with strong correctness guarantees
  • Build feedback loops that close the gap between offline evaluation and real-world behavior
  • Create observability and simulation systems that make failure modes visible and fixable
  • Work directly with researchers to turn experimental insights into stable, scalable systems
  • Contribute to sandboxed environments where agents can operate and self-validate safely
  • Continuously adapt orchestration systems as model capabilities evolve
Stack & Tools

Python Distributed Systems Eval Frameworks Agent Orchestration Observability Async Execution ML Pipelines

What They’re Looking For

Must-Have — Non-Negotiable

  • Experience building and operating complex systems in production — reliability under real-world pressure
  • Ability to debug complex systems and identify root causes of failures, not just symptoms
  • Comfort working in ambiguous, fast-moving environments where the problems are genuinely unsolved
Required
  • Familiarity with experimentation, evaluation, or data-driven product improvement loops
  • Experience working closely with researchers or data scientists — translating their needs into reliable infrastructure
  • Experience owning systems end-to-end, from design through production and iteration
  • Strong backend engineering foundation — distributed systems, async execution, observability
Even Better If
  • You’ve built or worked on agent harnesses, orchestration layers, or execution frameworks
  • You think in terms of control planes, feedback loops, and system-level optimization — not just features
  • You’re excited about diagnosing failure modes and iterating toward measurable improvements
  • You care deeply about production quality — not just making systems work, but making them reliable, safe, and scalable
  • You’re motivated by pushing the frontier of how intelligent systems behave in the real world
Why This Role Is Different
  • Safety is a first-class engineering concern. Not a compliance layer added at the end — a design constraint built into every system from the start.
  • You’ll work with researchers, not around them. This is a true research engineering role — your infrastructure directly enables the science.
  • Small team, outsized leverage. Every architectural decision you make shapes how intelligent systems behave in the world.
  • The mission is the point. This lab exists to get superintelligence right. If that drives you, there is no comparable environment.
Compensation

$300,000 – $600,000+ total compensation depending on level, with meaningful equity at an organization of this caliber. This is not typical startup equity — this is a lab building toward one of the most consequential outcomes in technology.

Top-of-Market Base
Meaningful Equity
On-Site
San Francisco
Small Elite Team

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Agentic Infrastructure Engineer — Frontier AI Lab
Agentic Infrastructure Engineer — Frontier AI Lab

Aionia Group • New York (NY)

On-site
USD 250,000 - 500,000
Competitive equity
On-site collaboration with a small team
High compensation package
Research Engineer — AI Alignment & Evaluation
Research Engineer — AI Alignment & Evaluation

W3 Sourcing • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Senior Software Engineer
Senior Software Engineer

Evolve Group • New York (NY)

On-site
USD 200,000 - 400,000
Software Engineer
Software Engineer

Venture Up • San Francisco (CA)

On-site
USD 350,000 - 600,000
Equity
Office in San Francisco FiDi
On-site with optional remote Sunday (½
Head of AI Red Teaming
Head of AI Red Teaming

Trajectory Labs, PBC • Berkeley (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Equity
Health coverage
401(k)
+1
Frontier Agents Engineer (Applied AI) Software
Frontier Agents Engineer (Applied AI) Software

Front Door Defense • New York (NY)

On-site
USD 180,000 - 225,000
Staff Frontier Agents Engineer (Applied AI) Software
Staff Frontier Agents Engineer (Applied AI) Software

Front Door Defense • New York (NY)

On-site
USD 252,000 - 315,000
Comprehensive health
Dental and vision coverage
Retirement benefits
+2
Applied Research Scientist
Applied Research Scientist

Fleet AI, Inc. • Buffalo (NY)

On-site
USD 150,000 - 210,000
Senior Frontier Agents Engineer (Applied AI) Software
Senior Frontier Agents Engineer (Applied AI) Software

Front Door Defense • New York (NY)

On-site
USD 216,000 - 270,000
Health, dental, vision coverage
Generous PTO and retirement benefits
Learning and development stipend
Senior Frontier Agents Engineer (Applied AI)
Senior Frontier Agents Engineer (Applied AI)

Scale AI • New York (NY)

On-site
USD 216,000 - 270,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2