AI Automation Engineer - End-to-End Evaluation

Niantic

San Francisco (CA)

Hybrid

USD 158,000 - 210,000

Full time

47 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Niantic Spatial is hiring an AI Automation Engineer to design and own the end-to-end evaluation system that produces reproducible evidence across releases. You will automate evaluations, instrument scorecards, and build agent workflows to simulate real-world use cases at scale.

You will work with a data manager to ensure reproducible datasets and drive automated diagnostics across the platform. The role emphasizes engineering-first evaluation, building scalable automation, and delivering trusted

Qualifications

  • Built and maintained production-grade automation or test infrastructure that other engineers relied on daily.
  • Experience evaluating systems where correctness is graded rather than binary, meaning quality, accuracy, or latency thresholds rather than pass/fail assertions.
  • Strong Python, with fluency in orchestration, CI-style pipelines, cloud storage, and reproducible environments.
  • Worked with backend services and APIs you did not own, integrating against them without becoming a bottleneck for their team.
  • Written up a technical finding clearly enough that a non-author could act on it without a meeting.
  • A bachelor's degree in a relevant field, or equivalent experience.

Responsibilities

  • Build the Evaluation Machine - Own the automation that executes Lab evaluations end to end: environment setup, run orchestration, artifact capture, and result collection.
  • Make Results Comparable - Instrument scorecards so a result can be compared across product versions, devices, and capture conditions.
  • Automate the Agent-Driven Layer - Build the agent workflows that exercise priority customer use cases at realistic scale, across the graded difficulty spectrum from easy to frontier.
  • Kill Manual Work Permanently - Convert one-off experiments into standing protocols that run on every relevant release.
  • Make Failures Actionable - Produce diagnostics precise enough that a finding reaches its owner with a reproducible case and data attached.
  • Keep Data From Being the Bottleneck - Work with the AI Data Manager so every run is reproducible from a known dataset state.

Skills

Python
Orchestration
CI/CD
Cloud storage
APIs

Education

Bachelor's degree

Job description

Niantic Spatial is hiring an AI Automation Engineer to design and own the end-to-end evaluation system that produces reproducible evidence across releases. You will automate evaluations, instrument scorecards, and build agent workflows to simulate real-world use cases at scale.

You will work with a data manager to ensure reproducible datasets and drive automated diagnostics across the platform. The role emphasizes engineering-first evaluation, building scalable automation, and delivering trusted

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Automation Engineer: End-to-End Evaluations + Equity
AI Automation Engineer: End-to-End Evaluations + Equity

Niantic Spatial, Inc. • San Francisco (CA)

On-site
USD 158,000 - 210,000
Annual bonus
Equity
Medical coverage
+3
AI Automation Engineer, Real-World Test Lab
AI Automation Engineer, Real-World Test Lab

Niantic • San Francisco (CA)

Hybrid
USD 158,000 - 210,000
AI Automation Engineer, Real-World Test Lab
AI Automation Engineer, Real-World Test Lab

Niantic Spatial, Inc. • San Francisco (CA)

On-site
USD 158,000 - 210,000
Annual bonus
Equity
Medical coverage
+3
Robotics Hardware Operations Engineer
Robotics Hardware Operations Engineer

Niantic • San Francisco (CA)

Hybrid
USD 142,000 - 193,000
AI Evaluation & Safety Engineer
AI Evaluation & Safety Engineer

Nuna • San Francisco (CA)

On-site
USD 150,000 - 230,000
AI Engineer, Computer Vision — Localization & Production
AI Engineer, Computer Vision — Localization & Production

Niantic • San Francisco (CA)

Hybrid
USD 166,000 - 221,000
Senior AI Test & Evaluation Engineer
Senior AI Test & Evaluation Engineer

Motion • Birmingham (AL), Northern (KY)

Hybrid
USD 120,000 - 180,000
Senior Software Engineer: AI Agent Training & Evaluations
Senior Software Engineer: AI Agent Training & Evaluations

Prolific • United States

Remote
USD 150,000 - 190,000
Remote working
Competitive salary
AI Systems Observability Engineer — Evaluation & Readiness
AI Systems Observability Engineer — Evaluation & Readiness

NTT DATA North America • Charlotte (NC)

On-site
USD 130,000 - 185,000
Remote AI Evaluation & Quality Assurance Engineer
Remote AI Evaluation & Quality Assurance Engineer

Equiliem • United States

Remote
USD 69,000 - 74,000