AI Engineer, Internal Systems

Menlo Ventures

United States

On-site

USD 140,000 - 180,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Wispr is seeking to design and build the verification and evaluation layer that lets agents finish their work with confidence. You will be the first to create end-to-end verification loops and fast, reliable testing environments to turn feedback into durable evaluation signals.

You will focus on improving how context, instructions, skills, and memory are delivered to agents, and you will operate the agent fleet as a production system to measure quality, adoption, and impact.

Qualifications

  • You've built internal tools that other engineers adopted, and you can explain how you measured their impact
  • You use agentic coding tools deeply and have specific opinions about where they succeed and fail
  • You've built evaluation or verification infrastructure such as test harnesses, CI systems, eval pipelines, or benchmarks
  • You can work hands-on across application code, infrastructure, and unfamiliar systems
  • You are empirical about your own work, attentive to subtle failure modes, and comfortable owning ambiguous problems

Responsibilities

  • Build end-to-end verification loops that let agents determine when their work is actually complete
  • Create fast, reliable testing environments and turn failures and human feedback into durable evaluation signals
  • Improve how context, instructions, skills, and memory are delivered to agents—and measure what makes them perform better
  • Operate the agent fleet as a production system and measure its quality, adoption, and impact

Job description

About Wispr

Wispr is an AI research and product company building the voice interface for computing. AI can now reason, code, and act. Yet, humans still do the work of the interface. We think that’s backwards.
Our first products are Flow, which lets you speak naturally in any application, and Notetaker, which builds context across conversations. We’re building toward an interface that can perceive, understand, and take action with earned trust. That means solving hard problems across models, systems, and product - and caring about the final human experience as deeply as the technology underneath it.



We’re a talent-dense team that holds strong opinions, tests them quickly, and builds technology that sparks joy. Our goal is to build the first voice interface used every day by a billion people.


About the role

Our agents already write code and respond to incidents around the clock, but they can't yet tell when they're done, so every loop still ends with a human checking. You'd be the first to build the verification and evaluation layer that lets them finish with confidence.


What you'll do


  • Build end-to-end verification loops that let agents determine when their work is actually complete


  • Create fast, reliable testing environments and turn failures and human feedback into durable evaluation signals


  • Improve how context, instructions, skills, and memory are delivered to agents—and measure what makes them perform better


  • Operate the agent fleet as a production system and measure its quality, adoption, and impact



You may be a fit if


  • You've built internal tools that other engineers adopted, and you can explain how you measured their impact


  • You use agentic coding tools deeply and have specific opinions about where they succeed and fail


  • You've built evaluation or verification infrastructure such as test harnesses, CI systems, eval pipelines, or benchmarks


  • You can work hands-on across application code, infrastructure, and unfamiliar systems


  • You are empirical about your own work, attentive to subtle failure modes, and comfortable owning ambiguous problems



Prior experience in voice, model training, or our product domains is not required.


Get to know us


  • Our Series B


  • Advancing HCI


  • Introducing Wispr Advanced Interfaces Lab


  • Technical challenges behind Flow



Logistics

We sponsor H1B, O1, EB1, L1, STEM OPT, and more. We can't guarantee sponsorship for every role, but if we make you an offer we’ll make every reasonable effort, with help from our immigration law firm.
The strongest candidates we meet rarely do, so don’t exclude yourself prematurely. We’re rethinking how humans interact with AI, and doing that well demands diversity of perspective and experience.
We're committed to a fair and accessible interview process. If you need any accommodations or adjustments, please let us know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Engineer, Internal Systems
AI Engineer, Internal Systems

Wispr Flow • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Product Engineer, Frontend
Product Engineer, Frontend

Wispr Flow • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Product Engineer, Frontend
Product Engineer, Frontend

Menlo Ventures • United States

On-site
USD 90,000 - 150,000
Platform Engineer, Enterprise
Platform Engineer, Enterprise

Wispr Flow • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Platform Engineer, Enterprise
Platform Engineer, Enterprise

Menlo Ventures • United States

On-site
USD 130,000 - 210,000
Product Engineer, Systems Architect
Product Engineer, Systems Architect

Wispr Flow • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Product Engineer, Systems Architect
Product Engineer, Systems Architect

Menlo Ventures • United States

On-site
USD 150,000 - 190,000
Product Engineer, Growth
Product Engineer, Growth

Menlo Ventures • United States

On-site
USD 120,000 - 180,000
Product Engineer, Growth
Product Engineer, Growth

Wispr Flow • San Francisco (CA)

On-site
USD 120,000 - 170,000
Forward Deployed Engineer
Forward Deployed Engineer

Wispr Flow • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000