Get more replies from employers
Send a job-specific resume in minutes.
A1 is hiring a Backend Engineer, AI, in Palo Alto to own the inference and orchestration layer powering AI interactions in our product. You will build and operate production systems that turn model capabilities into fast, stable APIs accessed by mobile and desktop clients.
You will design multi-step inference pipelines, manage tool calls and retries, and optimize routing, caching, batching, streaming, and state management to meet latency and throughput goals.
There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to conversations, errands, organising and workflows, with minimal prompting.
Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. The system must handle multi-step reasoning, interact with external tools, and remain reliable despite non-deterministic model behavior. Our objective is to help users complete tasks daily enjoyable with over ~90%* reduced time.
As a Backend Engineer, AI, you own the inference and orchestration layer that powers every AI interaction in the product. Your work sits between models and users, where latency, correctness, reliability, and cost directly impact real-world experience. Build and operate production systems that turn model capability into fast, stable, observable APIs used across mobile and desktop clients.
The best products today in the world were built by small, world class teams.
We are a high talent density and hands-on team. We make decisions collectively, move at rapid speed, striking a balance between shipping high quality work and learning.
Joining our team requires the ability to bring structure, exercise judgment, and execute independently. Our goal is to put in hands of our users a truly magical product