Production AI Engineer: LLM, RAG & Multi-Agent Systems

Zof AI

San Francisco, Northern (CA, KY)

Hybrid

USD 140,000 - 190,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Zof AI in San Francisco, CA is seeking an Applied AI Engineer to build product features on top of frontier model APIs. This role owns the model-facing layer of our products, including retrieval and RAG pipelines, agent planning and orchestration, multi-agent frameworks, and the context engineering that makes AI systems reliable in production.

You have shipped LLM-powered features to real users and treat quality, cost, and latency as engineering constraints.

Qualifications

  • Shipping LLM-powered features to production.
  • Strong software engineering foundation.
  • Knowledge of retrieval, RAG, and agent patterns.
  • Fluency with modern model APIs.
  • Ability to balance quality, cost, and latency.

Responsibilities

  • Design and build product features on top of frontier model APIs.
  • Build and tune retrieval and RAG pipelines end to end.
  • Design agent planning, tool use, and multi-agent orchestration.
  • Own prompt and context engineering as a disciplined, tested practice.
  • Wire evals into the development loop so quality is measured, not assumed.
  • Optimize the cost, quality, and latency of AI features in production.
  • Collaborate with product and engineering to ship reliably and fast.
  • Own model-layer systems from design through production.

Skills

LLM features
Software engineering
Model APIs
Retrieval & RAG

Tools

TypeScript
Python
Node.js
Postgres

Job description

Zof AI in San Francisco, CA is seeking an Applied AI Engineer to build product features on top of frontier model APIs. This role owns the model-facing layer of our products, including retrieval and RAG pipelines, agent planning and orchestration, multi-agent frameworks, and the context engineering that makes AI systems reliable in production.

You have shipped LLM-powered features to real users and treat quality, cost, and latency as engineering constraints.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Applied AI Engineer
Applied AI Engineer

Zof AI • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Remote AI/ML Engineer — LLMs, RAG & Multi-Agent
Remote AI/ML Engineer — LLMs, RAG & Multi-Agent

YO AI Labs • Phoenix (AZ)

Remote
USD 120,000 - 180,000
Applied AI Engineer - LLM Orchestration & Production AI
Applied AI Engineer - LLM Orchestration & Production AI

Radley James • San Francisco (CA)

On-site
USD 180,000 - 275,000
Production AI Engineer: LLMs, RAG & Observability
Production AI Engineer: LLMs, RAG & Observability

GeniusXLab LLC • United States

On-site
USD 140,000 - 210,000
Learning budget
Remote work
Competitive pay
Enterprise AI Research Engineer LLMs RAG & Agents
Enterprise AI Research Engineer LLMs RAG & Agents

Fabrion • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
Remote AI/ML Engineer for Secure LLM & Multi-Agent Systems
Remote AI/ML Engineer for Secure LLM & Multi-Agent Systems

YO AI Labs • Town of Texas (WI)

Remote
USD 120,000 - 180,000
AI Agent Architect - Production-Grade LLMs (Hybrid)
AI Agent Architect - Production-Grade LLMs (Hybrid)

Goliath Partners Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 220,000 - 275,000
Hybrid work model
Ownership in technical direction
Significant autonomy
AI Product Engineer: Production Agents & LLMs
AI Product Engineer: Production Agents & LLMs

Samba TV • San Francisco (CA)

On-site
USD 150,000 - 200,000
Remote AI/ML Engineer: Scalable LLMs & Multi-Agent Apps
Remote AI/ML Engineer: Scalable LLMs & Multi-Agent Apps

YO AI Labs • North Carolina

Remote
USD 120,000 - 160,000
Senior AI/ML Engineer: Agentic LLMs, RAG & Production ML
Senior AI/ML Engineer: Agentic LLMs, RAG & Production ML

Orbien LLC • Maryland

On-site
USD 150,000 - 230,000