Harness Engineer

AI Fabrik

San Francisco (CA)

On-site

USD 150,000 - 190,000

Full time

9 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

AI Fabrik is hiring to own our engineering harness—the tooling and AI workflows that run across the full delivery pipeline. You will identify where teams lose time, build AI-assisted harnesses, and measure success by faster shipping with fewer handoffs.

The role focuses on agent environments, release workflows, and performance-driven improvements to the end-to-end software delivery cycle. This onsite opportunity is based in Redwood City, CA.

Qualifications

  • 3+ years of professional software engineering experience in tooling, platform engineering, or DevOps.
  • Proven ability to measure and improve cycle time and deployment frequency.
  • Proficiency in Python or Go and experience with at least one major LLM API (OpenAI, Anthropic, or Google).
  • Strong understanding of production reliability patterns and CI/CD practices.

Responsibilities

  • Identify bottlenecks across the full software delivery lifecycle and build AI-assisted harnesses that remove them, measuring impact in cycle time, lead time, defect rate, and deployment frequency.
  • Build harnesses that accelerate from idea to implementation, including AI-assisted requirement elaboration and automated design review.
  • Design and maintain agent environments for code generation, automated refactoring, test authoring, and documentation; define constraints and tool permissions.
  • Build agent-assisted release workflows, automated pre-deployment checklists, rollback triggers, and progressive rollout controls.
  • Define good output at each stage through evaluation datasets and CI/CD–integrated quality gates.
  • Instrument harness components with logging and metrics to identify underperformance and future productivity gains.
  • Implement boundaries and guardrails for safe agent operation without sacrificing speed.
  • Apply token budgeting, caching strategies, and model tiering to control AI costs.

Skills

Developer tooling
Platform engineering
DevOps
Python/Go
LLM API integration
CI/CD pipelines
Observability
Documentation

Job description

AI Fabrik builds an edge inference delivery network for high-performance tokens, with faster time-to-market from grid to tokens. Our mission is to build the inference infrastructure we wished every enterprise already had — close to users, close to the cloud, and extremely resilient for real-time workloads. We are builders, architects, engineers, and researchers with hands-on experience in real-world AI deployment in production, and decades of data center experience that taught us exactly what needs to change.

AI Fabrik was incubated inside Gruve and backed by Mayfield, Xora (Temasek), Acclimate Ventures, Cisco Investments — existing investors from Gruve who followed us into this new chapter. We are deploying five initial production sites, with the first one coming online in July 2026.

About the Role

We're hiring someone to own our engineering harness — the tooling and AI workflows that run across the full delivery pipeline, from how requirements get written to how releases go out.

The job is practical: work out where teams are losing time between ideas and shipped software, and build the infrastructure that fixes it. AI agents are the main lever. The measure of success is whether teams are actually shipping faster and with fewer manual handoffs, not whether the agents themselves are impressive.

Key Responsibilities
  • Identify bottlenecks across the full software delivery lifecycle — requirements refinement, design review, implementation, testing, and production rollout — and build AI-assisted harnesses that remove them; measure impact in concrete terms: cycle time, lead time, defect rate, and deployment frequency
  • Build harnesses that help teams move faster from idea to implementation, including AI-assisted requirement elaboration, automated design review, specification validation, and artefact generation
  • Design and maintain the agent environments that support day-to-day development work: code generation, automated refactoring, test authoring, and documentation; define the constraints, context files, and tool permissions that keep agents accurate and scoped
  • Build agent-assisted release workflows, automated pre-deployment checklists, rollback triggers, and progressive rollout controls that reduce manual effort and release risk
  • Define what good output looks like at each stage of the pipeline and enforce it automatically through evaluation datasets, grading rubrics, and CI/CD-integrated quality gates
  • Instrument every harness component with structured logging, tracing, and quality metrics; use that data to identify where agents underperform, where humans are compensating, and where the next productivity gain is
  • Implement the permission boundaries, guardrails, and human-in-the-loop checkpoints that keep agents operating within safe bounds without sacrificing delivery speed
  • Apply token optimisation, caching strategies, model tiering, and budget controls to ensure productivity gains are not eroded by runaway LLM costs
Basic Qualifications
  • 3+ years of professional software engineering experience, with a strong background in developer tooling, platform engineering, or DevOps
  • Demonstrated focus on software delivery performance — experience measuring and improving cycle time, deployment frequency, or release quality
  • Proficiency in at least one mainstream programming language such as Python or Go
  • Hands-on experience integrating with at least one major LLM API (OpenAI, Anthropic, or Google)
  • Solid understanding of production reliability patterns: retry logic, circuit breakers, rate limiting, timeout handling, and graceful degradation
  • Experience with CI/CD pipeline design and the full software delivery lifecycle
  • Experience with monitoring, logging, and observability of production systems
  • Strong written communication skills — able to document architectural decisions, constraints, and harness behaviour clearly for engineering teams
Preferred Qualifications
  • Experience with context engineering: token budgeting, retrieval-augmented generation (RAG), and context assembly across multi-step workflows
  • Familiarity with agent frameworks (LangGraph, CrewAI, Claude Agent SDK, or equivalent) and agent configuration patterns
  • Knowledge of LLM security practices, including prompt injection defence, PII handling, and output filtering
  • Familiarity with the Model Context Protocol (MCP) and experience integrating MCP servers into development toolchains
  • Background in multi-agent orchestration: coordination patterns, shared state management, and parallel execution
  • Familiarity with responsible AI practices and awareness of safety, compliance, or governance considerations in production AI systems
  • Contributions to open-source developer tooling, harness infrastructure, or AI engineering projects
Why AI Fabrik

At AI Fabrik, we hire for impact. We want those who challenge how inference infrastructure is built and who excel at delivering it in production. We are builders, architects, engineers, and researchers. We move fast, work with rigor, and care deeply about what runs in the real world.

We are committed to building a diverse and inclusive team. AI Fabrik is an equal opportunity employer. We welcome applicants from all backgrounds and thank all who apply; however, only those selected for an interview will be contacted.

Please note that this is an onsite position based out of AI Fabrik’s Redwood City, California office.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Harness Engineer - W2
AI Harness Engineer - W2

Stealth Talent Solutions • United States

On-site
USD 150,000 - 185,000
AI Harness Engineer
AI Harness Engineer

e&e IT Consulting Services, Inc. • Harrisburg

On-site
USD 83,000 - 152,000
AI Software Engineer, Agent Harness
AI Software Engineer, Agent Harness

EnCharge AI • United States

Hybrid
USD 42,000 - 73,000
AI Engineer
AI Engineer

Rad Hires • United States

Remote
USD 83,000 - 165,000
Technical Program Manager
Technical Program Manager

AI Fabrik • San Francisco (CA)

On-site
USD 150,000 - 200,000
AI Systems Engineer, Codex Agents
AI Systems Engineer, Codex Agents

OpenAI • San Francisco (CA)

On-site
USD 230,000 - 385,000
Staff Software Engineer - Data Platform
Staff Software Engineer - Data Platform

Mosaic.tech • Mountain View (CA)

Hybrid
USD 180,000 - 220,000
Competitive salary
Comprehensive healthcare benefits
FSA
+5
Founding Engineer - Applied AI
Founding Engineer - Applied AI

Product Pulse • San Francisco (CA)

On-site
USD 120,000 - 160,000
Unlimited PTO
Full health insurance
Free lunch and dinner
+2
Senior Technical Content Strategist
Senior Technical Content Strategist

Split Software • San Francisco (CA)

Hybrid
USD 150,000 - 165,000
Competitive salary
Comprehensive healthcare benefits
Flexible Spending Account (FSA)
+5
Staff Software Engineer - Data Platform
Staff Software Engineer - Data Platform

Harness • Mountain View (CA)

Hybrid
USD 180,000 - 220,000
Competitive salary
Comprehensive healthcare benefits
Flexible Spending Account (FSA)
+5