Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Zof AI, Inc. in San Francisco, CA is hiring a Site Reliability Engineer to run the infrastructure for large agent workloads and safe, cost-effective operations.
This is a full-time on-site role reporting to the Infrastructure team, focused on Kubernetes, containers, CI/CD, and isolating untrusted code within sandboxed environments. The ideal candidate will have production-scale experience, strong security and reliability judgment, and a proactive approach to observability, cost control, and
Competitive salary
Competitive salary
Plus meaningful equity
All roles
San Francisco, CA
Site Reliability Engineer
San Francisco, CAFull-timeMid to SeniorOn-site
Zof AI is hiring for this role in San Francisco, CA. This is a full-time opportunity for candidates who want to contribute directly to the development of ambitious AI products in a high-performance environment.
Must be able to run infrastructure for large agent workloads and use AI tools to automate operational work.
Zof AI is seeking a Site Reliability Engineer to run the infrastructure that lets fleets of sandboxed agents execute customer code safely and cheaply. This role owns the execution layer of our control plane: Kubernetes and container orchestration, CI/CD pipelines, hard isolation for untrusted code, observability, and the cost controls that keep large agent fleets affordable. If you have worked as a Site Reliability Engineer, Platform Engineer, Cloud Engineer, or Infrastructure Engineer, this is that discipline at Zof AI. The ideal candidate has operated production infrastructure at scale and treats security, reliability, and cost per agent run as constraints they personally own.
DevOpsKubernetesCI/CDCloud InfrastructureReliability
Benefits may depend on role and final offer terms.