Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Zof AI seeks an experienced MLOps Engineer to own the platform for AI systems, spanning model serving, scaling, and reliability. You will push the platform that enables a small team to run large-scale AI workloads efficiently in production, with a focus on GPU throughput, observability, and cost control.
As a senior on-site engineer in San Francisco, you will design and operate the backend infrastructure, partner with software engineers, and drive incident practices, rollback automation, and
Zof AI is seeking a MLOps Engineer to own the platform our AI systems run on. This is a consolidated platform role spanning what the market posts as AI Infrastructure, MLOps, LLMOps, and agent platform engineering: model serving and scaling, GPU and compute efficiency, and the provisioning, observability, and reliability infrastructure that keeps production agents stable. The ideal candidate has operated real AI workloads in production and builds infrastructure that lets a small team run systems well above its weight.
Engineering · Senior · Full-time · On-site · San Francisco, CA
Experience operating production AI or ML systems at scale is required