Get more replies from employers
Send a job-specific resume in minutes.
Systems Limited is seeking an experienced MLOps engineer to own production serving and scaling for LLM/agentic workloads, focusing on inference infra, load balancing, and caching.
You will monitor token costs, build observability, manage canary rollouts, and collaborate with GenAI engineers on production readiness, cost governance, and incident response.
ABOUT
Owns production operations for LLM and agentic workloads — serving, cost, and observability for a fundamentally less predictable class of system than classical ML.