Stand out for this role — generate a tailored resume and cover letter in about a minute.
Systems Limited is seeking an experienced MLOps Engineer to own production serving for LLM/GenAI workloads in Lahore, Pakistan. You will scale inference infra, manage caching, and optimize token costs while ensuring reliable operations for client projects.
You will implement observability pipelines, model routing, and rollout strategies, collaborating with GenAI engineers to maintain production readiness and cost governance for evolving workloads.
ABOUT
Owns production operations for LLM and agentic workloads — serving, cost, and observability for a fundamentally less predictable class of system than classical ML.