A complete application in a minute — tailored resume and cover letter, ready to send.
Systems Limited is seeking an experienced MLOps engineer to own production serving and scaling for LLM/agentic workloads, including inference infrastructure, load balancing, and caching. You will drive cost awareness and implement observability for failure modes.
Collaborate with GenAI Engineers on rollout strategies, model versioning, canaries, and fallbacks. You will explain token-cost dynamics to clients and help set governance for LLM workloads.
Owns production operations for LLM and agentic workloads — serving, cost, and observability for a fundamentally less predictable class of system than classical ML.