LLM Production Engineer - Cost, Rollouts & Observability

Systems Limited

Lahore

On-site

PKR 4,000,000 - 7,000,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Systems Limited is seeking an experienced MLOps Engineer to own production serving for LLM/GenAI workloads in Lahore, Pakistan. You will scale inference infra, manage caching, and optimize token costs while ensuring reliable operations for client projects.

You will implement observability pipelines, model routing, and rollout strategies, collaborating with GenAI engineers to maintain production readiness and cost governance for evolving workloads.

Qualifications

  • 4–6 years in platform/MLOps engineering with hands-on LLM/GenAI production experience.
  • Deep understanding of LLM inference economics including token costs, batching, caching, and routing.
  • Experience with LLM observability tooling (tracing, eval pipelines, prompt/version management).
  • Familiarity with multiple hosting platforms and cost/performance tradeoffs (Azure AI Foundry, AWS Bedrock, Vertex AI, self-hosted options).
  • Experience building canary/rollback strategies for probabilistic systems.
  • Calm under pressure during live incidents affecting client-facing systems.

Responsibilities

  • Own production serving and scaling for LLM/agentic workloads (inference infra, load balancing, caching)
  • Monitor and control inference cost — token usage, retry/loop cost, model routing decisions
  • Build observability for LLM-specific failure modes: hallucination rate, latency spikes, prompt drift
  • Manage model/version rollout strategy (canary releases, fallback models, A/B testing)
  • Own incident response for LLM/agent production issues
  • Partner with GenAI Engineers and Agentic AI Architects on production-readiness reviews
  • Explain token-cost dynamics to client finance/business stakeholders
  • Collaborate closely with GenAI Engineers without needing a hard line between build and run
  • Support the practice in setting cost governance policy for LLM workloads

Skills

MLOps engineering
LLM production
Cost optimization
Observability tooling
Model routing

Tools

Azure AI Foundry
AWS Bedrock
Vertex AI
vLLM
TGI

Job description

Systems Limited is seeking an experienced MLOps Engineer to own production serving for LLM/GenAI workloads in Lahore, Pakistan. You will scale inference infra, manage caching, and optimize token costs while ensuring reliable operations for client projects.

You will implement observability pipelines, model routing, and rollout strategies, collaborating with GenAI engineers to maintain production readiness and cost governance for evolving workloads.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GenAI Frontline Engineer: Build Production LLM Apps
GenAI Frontline Engineer: Build Production LLM Apps

Systems Limited • Lahore

On-site
PKR 2,500,000 - 5,500,000
Senior AI Platform Engineer - Scalable MLOps & Infra
Senior AI Platform Engineer - Scalable MLOps & Infra

Systems Limited • Lahore

On-site
PKR 3,000,000 - 5,400,000
Forward Deployed Engineer - LLMOps
Forward Deployed Engineer - LLMOps

Systems Limited • Lahore

On-site
PKR 4,000,000 - 7,000,000
Forward Deployed Engineer - LLMOps
Forward Deployed Engineer - LLMOps

Systems Limited • Islamabad

On-site
PKR 2,000,000 - 3,000,000
Senior AI/ML Engineer – Remote, LLMs & Production AI
Senior AI/ML Engineer – Remote, LLMs & Production AI

Jlab • Islamabad

On-site
PKR 2,000,000 - 4,200,000
Senior MLOps Engineer: Production, FinOps & Observability
Senior MLOps Engineer: Production, FinOps & Observability

Systems Limited • Lahore

On-site
PKR 3,000,000 - 5,400,000
Senior AI Engineer: LLMs, MLOps & Production AI
Senior AI Engineer: LLMs, MLOps & Production AI

FabTechSol • Sialkot

On-site
PKR 2,000,000 - 2,750,000
Senior MLOps Engineer: Scalable ML Deployment & Monitoring
Senior MLOps Engineer: Scalable ML Deployment & Monitoring

TalentHue- Careers • Lahore

On-site
PKR 1,800,000 - 3,000,000
Senior AI Engineer - Production AI & LLMs (Hybrid)
Senior AI Engineer - Production AI & LLMs (Hybrid)

NorthBay Solutions LLC • Islamabad

Hybrid
PKR 4,800,000 - 9,600,000
Senior AI Engineer (Hybrid) - Production ML & LLMs
Senior AI Engineer (Hybrid) - Production ML & LLMs

NorthBay Solutions • Islamabad

Hybrid
PKR 1,800,000 - 3,200,000