ML Ops Engineer Chanakya

Sarvam

Bengaluru

On-site

INR 3,000,000 - 6,500,000

Full time

5 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Sarvam is building the bedrock of Sovereign AI for India, and seeks an MLOps Engineer to own model lifecycle across defence and strategic deployments. You will ensure the system is always on, accurate, and auditable, working with STR Deployment Engineers and Product teams to deploy new products.

You will design serving infra for on‑prem and cloud, build CI/CD for model updates, monitor latency and drift, and create evaluation dashboards. Strong ML Ops and Python skills are essential.

Qualifications

  • 3–5 years in ML engineering or MLOps with at least one production ML system in operation.
  • Deep expertise in model serving and deployment of ML models.
  • Experience fine-tuning models in constrained environments (on‑prem/air‑gap).
  • Containerisation with Docker/Kubernetes for edge deployments.
  • Experience with monitoring tools like Prometheus/Grafana.
  • Proficiency in Python and ML evaluation workflows.
  • Hands-on CI/CD tooling for ML (GitHub Actions, ArgoCD, DVC, etc.).

Responsibilities

  • Design and operate model serving infrastructure across on‑prem and cloud.
  • Build and maintain CI/CD pipelines for model updates and gated deployments.
  • Monitor production model performance and surface issues proactively.
  • Create evaluation infra and runbooks for field deployments.
  • Collaborate with Data Scientists on eval pipelines and infrastructure.
  • Manage containerised serving in constrained and edge environments.
  • Own incident response for model‑layer failures across deployments.

Skills

ML engineering
MLOps
Model serving
Docker
Kubernetes
Prometheus
Python
CI/CD for ML
LLM deployment
Edge/on‑prem deployment

Tools

Triton Inference Server
vLLM
GGUF/AWQ/GPTQ
ArgoCD
DVC

Job description

Job Description:

About Sarvam

Sarvam is building the bedrock of Sovereign AI for India. The company is developing Indias full-stack sovereign AI platform, building across research, models, infrastructure and applications with a singular focus on making AI genuinely work for India. Sarvam works with leading enterprises and public institutions and is backed by Lightspeed, Peak XV, and Khosla Ventures. Sarvam partners with Indias leading brands, including Tata Capital, SBI Life, CRED, IDFC, and LIC.

About the Role

The MLOps Engineer owns the model lifecycle across all defence and strategic sector deployments — from serving infrastructure and monitoring to evaluation pipelines and environment management. You ensure the system is always on, always accurate, and always auditable.

You will work across both layers: supporting Strategic Deployment Engineers in the field, and owning the model deployment infrastructure for new products being built by the product engineering team. The standards here are uncompromising — a model failure is not a UX problem, it is an operational risk.

What Youll Do
  • Design and operate model serving infrastructure across on-prem and cloud deployments

  • Build and maintain CI/CD pipelines for model updates, rollbacks, and evaluation-gated deployments

  • Monitor model performance in production — latency, accuracy drift, throughput, failure modes — and build systems that surface issues before clients do

  • Build evaluation infrastructure: harnesses, A/B testing, and model comparison tooling for field and lab use

  • Manage containerised model serving in constrained, air-gapped, and edge environments

  • Collaborate with Data Scientists on eval pipelines; own the infrastructure layer underneath

  • Create runbooks and operational playbooks that Strategic Deployment Engineers can use in the field

  • Own incident response for model-layer failures across all active deployments

What Were Looking For
  • 3–5 years in ML engineering or MLOps with at least one production LLM or ML system in continuous operation

  • Deep expertise in model serving: vLLM, TGI, Triton Inference Server, or equivalent; experience with quantised model formats (GGUF, AWQ, GPTQ)

  • Experience fine-tuning and adapting models in constrained, on-prem, or air-gapped environments, including managing data pipelines and compute limitations specific to the environment

  • Containerisation experience with Docker, Kubernetes, or lightweight alternatives (K3s, K0s) for constrained and edge environments; familiarity with deploying across heterogeneous hardware and infrastructure configurations

  • Monitoring and observability using Prometheus, Grafana, or equivalent; ability to build custom eval dashboards

  • Python fluency; familiarity with fine-tuning workflows and model evaluation frameworks

  • Hands-on experience with CI/CD tooling for ML pipelines: GitHub Actions, ArgoCD, DVC, or similar

Signals We Look For
  • Youve kept a production ML system running under load — and debugged it when it broke

  • You dont wait for things to fail; you build systems that tell you when theyre about to

  • You write documentation that actually gets used, by people who arent you

Who You Are
  • You treat uptime and correctness as equally non-negotiable

  • You understand that operational reliability is a form of trust-building

  • Youre as comfortable optimising inference throughput as you are writing a field runbook for a deployment engineer

  • You take ownership of model and stack health across every active deployment — not just the ones you set up

Why Sarvam?

Sarvam is a fast-moving, high talent-density team building full-stack AI for India, working on problems that push the frontiers of AI with real population-scale impact.

  • Work alongside researchers, engineers, builders, and business leaders who move fast and hold each other to a very high bar

  • High ownership and high impact, from day one

  • Everything we do is AI-first, from the way we build and ship to the way we think about problems

  • You can work on problems that could change how an entire country learns, works, and communicates

If you want to work on problems at the frontier of AI in India, Sarvam is the place to be.

Requirements:

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Ops Engineer, Chanakya
ML Ops Engineer, Chanakya

Sarvam • Delhi

On-site
INR 1,200,000 - 2,000,000
High ownership
High impact from day one
AI-first approach
ML Ops Engineer, Chanakya
ML Ops Engineer, Chanakya

Sarvam • Delhi

On-site
INR 1,500,000 - 2,000,000
High ownership in projects
Collaborative team environment
Opportunity to work on impactful AI solutions
Strategic Deployment Engineer Chanakya
Strategic Deployment Engineer Chanakya

Sarvam • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Embedded Data Scientist Chanakya
Embedded Data Scientist Chanakya

Sarvam • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Embedded Data Scientist, Chanakya
Embedded Data Scientist, Chanakya

Sarvam • Delhi

On-site
INR 1,000,000 - 2,000,000
GTM & Strategy, Model API
GTM & Strategy, Model API

Sarvam • Bengaluru

On-site
INR 800,000 - 1,100,000
Strategic Deployment Engineer, Chanakya
Strategic Deployment Engineer, Chanakya

Sarvam • Delhi

On-site
INR 1,200,000 - 1,800,000
Platform Engineer - AI Infrastructure
Platform Engineer - AI Infrastructure

Sarvam • Chennai District

On-site
INR 4,000,000 - 7,000,000
Product Operations & Analytics Intern
Product Operations & Analytics Intern

Sarvam AI • Bengaluru

On-site
INR 600,000 - 1,000,000
Embedded Data Scientist, Chanakya
Embedded Data Scientist, Chanakya

Sarvam AI • Delhi

On-site
INR 1,500,000 - 3,000,000