AI/ML Platform Engineer IV

The Standard India

Bengaluru

Hybrid

INR 4,200,000 - 7,800,000

Full time

36 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Annual incentive bonus
Paid time off
Learning and leadership enablement
Career advancement opportunities

Job summary

The Standard India is seeking an AI/ML Platform Engineer IV to stand up the multi-provider model gateway, model registry, and LLMOps lifecycle for platform-owned models on our Data & AI Platform. You will enable access, governance, and cost-management across models from Anthropic, OpenAI, and open-source providers.

Hybrid role with on-site Bengaluru base; you will operate GPU inference infrastructure and governance tooling, driving scalable MLOps across the enterprise.

Qualifications

  • 10+ years of software or ML engineering experience.
  • Expert Python and strong CI/CD and containers.
  • Hands-on with Anthropic and OpenAI APIs.
  • GPU inference serving with NVIDIA NIM or open-source like vLLM.
  • Experience on Kubernetes GPU scheduling and tuning.
  • MLflow and model registry experience.
  • LLM evaluation with Langfuse/Arize.
  • Guardrails or safety controls for generative systems.
  • Feature store experience and drift monitoring.

Responsibilities

  • Design and operate the multi-provider model gateway across provider models.
  • Stand up the MCP/tool registry and AI Gateway with permissions.
  • Set up GPU inference infra on NVIDIA Grace Blackwell environments.
  • Manage GPU environments: Kubernetes, CUDA stack, tuning.
  • Run the model registry with versioning and risk status.
  • Build LLMOps: evaluation, cost monitoring, and versioning.
  • Establish LLM observability with Langfuse/Arize.
  • Enforce guardrails at runtime for governance.
  • Develop MLOps automation for platform models.
  • Operate serving/training environments on shared substrate.
  • Provide evidence for validation and mentor juniors.

Skills

Python
CI/CD
Containers
GPU Inference
Kubernetes
Model Registry
Langfuse/Arize
Guardrails
MLflow

Education

B.Tech / B.E.
Master's degree preferred

Tools

NVIDIA NIM
vLLM
MLflow
Unity Catalog
Databricks ML
Neo4j

Job description

The next part of your journey is right around the corner — with Standard India A genuine desire to make a difference in the lives of others is the foundation for everything we do. With a customer-first mindset and an intentional focus on building strong teams across the nation, we’ve been able to uphold our legacy of financial stability while investing in new, innovative technologies that support the needs of our customers. Our high-performance culture, focused on operational excellence thrives thanks to remarkable people united by compassion and a customer-first commitment. Are you ready to make a difference?

AI/ML Platform Engineer IV
Job Summary

The Standard is seeking an AI/ML Platform Engineer IV to stand up the multi-provider model gateway, the model registry and the LLMOps/MLOps lifecycle for platform-owned models on our Data & AI Platform. In this role, you will define how models are accessed, governed, evaluated and cost-managed across the enterprise — routing across Anthropic, OpenAI and open-source models, operating the model registry with risk status, and building evaluation and observability into every model change. You will also set up and operate our LLMOps environment including AI gateway, GPU inference infrastructure: deploying NVIDIA NIM microservices and open-source serving frameworks on NVIDIA Grace Blackwell (GB300) environments to host open-source models on-premises. Your initial assignment is the MCP Gateway task force. Actuarial pricing and reserving models remain governed by Actuarial's own standards; this role covers platform-owned models only. This role sits on The Standard's Data & AI Platform team, which owns the enterprise's shared data and AI foundation — governed once and consumed by many — built on Azure Databricks, Neo4j, Confluent Kafka, Kong, and Anthropic and OpenAI models behind a governed gateway.

Ability to work on-site in Bengaluru, India is a requirement of the role.

Mode of Working

This role follows a hybrid work model, with a primary base in Bengaluru. While on-site presence is essential for key engagements, flexibility is offered for remote work based on business needs and team alignment.

Key Responsibilities
  • Design and operate the multi-provider model gateway — routing, failover, rate limiting, caching and usage attribution (LiteLLM-style patterns) — across Anthropic, OpenAI and self-hosted open-source models.
  • Stand up and operate the MCP/tool registry, AI Gateway — a governed catalog of MCP servers and tools with scoped permissions — and A2A coordination surfaces.
  • Set up and operate GPU inference infrastructure on NVIDIA Grace Blackwell (GB300) environments: deploy NVIDIA NIM microservices and open-source serving frameworks such as vLLM, SGLang and NVIDIA Dynamo for distributed serving.
  • Manage GPU environments end to end: Kubernetes GPU scheduling (GPU Operator, device plugins, MIG/time-slicing), driver and CUDA stack lifecycle, capacity planning, and throughput/latency/utilization tuning.
  • Run the model registry with versioning, lineage, approval workflow and model-risk status for platform-owned models.
  • Build LLMOps capabilities: prompt and context versioning, evaluation harnesses, regression testing and token-cost monitoring.
  • Stand up LLM observability and evaluation with tools such as Langfuse, Arize Phoenix and RAGAS.
  • Integrate guardrails enforcement so governance policy executes at runtime.
  • Build MLOps automation for platform-owned predictive models: training pipelines, feature store integration, drift detection and retraining triggers on Databricks.
  • Operate model serving and training/evaluation environments on the shared substrate, benchmarking self-hosted models against provider APIs for cost and performance.
  • Produce reproducible evidence for independent model validation and mentor junior AI/ML Platform Engineers.
Skills And Background You’ll Need
Education

B.Tech / B.E. in Software Engineering, Computer Science, Data Science or Information Systems, or related field is required. Master’s degree is preferred.

Experience
  • Typically requires 10+ years of software or ML engineering experience, including significant production ML/AI platform work.
  • Expert Python and strong engineering fundamentals: testing, CI/CD and containers.
  • Hands-on experience with Anthropic and OpenAI APIs and multi-provider gateway/routing patterns (LiteLLM-style).
  • Hands-on GPU inference serving experience with NVIDIA NIM and/or open-source frameworks such as vLLM, SGLang.
  • Experience operating GPU infrastructure on Kubernetes: GPU Operator, device plugins, MIG or time-slicing, and utilization/throughput tuning.
  • MLflow and model registry experience, plus model serving and monitoring in production.
  • LLM evaluation and observability experience with Langfuse, Arize Phoenix, RAGAS or equivalent.
  • Experience implementing guardrails or safety controls for generative systems.
  • Feature store experience and drift monitoring for predictive models.
Preferred Experience
  • Experience deploying on Grace Blackwell-class systems (GB200/GB300 NVL72), NVIDIA AI Enterprise, NVIDIA Dynamo for disaggregated serving, or low-precision formats such as NVFP4/FP8.
  • Databricks ML stack depth (Model Serving, Mosaic AI tooling) and Unity Catalog model governance.
  • Retrieval engineering with vector stores or GraphRAG (Neo4j).
  • Exposure to NIST AI Risk Management Framework or AI TRiSM in a regulated industry.
  • Fine-tuning open-source models; GPU FinOps and inference cost optimization.
  • Good to have Databricks Machine Learning Professional or Azure AI Engineer certification.
Key Behaviors Of a Successful Candidate
  • Adaptability — Seeks information to understand the rationale for and importance of the change.
  • Customer Focus — Displays an interest in the customer by trying to understand their concerns and issues; draws on customer insight to help others best meet current and future customer needs.
  • Improvement Mindset — Demonstrates curiosity by asking questions regarding current approaches/methods and identifying potential changes.
Why Join Standard India?
  • A rich benefits package that supports your health, well-being, and financial goals
  • An annual incentive bonus plan tied to both individual and organizational performance
  • Paid time off including earned leave, sick/casual leave, and public holidays, in accordance with the company’s leave policy
  • A supportive, responsive management approach that empowers you to grow
  • Opportunities for career advancement through learning, coaching, and leadership enablement
  • A culture where your voice matters - your ideas shape our future, and we win as one
  • Purpose-driven work that makes a difference for our customers, communities, and each other

Eligibility to participate in an incentive program is subject to the rules governing the program and plan. Any award depends on a variety of factors including individual and organizational performance.

StanCorp Global Services India Private Limited is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, religion, colour, sex, gender identity, sexual orientation, age, disability, caste, HIV status or any other characteristic protected by applicable laws of India.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Model Validation Lead
AI Model Validation Lead

The Standard India • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Rich benefits package
Annual incentive bonus plan
Paid time off including sick/casual leave
+1
Software Engineer III — AI Gateway & Integration
Software Engineer III — AI Gateway & Integration

The Standard India • Bengaluru

Hybrid
INR 2,800,000 - 4,200,000
DevOps Platform Engineer IV
DevOps Platform Engineer IV

The Standard India • Bengaluru

Hybrid
INR 3,600,000 - 6,000,000
Benefits package
Annual incentive bonus
Paid time off
Data Platform Engineer III
Data Platform Engineer III

The Standard India • Bengaluru

Hybrid
INR 4,500,000 - 6,500,000
Manager of Engineering - Data Platform
Manager of Engineering - Data Platform

The Standard India • Bengaluru

Hybrid
INR 3,500,000 - 7,000,000
Software Engineer III — Agent Integration & SDK
Software Engineer III — Agent Integration & SDK

The Standard India • Bengaluru

Hybrid
INR 2,000,000 - 3,200,000
Annual incentive bonus
Paid time off
Health and wellness programs
+2
Software Engineer III - BenTech Integrations
Software Engineer III - BenTech Integrations

The Standard India • Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Health benefits
Annual bonus
Paid time off
+1
Principal Software Engineer V
Principal Software Engineer V

The Standard India • Bengaluru

Hybrid
INR 4,200,000 - 6,600,000
Annual incentive bonus
Paid time off
Career progression
+1
Software Engineer III - Data Services
Software Engineer III - Data Services

The Standard India • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Annual incentive bonus
Paid time off
Learning & development programs
Scrum Master III - Data & AI Platform
Scrum Master III - Data & AI Platform

The Standard India • Bengaluru

Hybrid
INR 2,500,000 - 3,500,000
Annual incentive bonus
Paid time off including earned leave &
Learning and leadership enablement
+2