AI/ML Ops Engineer

revolte.ai

Chennai District

On-site

INR 900,000 - 1,800,000

Full time

32 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Revolte is building AI for Software Engineering, an AI-native platform that helps engineering teams deliver software faster with quality and governance.

We are seeking an AI/ML Ops Engineer to design and operate scalable ML pipelines, deployment systems, and model infrastructure across cloud and Kubernetes environments, collaborating with AI, backend, and platform engineers. You will own CI/CD for AI models and ensure observability and cost efficiency in production.

Qualifications

  • 1–3 years of experience in MLOps, ML engineering, AI infrastructure, DevOps, SRE, or platform engineering.
  • Strong software engineering and automation skills.
  • Experience taking ML/AI workloads from experimentation to production.
  • Cloud, containers, and distributed systems understanding; deployment workflows.
  • Strong problem-solving and production troubleshooting abilities.
  • Ability to collaborate with AI, backend, platform, and product teams.
  • Ownership mindset; thrives in fast-moving environments.
  • Interest in AI/LLMs, AI agents, production AI systems.

Responsibilities

  • Build and maintain production-grade MLOps and LLMOps pipelines for AI apps.
  • Automate model training, evaluation, versioning, deployment, monitoring, and rollback workflows.
  • Create CI/CD pipelines for AI/ML models and services.
  • Deploy and operate AI workloads across cloud, containerised, and Kubernetes environments.
  • Develop reliable model-serving and inference infrastructure with focus on performance and cost.
  • Implement monitoring and observability for model performance and infra health.
  • Collaborate with AI engineers to optimise models and inference workloads.
  • Manage data, model, and feature pipelines across dev, staging and prod.
  • Troubleshoot production issues in AI apps, models, and data pipelines.
  • Establish best practices for reproducibility, security, reliability, and governance of AI systems.

Skills

Python
Go
Bash
CI/CD
Cloud platforms
Kubernetes
Prometheus
Grafana
MLflow
Terraform
Model monitoring
Troubleshooting
Collaboration

Tools

Docker
Kubernetes
GitHub Actions
GitLab CI
Jenkins
Kubeflow
SageMaker
Vertex AI
vLLM
KServe

Job description

Revolte is building AI for Software Engineering .

We are creating an AI-native platform that helps engineering teams execute the software delivery lifecycle with speed, quality, and control - from planning and implementation to testing, release, governance, and production feedback.

We're looking for an AI/ML Ops Engineer to build and operate the systems that take AI and ML capabilities from development to reliable production. You'll work closely with AI, backend, and platform engineers to build scalable pipelines, deployment systems, model infrastructure, and observability for AI-powered applications.

Responsibilities
  • Build and maintain production-grade MLOps and LLMOps pipelines for AI applications.
  • Automate model training, evaluation, versioning, deployment, monitoring, and rollback workflows.
  • Build CI/CD pipelines for AI/ML models, services, and supporting infrastructure.
  • Deploy and operate AI workloads across cloud, containerised, and Kubernetes environments.
  • Build reliable model-serving and inference infrastructure with a focus on performance, scalability, and cost.
  • Implement monitoring and observability for model performance, latency, reliability, and infrastructure health.
  • Work with AI Engineers to optimise models and inference workloads for production.
  • Manage data, model, and feature pipelines across development, staging, and production.
  • Troubleshoot production issues across AI applications, models, data pipelines, and infrastructure.
  • Establish best practices for reproducibility, security, reliability, and governance of AI systems.
  • Evaluate and adopt emerging AI infrastructure, MLOps, and LLMOps technologies.
What We Are Looking For
  • 1-3 years of experience in MLOps, ML Engineering, AI Infrastructure, DevOps, SRE, or Platform Engineering.
  • Strong software engineering and automation skills.
  • Experience taking ML/AI workloads from experimentation to production.
  • Strong understanding of cloud infrastructure, containers, distributed systems, and deployment workflows.
  • Strong problem-solving and production troubleshooting skills.
  • Ability to collaborate effectively with AI, backend, platform, and product teams.
  • Strong ownership mindset and willingness to work in a fast-moving environment.
  • Genuine interest in AI, LLMs, AI agents, and production AI systems.
Technical Skills
  • Strong proficiency in Python, Go, or Bash .
  • Hands-on experience with AWS, Azure, or Google Cloud .
  • Experience with Docker and Kubernetes .
  • Experience with MLflow, Kubeflow, SageMaker, Vertex AI , or similar MLOps platforms.
  • Experience building CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or similar.
  • Experience with Terraform, Pulumi, or other Infrastructure as Code tools.
  • Understanding of model training, evaluation, versioning, deployment, and monitoring.
  • Experience with model serving/inference frameworks such as vLLM, KServe, Triton, or BentoML is a plus.
  • Familiarity with LLMOps, RAG, embeddings, vector databases, and AI agent workloads.
  • Experience with Prometheus, Grafana, OpenTelemetry, or similar observability tools.
  • Strong understanding of Linux, networking, cloud security, IAM, and containerised systems.
  • Experience with databases, object storage, message queues, and distributed systems is desirable.
Why Join Revolte
  • Build the infrastructure behind the next generation of AI-powered software engineering.
  • Work at the intersection of AI, ML, cloud infrastructure, and distributed systems .
  • Solve challenging problems around model reliability, inference, scalability, observability, and cost.
  • Work closely with AI, backend, and platform engineers.
  • Take ownership of systems from architecture and automation through production.
  • Work with modern AI and cloud technologies and continuously expand your technical expertise.
  • Join a high-ownership AI-first company where your work directly shapes the engineering foundation.
  • Grow with Revolte as we build technology that changes how software engineering is done.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Member of Technical Staff: ML Systems and Infrastructure
Senior Member of Technical Staff: ML Systems and Infrastructure

DevRev • Bengaluru

On-site
INR 1,500,000 - 2,300,000
AI-MLOps SRE Lead Engineer
AI-MLOps SRE Lead Engineer

Randstad • Hyderabad

Hybrid
INR 2,400,000 - 3,600,000
Senior DevOps Engineer (Kubernetes & AI Infra)
Senior DevOps Engineer (Kubernetes & AI Infra)

Navikenz • Bengaluru

Hybrid
INR 1,500,000 - 2,500,000
AI-ML Engineer
AI-ML Engineer

KanthamAi • Mumbai

On-site
INR 2,000,000 - 3,000,000
Software DevOps Engineer - Gen Software and ML software
Software DevOps Engineer - Gen Software and ML software

The ePlane Company • Chennai District

On-site
INR 1,400,000 - 2,100,000
DevOps Specialist Engineer SRE, Cloud & Applied AI
DevOps Specialist Engineer SRE, Cloud & Applied AI

Clarus Advisers • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Software Development Engineer
Software Development Engineer

Ciroos • Gurugram District

On-site
AI / MLOps PLATFORM ENGINEER
AI / MLOps PLATFORM ENGINEER

Vrinda International • Bengaluru

On-site
INR 3,500,000 - 5,200,000
Lead ML Ops
Lead ML Ops

Evergent • Hyderabad

On-site
INR 1,500,000 - 2,000,000
AI Engineer (Python, GenAI/LLMs + ML Fundamentals)
AI Engineer (Python, GenAI/LLMs + ML Fundamentals)

Solutions By Text • Bengaluru

On-site
INR 2,500,000 - 4,000,000