AI Ops Engineer

OneMagnify

Chennai District

On-site

INR 1,500,000 - 2,200,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Medical Insurance
PF and Gratuity
Paid holidays

Job summary

OneMagnify seeks an AI Operations Engineer to design, automate, and secure scalable AI infrastructure in cloud environments. You will drive MLOps/LLMOps pipelines, CI/CD automation, and observability to ensure safe, reliable production AI deployments.

Ideal candidates bring 3–5 years in cloud engineering or SRE, strong Terraform and Kubernetes skills, and hands-on AI model deployment experience within a dynamic digital agency setting.

Qualifications

  • 3–5 years of cloud engineering, DevOps, or SRE experience.
  • Hands-on AI/ML production deployment experience.
  • Experience with CI/CD for AI workloads and containerization.
  • Proficiency with Terraform and cloud‑native tooling.

Responsibilities

  • Design and operate scalable AI infrastructure on the cloud.
  • Build and maintain CI/CD pipelines for AI applications.
  • Implement observability, security, and governance.
  • Collaborate with AI engineers and data scientists.

Skills

Cloud Platform Engineering
DevOps / SRE
CI/CD for AI workloads
Docker/Kubernetes
Terraform
GCP (Google Cloud)
Python scripting

Education

Master's or Bachelor's in CS/related

Tools

Terraform
Cloud Run
GKE
Cloud Build
BigQuery

Job description

OneMagnify is an AI native, platform-enabled B2B digital agency operating at the intersection of data, technology, and creativity. We help complex organizations drive measurable business outcomes by building smarter customer experiences and delivering highly integrated solutions across digital, media, and technology. By combining deep industry expertise with advanced analytics and artificial intelligence, we enable our clients to make better decisions, move faster, and compete more effectively in dynamic markets.

We are seeking a motivated and talented AI Operations Engineer with strong cloud infrastructure and AI operations expertise to build, automate, and secure the platform that powers advanced artificial intelligence solutions.

The Impact You’ll Have:

In this role, you will support quantitative analytics and AI engineering teams by designing, automating, and operating end-to-end cloud and AI infrastructure. Responsibilities include managing CI/CD pipelines, containerized microservices, observability platforms, and governance controls to ensure AI applications and models run safely, reliably, and at scale in production environments.

What you’ll do:
  • Cloud Platform Engineering: Architect and operate highly available, multi-service AI infrastructure on the cloud, managing the full lifecycle of compute, storage, networking, and security for AI workloads.
  • AI-Ops / LLMOps Pipelines: Establish robust MLOps and LLMOps pipelines covering the end-to-end lifecycle of Generative AI tools — including model deployment, version control, prompt and artifact tracking, automated evaluation, and continuous monitoring for performance, drift, and hallucination mitigation.
  • CI/CD Automation: Design and maintain automated build, test, and deployment pipelines for full-stack AI applications, ensuring seamless and secure continuous integration and delivery across front-end, back-end, and AI components.
  • Infrastructure as Code: Manage cloud infrastructure using Infrastructure as Code (e.g., Terraform, Cloud Build, Kubernetes manifests) to deliver reproducible, auditable, and scalable environments.
  • Containerization & Orchestration: Build and operate containerized microservices (Docker/Kubernetes), managing scaling, rolling deployments, resource optimization, and service resilience for AI workloads.
  • Observability & Reliability: Implement comprehensive monitoring, logging, tracing, alerting, and SRE practices to ensure platform reliability, availability, and performance of AI applications in production.
  • Security & Governance: Embed security across the platform — managing user identities and controlling access rights, safeguarding sensitive credentials, network policies, and data protection — ensuring all AI workloads meet Ford’s strict data privacy, security, and compliance standards.
  • Developer Enablement: Work closely with AI engineers, software engineers, and analytical modelers to provide self-service tooling, environments, and automated workflows that remove friction from development to production.
What you’ll need:
  • Education: Master's or Bachelor's degree in Computer Science, Software Engineering, Cloud Computing, Data Engineering, or a related technical discipline.
  • DevOps/Cloud Engineering Experience: 3–5 years of overall experience in cloud engineering, DevOps, or SRE, with at least 1–2 years of dedicated, hands‑on experience deploying and operating AI, ML, and Generative AI applications in production.
  • AI/MLOps Engineering: Proven track record of implementing CI/CD for AI workloads, containerization (Docker/Kubernetes), and cloud infrastructure management (Terraform, Cloud Build, GKE).
  • Automation & Reliability: Demonstrated experience with infrastructure automation, incident response, and building observable, self‑healing production systems.
  • Cloud Platform: Extensive hands‑on experience with a leading cloud provider (e.g., Google Cloud Platform), including Cloud Run, Cloud Build, GKE, GCS, BigQuery, IAM, VPC networking, and Secret Manager.
  • DevOps / SRE Practices: Strong proficiency in CI/CD tooling, GitOps, containerization (Docker), orchestration (Kubernetes), Infrastructure as Code (Terraform), and cloud‑native monitoring and logging.
  • AI-Ops / LLMOps: Proficiency in tools and platforms for model deployment, prompt/model versioning, evaluation, tracing, and monitoring of LLM and GenAI outputs.
  • Scripting & Automation: Hands‑on experience with a programming/scripting language such as Python, or Bash for automating infrastructure and operational tasks, and for building tooling that serves collaborators.
  • Software Engineering Fundamentals: Solid understanding of full‑stack application architecture and modern deployment patterns for AI tools, with the ability to integrate front‑end, back‑end, and AI services reliably.
  • Security & Compliance: Familiarity with cloud security best practices, identity and access management, secrets management, and compliance standards in regulated environments.
  • Analytics Workflow Understanding: Awareness of the typical workflows of data scientists and modelers (data wrangling, feature engineering, model validation) so you can build reliable platforms and pipelines that serve them.
Future-Ready Skills (Nice to Have):
  • Experience in integrated marketing, digital agency, marketing services, or consulting environments preferred.
  • Previous exposure to the Banking, Financial Services, or Credit Analytics industries.
  • Experience with Machine Learning engineering and model serving frameworks, or relevant cloud certifications (e.g., Google Cloud Professional DevOps Engineer / Cloud Architect).
Benefits
  • Benefits We offer a comprehensive benefits package including Medical Insurance, PF, Gratuity, paid holidays, and more.

We are an equal opportunity employer

We believe that Innovative ideas and solutions start with unique perspectives. That’s why we’re committed to providing every employee a workplace that’s free of discrimination and intolerance.

We’re proud to be an equal opportunity employer and actively search for like-minded people to join our team.

Whether it’s awareness, advocacy, engagement, or efficacy, we move brands forward with work that delivers results. Through meaningful analytics, engaging communications and innovative technology solutions, we help clients tackle their most ambitious projects and overcome their biggest challenges.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Manager, Artificial Intelligence
Manager, Artificial Intelligence

OneMagnify • Chennai District

On-site
INR 3,500,000 - 7,000,000
Medical, dental, and vision coverage
401(k) retirement plan
Paid holidays
+3
AI Engineer, Smart Claims
AI Engineer, Smart Claims

Amplify Health • Dadri

On-site
INR 11,472,000 - 17,208,000
Senior Engineer, Enterprise BI & Analytics
Senior Engineer, Enterprise BI & Analytics

OneMagnify • Chennai District

On-site
INR 1,500,000 - 2,800,000
Medical Insurance
PF
Gratuity
+1
Mid-Senior Data Scientist
Mid-Senior Data Scientist

OneMagnify • Chennai District

On-site
INR 1,500,000 - 2,500,000
Medical Insurance
PF
Gratuity
+1
Senior AI/ML Engineer (MLOps and LLMOps platforms )
Senior AI/ML Engineer (MLOps and LLMOps platforms )

Optum • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Generative AI Architect
Generative AI Architect

Confidential • Bengaluru

Hybrid
INR 4,000,000 - 7,000,000
Hybrid work options
Learning budget
GPU infrastructure access
+1
Financial Risk Credit Modeler
Financial Risk Credit Modeler

OneMagnify • Chennai District

On-site
INR 1,200,000 - 1,800,000
Medical Insurance
Provident Fund (PF)
Gratuity
+1
AI/ML Engineer
AI/ML Engineer

Jash Data Sciences Pvt. Ltd. • Pune District

On-site
INR 800,000 - 1,200,000
Competitive salary
Learning opportunities
Exposure to latest AI technologies
AIML Engineer Digital And AI
AIML Engineer Digital And AI

IQ-EQ • Hyderabad

On-site
INR 1,400,000 - 2,800,000
Health insurance
Paid time off
Professional development
+1
Senior AI-ML Platform Engineer / Tech Lead
Senior AI-ML Platform Engineer / Tech Lead

Merck Group • Bengaluru

On-site
INR 4,000,000 - 7,000,000