Senior/Lead DevOps + MLOps Engineer

Luxoft Germany

Karnataka

On-site

INR 1,200,000 - 2,000,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Luxoft Germany seeks a Senior DevOps/Platform Engineer to evolve and operate a core cloud platform on AWS, enabling Enterprise-grade Agentic AI workloads while maintaining stability and security. You will design, implement, and optimize infrastructure, IaC, and CI/CD processes across traditional apps and AI components.

You will collaborate with AI/ML, security, and application teams to ensure scalable, observable, and compliant delivery, shaping platform standards and developer experience.

Qualifications

  • 8+ years of experience in DevOps, Platform Engineering, or SRE roles with deep AWS expertise (EC2, EKS/ECS, VPC, IAM, S3, RDS/DynamoDB, CloudWatch, Lambda).
  • Strong Linux administration and privilege escalation knowledge (sudo, etc.).
  • Hands-on Terraform experience: modular architectures, state management, multi-account setups.
  • Scripting in Python and Bash for automation.
  • Experience with container technologies: Docker, Kubernetes, and image lifecycle management.

Responsibilities

  • Design, build, and operate a shared AWS cloud platform supporting traditional services and agentic AI workloads.
  • Own and evolve IaC using Terraform across environments.
  • Extend infra for Enterprise-grade Agentic AI systems with secure data access and governance.
  • Build platform abstractions, templates, and tooling for safe consumption of agentic capabilities.
  • Develop and maintain CI/CD pipelines for traditional & AI-driven components.
  • Implement platform observability (metrics, logs, traces) across services and agent workloads.
  • Enforce security and compliance (IAM, secrets, encryption, least privilege).
  • Collaborate with application, AI/ML, and security teams to improve developer experience and reliability.
  • Lead architecture discussions and platform standards.

Skills

AWS expertise
Terraform
IaC (Terraform)
Docker
Kubernetes
CI/CD pipelines
Linux administration
Networking & security concepts
Scripting (Python/Bash)
IAM & security best practices
Agentic AI platform experience
Observability & telemetry

Tools

TF modules design
Git
Artifactory/Nexus
EC2/EKS/ECS
CloudWatch

Job description

Project description

We are seeking a Senior DevOps / Platform Engineer with deep AWS expertise to evolve and operate our core cloud platform, while enabling Enterprise-grade Agentic AI capabilities on top of established infrastructure. This role is focused on platform stability, scalability, and developer enablement, ensuring that traditional services and emerging agentic systems coexist securely and reliably. You will play a key role in transforming our platform into a foundation that supports AI-augmented and agentic SDLC workflows, without compromising operational excellence.

Responsibilities
  • Design, build, and operate a shared AWS cloud platform that supports both traditional services and agentic AI workloads.
  • Own and evolve Infrastructure as Code (IaC) using Terraform, ensuring consistency, security, and repeatability across environments.
  • Extend existing infrastructure to support Enterprise-grade Agentic AI systems, including:o Execution runtimes for autonomous and semi-autonomous agentso Secure access to data, services, and APIso Platform-level guardrails for safety, governance, and cost control
  • Build platform abstractions, templates, and tooling that enable teams to safely consume agentic capabilities.
  • Support and integrate agentic SDLC tools and processes, including AI-assisted development, testing, and release automation.
  • Develop and maintain CI/CD pipelines for both traditional applications and AI-driven components.
  • Implement platform-level observability (metrics, logs, traces) across services and agent workloads.
  • Enforce security and compliance best practices (IAM, secrets management, encryption, least privilege).
  • Collaborate closely with application, AI/ML, and security teams to improve developer experience and platform reliability.
  • Act as a technical leader in architecture discussions, platform standards, and operational readiness.
SKILLS
Must have
  • 8+ years of experience in DevOps, Platform Engineering, or Site Reliability Engineering roles. Deep, hands-on AWS expertise, including:EC2, EKS/ECS, VPC, IAM, S3, RDS/DynamoDB, CloudWatch, Lambda
  • Strong understanding of VPC design, Security Groups, Route Tables, NACLs, and hybrid networking concepts.
  • Ability to troubleshoot network connectivity issues, ingress/egress rules, and cloud security configurations.
  • Strong production experience with Terraform, including:-Designing modular Terraform architectures-Managing state, environments, and multi-account setups
  • Strong Linux administration skills:-User and permissions management-Process and service management-Package management-System troubleshooting
  • Solid understanding of Linux commands and privilege escalation concepts (sudo, sudo su, sudoers).
  • Hands-on scripting experience:-Python-Bash/Shell scripting
  • Ability to automate operational tasks and troubleshooting workflows.
  • Experience with container technologies:-Docker-Kubernetes (EKS preferred)-Container image lifecycle management and versioning/tagging strategies
  • Experience with artifact repositories:-JFrog Artifactory, Nexus or equivalent-Understanding of image tagging, versioning, and promotion strategies
  • Proven experience rolling out Enterprise-grade Agentic AI infrastructure on top of existing platforms, including:-Supporting agent execution within established networking, security, and compliance boundaries-Enabling scalability, observability, and governance for agent behavior
  • Hands-on experience supporting agentic SDLC tools and processes:-AI-assisted coding, testing, and deployment workflows-Agent-based automation within CI/CD and operational processes
  • Experience with AWS AI services such as: Amazon Bedrock, SageMaker, Related AI/ML platform services.
  • Solid understanding of Linux, networking, and cloud security fundamentals.
  • Familiarity with MLOps or AI platform components Model serving, Vector databases, Feature stores
Nice to have
  • Experience designing internal developer platforms (IDPs).
  • Knowledge of policy-as-code and governance frameworks (OPA, SCPs, tagging strategies).
  • AWS certifications (Solutions Architect, DevOps Engineer).
  • Experience operating platforms at enterprise scale or in regulated environments.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior/Lead DevOps + MLOps Engineer
Senior/Lead DevOps + MLOps Engineer

Luxoft Germany • Hyderabad

On-site
INR 11,572,000 - 17,358,000
Senior DevOps + MLOps Engineer
Senior DevOps + MLOps Engineer

Luxoft • Bengaluru

On-site
INR 2,600,000 - 4,200,000
Senior DevOps + MLOps Engineer
Senior DevOps + MLOps Engineer

Luxoft • Gurugram District

On-site
INR 1,800,000 - 2,400,000
Senior DevOps + MLOps Engineer
Senior DevOps + MLOps Engineer

Luxoft • Hyderabad

On-site
INR 4,200,000 - 6,400,000
Senior Platform Engineer
Senior Platform Engineer

EPAM Systems • Hyderabad

On-site
INR 2,600,000 - 5,200,000
Senior Platform Engineer
Senior Platform Engineer

EPAM Systems • Maharashtra

On-site
INR 2,400,000 - 4,200,000
Senior Platform Engineer
Senior Platform Engineer

EPAM Systems • Coimbatore District

On-site
INR 1,800,000 - 3,000,000
Senior Platform Engineer
Senior Platform Engineer

EPAM Systems • Bengaluru

On-site
INR 3,000,000 - 5,400,000
AWS DevOps Engineer
AWS DevOps Engineer

Luxoft • Gurugram District

On-site
INR 900,000 - 1,500,000
AWS certifications
AIOps exposure
Enterprise-scale Kubernetes
+2
AWS DevOps Engineer
AWS DevOps Engineer

Luxoft • Hyderabad

On-site
INR 4,000,000 - 7,000,000