Cloud Operations Lead – SRE / DevOps / Platform Engineering

PeoplePilot

Pune District

On-site

INR 2,600,000 - 5,200,000

Full time

11 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

PeoplePilot is seeking an experienced Cloud Operations Lead with a strong background in SRE, DevOps, and Platform Engineering. You will lead a small team while hands-on managing cloud platforms, reliability, security, and production support. Expect ~80% focus on cloud operations, with leadership responsibilities.

This role emphasizes operating across AWS-based platforms, Terraform/IaC, and CI/CD tooling, driving automation and governance in an enterprise environment.

Qualifications

  • Proven experience in Linux administration and operations.
  • Hands-on AWS cloud operations including IAM, EC2, Networking and EKS.
  • Kubernetes administration in production environments.
  • Strong Terraform/IaC skills for infrastructure provisioning.
  • CI/CD tooling knowledge (GitHub Actions, Jenkins).
  • Experience with monitoring/observability tools (Datadog, Prometheus, Grafana).
  • Incident management and RCA capabilities.
  • Security operations including vulnerability remediation and secrets rotation.
  • Experience navigating enterprise change-management processes.

Responsibilities

  • Lead cloud operations and production support across AWS platforms.
  • Troubleshoot Linux systems, cloud infra, networking, and Kubernetes runtimes.
  • Drive reliability through monitoring, automation, and incident handling.
  • Build and maintain IaC using Terraform, Ansible, and Helm.
  • Support and optimize CI/CD pipelines and deployment automation.
  • Design monitoring, alerts, dashboards, and runbooks for platforms.
  • Lead vulnerability remediation, access governance, and platform hardening.
  • Automate provisioning, OS/AMI upgrades, and run day-2 activities.
  • Support deployments, release management, and change control.

Skills

Linux Administration
AWS Cloud Operations
Kubernetes Administration
Terraform & IaC
CI/CD
Monitoring & Observability
Incident Management
Security Operations
Change Management

Tools

GitHub Actions
Jenkins
Terraform
Ansible
Helm
Datadog
Prometheus
Grafana
Nagios

Job description

Shift: Overlap with US & EU Business Hours

Immediate Joiners required (15 days notice)

Role Summary

We are seeking an experienced Cloud Operations Lead with a strong background in Site Reliability Engineering (SRE), DevOps, and Platform Engineering. The ideal candidate will be responsible for ensuring the reliability, security, and operational excellence of cloud-based platforms and services while leading a small team of engineers.

This is a hands-on role with approximately 80% focus on Cloud Operations, Production Support, Reliability, and Platform Ownership, combined with leadership responsibilities.

Key Responsibilities
  • Lead cloud operations and production support activities across AWS-based platforms.
  • Manage and troubleshoot Linux systems, cloud infrastructure, networking, and Kubernetes environments.
  • Drive operational excellence through monitoring, observability, automation, and incident management.
  • Build and maintain Infrastructure as Code (IaC) using Terraform, Ansible, and Helm.
  • Support and optimize CI/CD pipelines using GitHub Actions, Jenkins, and deployment automation tools.
  • Design and implement monitoring, alerting, dashboards, runbooks, and operational standards.
  • Lead vulnerability remediation, secrets management, access governance, and platform hardening initiatives.
  • Automate infrastructure provisioning, OS/AMI upgrades, and day-2 operational activities.
  • Support production deployments, release management, and change control processes.
  • Collaborate with engineering teams on onboarding, platform readiness, access management, and operational best practices.
  • Mentor and guide junior engineers while driving continuous service improvement.
Required Skills (Non-Negotiable)
  • Strong Linux Administration and Troubleshooting
  • AWS Cloud Operations (IAM, EC2, Networking, EKS)
  • Kubernetes Administration and Production Support
  • Terraform and Infrastructure as Code
  • CI/CD Tools (GitHub Actions, Jenkins)
  • Monitoring & Observability (Datadog, Prometheus, Grafana, SignalFx, Nagios, or similar)
  • Incident Management, Root Cause Analysis, and Production Support
  • Security Operations including vulnerability remediation, access management, and secrets rotation
  • Experience working in enterprise environments with formal change management processes
Preferred Skills
  • DNS, Proxy, Edge Services, and Networking Platforms
  • Teleport, Bastion Hosts, Service Accounts, and Access Management Solutions
  • Container Security and Supply Chain Security
  • AMI/Image Lifecycle Management
  • AI-enabled Operations, Custom Agentic AI, or Hyperscaler AI Services
  • Lead a team of cloud/platform engineers.
  • Drive operational governance, service reliability, and process standardization.
  • Promote automation-first and reliability-first engineering practices.
  • Partner with stakeholders across Cloud, Infrastructure, Security, and Application teams.
Nice to Have
  • Experience in SRE, Platform Engineering, or Managed Services environments.
  • Exposure to AI-powered operations, observability, or automation solutions.
  • Experience supporting large-scale distributed systems and cloud-native applications.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Operations Lead
Cloud Operations Lead

NewVision Software • Pune District

On-site
INR 3,000,000 - 5,500,000
Staff Software Engineer
Staff Software Engineer

OG1121 Meridium Services and Labs Private Limited • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Relocation Assistance
Team Lead Cloud Engineer
Team Lead Cloud Engineer

Fynxt • Chennai District

On-site
INR 4,000,000 - 6,000,000
Site Reliability Engineer
Site Reliability Engineer

WorkSpan • Bengaluru

On-site
INR 2,200,000 - 3,600,000
DevOps & SRE Lead
DevOps & SRE Lead

Syngentagroup • Pune District

On-site
INR 4,000,000 - 7,000,000
SRE Engineer
SRE Engineer

Prodapt Solutions Private Limited • Chennai District

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineer
Site Reliability Engineer

Innodata Inc. • India

On-site
INR 2,400,000 - 4,000,000
Senior Director of DevOps
Senior Director of DevOps

GreyOrange • Gurugram District

On-site
INR 6,000,000 - 9,000,000
Manager - Cloud and Devops Engineer
Manager - Cloud and Devops Engineer

PepsiCo • Hyderabad

On-site
INR 1,800,000 - 2,900,000
Lead DevOps Engineer
Lead DevOps Engineer

Lenskart • Gurugram District

On-site
INR 1,200,000 - 2,400,000