Senior SRE Engineer

EPAM Systems India Pvt Ltd

Chennai District

On-site

INR 1,800,000 - 3,200,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

EPAM Systems India Pvt Ltd is seeking a Senior SRE Engineer to join our in-office team. You will ensure reliability, scalability, and performance of critical cloud services, designing HA/DR strategies and building cloud infrastructure via IaC.

You will implement and optimize CI/CD pipelines, monitor systems with observability tools, and develop automation to streamline operations. Strong AWS, container, and scripting skills with leadership capabilities are required.

Qualifications

  • 4-8 years of overall IT experience.
  • 4+ years in SRE/DevOps/Cloud Infra roles.
  • Expertise in AWS services (EC2, S3, RDS, IAM, VPC, Lambda).
  • Infrastructure-as-Code using Terraform, AWS CDK, or CloudFormation.
  • CI/CD tooling such as Jenkins, GitHub Actions, or GitLab CI.
  • Containerization and orchestration with Docker, Kubernetes, ECS/EKS.
  • Observability with Datadog, New Relic, Prometheus, Grafana, ELK, CloudWatch.
  • Scripting in Python, Bash, or Go.
  • Networking, security, and IAM in cloud environments.
  • Excellent communication, problem-solving, and leadership to influence across teams.

Responsibilities

  • Design high-availability and disaster recovery strategies for critical workloads.
  • Build and manage cloud infrastructure using infrastructure-as-code tools.
  • Implement and optimize CI/CD pipelines for automated deployments.
  • Monitor performance and reliability through observability platforms.
  • Develop scripts and automation for operational efficiency.
  • Ensure robust networking, security, and IAM practices in cloud environments.
  • Lead cross-team collaboration to resolve reliability challenges.
  • Drive continuous improvements in infra and platform automation.

Skills

AWS EC2/S3/RDS/IAM/VPC/Lambda
Terraform
CloudFormation
CI/CD (Jenkins/GitHub Actions/GitLab)
Docker & Kubernetes
Observability (Datadog/New Relic/Prom/
Grafana/ELK/CloudWatch
Scripting (Python/Bash/Go)
Networking & IAM security
Leadership & communication

Tools

Terraform
AWS CDK
CloudFormation
Jenkins
GitHub Actions
GitLab CI
Docker
Kubernetes
ECS/EKS
Datadog
New Relic
Prometheus
Grafana
ELK
CloudWatch

Job description

We are seeking a Senior SRE Engineer to join our team on a full-time, in-office basis, responsible for ensuring the reliability, scalability, and performance of critical cloud infrastructure and services.

Responsibilities
  • Design and maintain high-availability and disaster recovery strategies for critical workloads
  • Build and manage cloud infrastructure using Infrastructure-as-Code tools
  • Implement and optimize CI/CD pipelines for automated deployments
  • Monitor system performance and reliability through observability platforms
  • Develop scripts and automation tools to streamline operations
  • Ensure robust networking, security, and identity/access management practices
  • Lead cross-team collaboration to resolve complex reliability challenges
  • Drive continuous improvement initiatives across infrastructure and platform automation
Requirements
  • 4-8 years of overall experience in IT
  • 4+ years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles
  • Expertise in AWS services such as EC2, S3, RDS, IAM, VPC, and Lambda
  • Knowledge of Infrastructure-as-Code using Terraform, AWS CDK, or CloudFormation
  • Background in CI/CD tools such as Jenkins, GitHub Actions, or GitLab CI
  • Proficiency in containerization and orchestration technologies including Docker, Kubernetes, and ECS/EKS
  • Competency in monitoring and observability tools such as Datadog, New Relic, Prometheus, Grafana, ELK, and CloudWatch
  • Skills in scripting or programming languages such as Python, Bash, or Go
  • Understanding of networking, security, and identity/access management in cloud environments
  • Excellent communication, problem-solving, and leadership skills with the ability to influence across teams
Nice to have
  • AWS or other Cloud Certification such as Solutions Architect or DevOps Engineer
  • Familiarity with AIOps, Serverless Architectures, and event-driven systems
  • Understanding of FinOps practices and cost optimization frameworks, along with SaaS monitoring tools such as Sumo Logic and PagerDuty
  • Exposure to Atlassian tools including Jira, Confluence, and Bitbucket, plus experience with SQL/NoSQL databases
  • Showcase of leading cross-functional reliability initiatives or platform-wide automation projects
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Bengaluru

On-site
INR 2,500,000 - 4,200,000
Senior SRE Engineer
Senior SRE Engineer

EPAM Systems • India

Hybrid
INR 1,800,000 - 2,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

UST • Pune District

On-site
INR 1,800,000 - 3,000,000
Senior Site Reliability Lead
Senior Site Reliability Lead

Generac • Pune District

On-site
INR 3,000,000 - 6,500,000
SRE Engineer
SRE Engineer

Prodapt Solutions Private Limited • Chennai District

On-site
INR 1,800,000 - 3,000,000
SRE+AWS Devops
SRE+AWS Devops

Virtusa • Bengaluru Urban

On-site
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

NCR Voyix • Chennai District

On-site
INR 3,000,000 - 5,400,000
Site Reliability Engineer
Site Reliability Engineer

Innodata Inc. • India

On-site
INR 2,400,000 - 4,000,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Site Reliability Engineer (SRE) – Core IT Infrastructure
Site Reliability Engineer (SRE) – Core IT Infrastructure

TECEZE • Chennai District

On-site
INR 1,000,000 - 2,000,000