AWS Site Reliability Engineering Lead

Infosys

Bengaluru

On-site

INR 1,500,000 - 2,100,000

Full time

8 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Infosys is seeking an AWS SRE Lead to manage and optimize a highly available AWS cloud platform. You will lead DevOps practices, improve incident response, and drive automation using IaC and CI/CD pipelines.

The role requires 5–9 years in AWS and SRE, with strong experience in monitoring, automation, and incident management. On-site Bengaluru location with global collaboration and security/compliance focus.

Qualifications

  • 5 to 9 years of experience in AWS Cloud and SRE/Production Support.
  • Strong knowledge of AWS services such as EC2, S3, RDS, IAM, VPC, and Route 53.
  • Experience with Linux/Unix administration.
  • Proficiency in Python or Shell Scripting.
  • Hands-on experience with monitoring tools such as CloudWatch, Grafana, or Prometheus.
  • Experience with incident management and troubleshooting production environments.
  • Knowledge of networking concepts including DNS, Load Balancers, TCP/IP, and VPN.
  • Experience with Git and CI/CD tools.

Responsibilities

  • Manage and maintain highly available and scalable AWS cloud infrastructure.
  • Monitor applications, servers, and cloud resources to ensure system reliability and performance.
  • Automate operational tasks using Python, Shell scripting, and IaC.
  • Troubleshoot production issues and perform RCA.
  • Implement and support CI/CD pipelines for seamless deployments.
  • Manage incident, problem, and change management activities.
  • Configure and maintain monitoring and alerting tools.
  • Collaborate with development and DevOps teams to improve system stability.
  • Ensure security, compliance, backup, and disaster recovery standards.
  • Participate in on-call support and production support activities.

Skills

AWS Cloud
SRE/Prod Support
EC2
S3
RDS
IAM
VPC
Route53
Linux Admin
Python
Shell Scripting
CloudWatch
Grafana
Prometheus
Incident Mgmt
CI/CD
Git

Tools

Docker
Kubernetes
Terraform
CloudFormation
Jenkins
ELK
Splunk
Datadog
New Relic

Job description

AWS SRE Lead Technology->Cloud Platform->Amazon Webservices DevOps,Technology->DevOps->Site Reliability Engineering(SRE)

Key Responsibilities
  • Manage and maintain highly available and scalable AWS cloud infrastructure.
  • Monitor applications, servers, and cloud resources to ensure system reliability and performance.
  • Automate operational tasks using Python, Shell scripting, and Infrastructure as Code (IaC).
  • Troubleshoot production issues and perform root cause analysis (RCA).
  • Implement and support CI/CD pipelines for seamless application deployments.
  • Manage incident, problem, and change management activities.
  • Configure and maintain monitoring and alerting tools.
  • Work closely with development and DevOps teams to improve system stability.
  • Ensure security, compliance, backup, and disaster recovery standards are met.
  • Participate in on-call support and production support activities.
Minimum Qualifications
  • 5 to 9 years of experience in AWS Cloud and SRE/Production Support.
  • Strong knowledge of AWS services such as EC2, S3, RDS, IAM, VPC, and Route 53.
  • Experience with Linux/Unix administration.
  • Proficiency in Python or Shell Scripting.
  • Hands-on experience with monitoring tools such as CloudWatch, Grafana, or Prometheus.
  • Experience with incident management and troubleshooting production environments.
  • Knowledge of networking concepts including DNS, Load Balancers, TCP/IP, and VPN.
  • Experience with Git and CI/CD tools.
Preferred Qualifications
  • Experience with Docker and Kubernetes.
  • Experience with Terraform or CloudFormation.
  • Knowledge of Jenkins and DevOps practices.
  • Experience with ELK, Splunk, Datadog, or New Relic.
  • Familiarity with ServiceNow and ITIL processes.
  • Experience in High Availability (HA) and Disaster Recovery (DR) environments.
  • AWS Certification is an added advantage.
  • Experience working in Agile/Scrum environments.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE+AWS Devops
SRE+AWS Devops

Virtusa • Bengaluru Urban

On-site
INR 1,500,000 - 2,500,000
Senior AWS Site Reliability Engineer
Senior AWS Site Reliability Engineer

Infosys • Bengaluru

On-site
INR 1,100,000 - 1,700,000
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Lyzr AI • Bengaluru

On-site
INR 1,000,000 - 2,000,000
Senior SRE Engineer
Senior SRE Engineer

TymblHub • Chennai District

On-site
INR 2,500,000 - 4,500,000
Lead SRE
Lead SRE

Cvent, Inc. • India

On-site
INR 2,500,000 - 4,500,000
Site Reliability Engineer_ AWS certified
Site Reliability Engineer_ AWS certified

PwC Acceleration Center India • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Chennai District

On-site
INR 2,500,000 - 4,000,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Site Reliability Engineer
Site Reliability Engineer

Saika Technologies Inc. • Hyderabad, Bengaluru

Hybrid
INR 3,000,000 - 4,200,000
Site Reliability Engineer
Site Reliability Engineer

Gemini Solutions • Gurugram District

On-site
INR 2,500,000 - 4,000,000