Senior Software Engineer - SRE

Incedo

Bengaluru

On-site

INR 2,500,000 - 4,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Incedo in Bengaluru is seeking a DevOps/SRE professional to participate in 24x7 follow-the-sun operations and On-Call support, addressing infrastructure alerts and incidents via triage, escalation, and service restoration.

You will monitor KPIs, perform daily triage, troubleshoot critical incidents with cross-functional teams, and drive process improvements to improve ticket routing, response times, and RCA closures. This role emphasizes ownership and customer outcomes.

Qualifications

  • Engineering degree in Computer Science or a related technical field.
  • 3-4 years of experience in DevOps, SRE, Cloud Operations, Platform Engineering, Production Engineering, or related roles.
  • Hands-on experience with one or more major cloud providers: AWS, Google Cloud Platform, or Microsoft Azure.
  • Hands-on experience with Kubernetes and containerised environments.
  • Experience with Infrastructure as Code, preferably Terraform, along with Jenkins, CI/CD, GitOps, and automation practices.
  • Strong Linux fundamentals and practical knowledge of networking, troubleshooting, and distributed systems.
  • Experience with monitoring, observability, alerting, Incident Management, and production troubleshooting.
  • Understanding of On-Call operations, RCA, SLA, SLO, and production reliability practices.
  • Strong analytical and troubleshooting skills with the ability to work effectively during customer-critical situations and production incidents.
  • Good communication and cross-functional collaboration skills.
  • Strong ownership mindset with a focus on customer outcomes, operational excellence, and continuous improvement.

Responsibilities

  • Participate in 24x7 follow-the-sun operations and On-Call support, handling alerts and incidents through triage, escalation, service restoration, and handoffs.
  • Monitor KPIs such as Customer Issue aging, DOHD ticket aging, response time, backlog burn-down, escalations, go-live incidents, and RCA SLA compliance.
  • Daily triage of new, aging, blocked, and escalated work with clear ownership and timely closure.
  • Troubleshoot customer-critical and platform incidents in collaboration with Engineering, Cloud Platform, Customer Enablement, and Support to restore service.
  • Improve DOHD processes for Customer Issues, CE questions, and requests to increase ticket capture, improve routing, and reduce response times.
  • Support internal RCA processes through DOHD workflows, SLA tracking, and timely closure of corrective and preventive actions.
  • Skills That Are Nice to Have: automation using Python, APIs, or workflow automation; AI-assisted support, knowledge retrieval, and self-service solutions.

Skills

DevOps
SRE
Cloud Operations
Platform Engineering
Production Engineering
Kubernetes
Terraform
CI/CD
GitOps
Linux
Networking
Monitoring
Incident Management
RCA
SLA/SLO
Communication
Ownership

Education

Engineering degree in Computer Science or related field

Tools

AWS
Google Cloud Platform
Microsoft Azure
Jenkins
CI/CD pipelines
Terraform

Job description

Role & responsibilities

Participate in 247 follow-the-sun operations and On-Call support, responding to infrastructure alerts and incidents through triage, escalation, service restoration, and effective handoffs.

  • Monitor Customer Experience and Operational KPIs, including Customer Issue and DOHD ticket aging, response time, backlog burn-down, escalations, go-live incidents, and RCA SLA compliance.
  • Participate in daily triage of new, aging, blocked, and escalated work, ensuring clear ownership and timely closure.
  • Troubleshoot customer-critical and platform incidents, collaborating with Engineering, Cloud Platform, Customer Enablement, and Support to restore service.
  • Improve DOHD processes for Customer Issues, CE questions, and requests to increase ticket capture, improve routing, and reduce response times.
  • Support internal RCA processes through DOHD workflows, SLA tracking, and timely closure of corrective and preventive actions.
Preferred candidate profile

- Engineering degree in Computer Science or a related technical field.

  • 3-4 years of experience in DevOps, SRE, Cloud Operations, Platform Engineering, Production Engineering, or a related role.
  • *Engineering degree in Computer Science or a related technical field. Production environment.
  • Hands-on experience with one or more major cloud providers: AWS, Google Cloud Platform, or Microsoft Azure.
  • Hands-on experience with Kubernetes and containerised environments.
  • Experience with Infrastructure as Code, preferably Terraform, along with Jenkins, CI/CD, GitOps, and automation practices.
  • Strong Linux fundamentals and practical knowledge of networking, troubleshooting, and distributed systems.
  • Experience with monitoring, observability, alerting, Incident Management, and production troubleshooting.
  • Understanding of On-Call operations, RCA, SLA, SLO, and production reliability practices.
  • Strong analytical and troubleshooting skills with the ability to work effectively during customer-critical situations and production incidents.
  • Good communication and cross-functional collaboration skills.
  • Strong ownership mindset with a focus on customer outcomes, operational excellence, and continuous improvement.

Skills That Are Nice to Have:

  • Experience automating repetitive workflows using Python, APIs, or workflow automation.
  • Experience with AI-assisted support, knowledge retrieval, and self-service solutions.
  • Familiarity with SRE, ITIL, Incident Management, and Problem Management practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Lead
SRE Lead

Acldigital • Ahmedabad District

On-site
INR 1,500,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Lloyds Technology Centre • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Bengaluru

On-site
INR 2,500,000 - 4,200,000
SRE Lead
SRE Lead

Bounteous • Gurugram District

Hybrid
INR 1,200,000 - 2,400,000
Site Reliability Engineer (SRE)- Python
Site Reliability Engineer (SRE)- Python

Apexon • Bengaluru

On-site
INR 1,100,000 - 1,500,000
Lead SRE
Lead SRE

Baazi Games • New Delhi

On-site
INR 3,500,000 - 6,000,000
Senior SRE Technical Specialist
Senior SRE Technical Specialist

United States Digital Space LLC • Karnataka

On-site
INR 1,500,000 - 2,000,000
AWS SRE Professional
AWS SRE Professional

Infosys • Bengaluru

On-site
INR 900,000 - 1,300,000
Site Reliability Engineer
Site Reliability Engineer

Epam Systems • Bengaluru

Hybrid
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer (SRE) / DevOps Engineer
Senior Site Reliability Engineer (SRE) / DevOps Engineer

Umanist Staffing LLC • Maharashtra

On-site
INR 3,500,000 - 5,500,000