Site Reliability Engineer (SRE) Support

Cognizant

Bengaluru Urban

On-site

INR 1,200,000 - 1,800,000

Full time

21 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Cognizant in Bangalore, India is seeking an experienced Site Reliability Engineer (SRE) to support and maintain mission-critical applications on Azure cloud. The role focuses on production support, incident management, RCA, and ensuring platform reliability.

The candidate will monitor systems with Dynatrace and Splunk, lead incident triage, and collaborate with cross-functional teams to drive improvements in automation, observability, and operational excellence. Shift-based work is required.

Qualifications

  • Experience in Production Support / SRE environment.
  • Hands-on expertise in managing critical production incidents.
  • Strong troubleshooting and analytical skills.
  • Ability to work in a shift-based support model.
  • Excellent communication and stakeholder management.

Responsibilities

  • Provide L2/L3 production support for enterprise applications and services.
  • Monitor health using Dynatrace, Splunk, and observability platforms.
  • Manage and resolve production incidents within defined SLA timelines.
  • Lead incident triaging and coordination with cross-functional teams during major incidents.
  • Perform detailed RCA and implement preventive actions to avoid recurrence.
  • Support Java-based applications running on Azure cloud.
  • Analyze application and system logs and performance metrics to identify issues proactively.
  • Collaborate with development, infra, and business teams to ensure service stability.
  • Drive continuous improvements in monitoring, alerting, automation and runbooks.
  • Participate in deployment validations, release support and change activities.
  • Create and maintain support documentation, runbooks, and operational procedures.

Skills

Incident Management
Root Cause Analysis
Service Reliability and Availability
Azure
Azure Infrastructure
Java Application Support
Linux Administration
Dynatrace
Splunk
Observability Tools
APM
Log Monitoring and Analysis
Alert Management
Shell/Python Scripting
ITIL Processes
DevOps and CI/CD
Kubernetes and Containers
Automation & Operational Excellence
Communication & Stakeholder Management

Education

Bachelor's or Master's Degree in Computer Science/IT/Engineering

Tools

Dynatrace
Splunk
Kubernetes

Job description

Job Description - Site Reliability Engineer (SRE) Support


Role: Site Reliability Engineer (SRE) - Production Support


Experience: 6 to 13 Years


Location: Bangalore


Work Model: 5 Days Work from Office (Mandatory)


Shift: 12:00 PM to 10:00 PM



Role Overview

We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain mission-critical applications and cloud infrastructure. The ideal candidate should possess strong expertise in Azure Cloud, Java application support, Linux administration, Incident Management, Observability, Dynatrace, Splunk, and Root Cause Analysis (RCA). This is a hands‑on production support role focused on ensuring platform reliability, availability, and operational excellence.



Key Responsibilities


  • Provide L2/L3 production support for enterprise applications and services.

  • Monitor application and infrastructure health using Dynatrace, Splunk, and observability platforms.

  • Manage and resolve production incidents within defined SLA timelines.

  • Lead incident triaging, troubleshooting, and coordination with cross-functional teams during major incidents.

  • Perform detailed Root Cause Analysis (RCA) and implement preventive actions to avoid recurring issues.

  • Support Java-based applications running on Azure cloud environments.

  • Analyze application logs, system logs, and performance metrics to identify issues proactively.

  • Work closely with development, infrastructure, and business teams to ensure service stability.

  • Drive continuous improvements in monitoring, alerting, automation, and operational processes.

  • Participate in deployment validations, release support, and change activities.

  • Create and maintain support documentation, runbooks, and operational procedures.



Mandatory Skills

Site Reliability Engineering (SRE)


  • Incident Management

  • Problem Management

  • Root Cause Analysis (RCA)

  • Service Reliability and Availability Management


Cloud


  • Microsoft Azure

  • Azure Infrastructure and Services

  • Cloud Operations & Support


Application Support


  • Java Application Support

  • Performance Analysis and Troubleshooting

  • Production Support


Linux


  • Linux Administration

  • System Monitoring and Troubleshooting


Monitoring & Observability


  • Dynatrace

  • Splunk

  • Observability Tools

  • Application Performance Monitoring (APM)

  • Log Monitoring and Analysis

  • Alert Management



Preferred Skills


  • Shell Scripting / Python Scripting

  • ITIL Processes
  • DevOps and CI/CD Concepts

  • Kubernetes and Container Technologies

  • Automation and Operational Excellence Practices



Desired Candidate Profile


  • Strong experience in Production Support / SRE environment.

  • Hands-on expertise in managing critical production incidents.

  • Excellent troubleshooting and analytical skills.

  • Ability to work effectively in a shift-based support model.

  • Strong communication and stakeholder management skills.

  • Experience supporting high-availability enterprise applications.



Qualification


  • Bachelor's Degree or Master's Degree in Computer Science, Information Technology, Engineering, or related discipline.



Summary

Experience: 6-13 Years


Location: Bangalore


Shift Timing: 12 PM - 10 PM


Work Model: 5 Days Work from Office


Role Type: Production Support / Site Reliability Engineering (SRE)


Primary Skills: Azure, Java Support, Linux, Dynatrace, Splunk, Incident Management, Observability, RCA, Production Support.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE) Support
Site Reliability Engineer (SRE) Support

Cognizant • Bengaluru

On-site
INR 2,400,000 - 4,200,000
SRE Lead
SRE Lead

3across • Bengaluru

Hybrid
INR 1,500,000 - 2,300,000
SRE Reliability Engineer
SRE Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Senior Site Reliability Engineer (SRE) Engineer
Senior Site Reliability Engineer (SRE) Engineer

Umanist Staffing • Pune District

On-site
INR 2,250,000 - 2,750,000
SRE
SRE

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 2,100,000
Senior Site Reliability Engineer (SRE) Engineer
Senior Site Reliability Engineer (SRE) Engineer

Umanist NA • Maharashtra

On-site
INR 1,500,000 - 2,500,000
Assistant Manager - Azure Site Reliability Engineer
Assistant Manager - Azure Site Reliability Engineer

Promaynov Advisory Services Pvt. Ltd • Bengaluru

On-site
INR 1,400,000 - 2,100,000
Software Engineer-DevOps
Software Engineer-DevOps

SMC Squared India • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Azure Site Reliability Engineer (SRE) + SQL - SaaS Operations
Azure Site Reliability Engineer (SRE) + SQL - SaaS Operations

Zensar Technologies • Pune District

On-site
INR 2,000,000 - 2,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000