Site Reliability Engineer

3across

Bengaluru

Hybrid

INR 3,500,000 - 6,000,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

3across is seeking a Senior Site Reliability Engineer (SRE) to manage reliability and performance of critical production systems. The role demands hands-on work with Linux/Unix, Python and Shell scripting, and leadership in incident management.

The candidate will drive incident resolution, perform RCA, and collaborate with cross-functional teams to improve automation and resiliency across services in a Bengaluru-based environment.

Qualifications

  • 7–12 years of relevant experience in SRE, production support or reliability engineering.
  • Strong hands-on troubleshooting and scripting capabilities.

Responsibilities

  • Drive resolution and restoration of critical production incidents.
  • Identify incident severity, business impact, and escalation procedures.
  • Troubleshoot issues in Linux/Unix production environments.
  • Develop Python and Shell/Bash scripts for automation.
  • Perform root cause analysis and contribute to problem management.
  • Coordinate with cross-functional teams during incidents.
  • Ensure incident documentation and post-incident reviews.
  • Mentor technical teams during production support activities.

Skills

SRE
Python & Shell Scripting
Linux/Unix
Team Handling & Incident Mgmt

Tools

Monitoring/Observability Tools
Cloud Platforms
Chaos Engineering Tools
ServiceNow (ITSM)

Job description

Senior Site Reliability Engineer (SRE)

Experience: 712 Years
Job Type: Full Time

Job Description

We are looking for an experienced Site Reliability Engineer (SRE) to manage and improve the reliability, availability, and performance of critical production systems. The ideal candidate should have strong hands‑on experience in SRE, Linux/Unix, Python and Shell scripting, along with experience in handling technical teams and critical production incidents.

Key Responsibilities
  • Drive resolution and restoration of critical production incidents and ensure timely service recovery.
  • Identify incident severity, business impact, risks, and appropriate escalation procedures.
  • Troubleshoot complex issues across Linux/Unix-based production environments.
  • Develop and maintain Python and Shell/Bash scripts for automation and operational efficiency.
  • Perform root cause analysis (RCA) and contribute to Problem Management activities.
  • Investigate and resolve application, infrastructure, and data‑related production issues.
  • Coordinate with cross‑functional technology teams during critical incidents.
  • Ensure proper incident documentation, timelines, impact analysis, and resolution details.
  • Identify opportunities for automation, process improvement, and operational efficiency.
  • Support resiliency initiatives and contribute to improving system reliability.
  • Conduct incident reviews and identify recurring/known issues.
  • Handle and guide a team of technical professionals during production support activities.
  • Ensure adherence to operational, audit, compliance, and change‑management processes.
Required Skills
  • Site Reliability Engineering (SRE)
  • Python & Shell/Bash Scripting
  • Linux / Unix
  • Team Handling & Production Incident Management
Preferred Skills
  • Major Incident Management
  • Problem Management & RCA
  • SQL
  • Monitoring/Observability tools
  • Automation
  • ITIL / Service Management
  • Cloud technologies
  • Chaos Engineering / Resiliency
  • ServiceNow or similar ITSM tools
Candidate Profile
  • 7-12 years of relevant experience in SRE / Production Support / Application Support / Reliability Engineering.
  • Strong hands‑on troubleshooting and scripting capabilities.
  • Proven experience handling critical production incidents.
  • Experience in leading or mentoring technical teams.
  • Strong communication, analytical, and problem‑solving skills.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Lloyds Technology Centre • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Unified Consultancy Services • Karnataka

Hybrid
INR 1,500,000 - 3,000,000
Site Reliability Engineer (SRE) / DevOps Engineer
Site Reliability Engineer (SRE) / DevOps Engineer

New Era Technology • Gurugram District

On-site
INR 1,500,000 - 2,100,000
Senior Associate Site Reliability Engineer
Senior Associate Site Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Hyderabad

On-site
INR 1,400,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Snapmint • Gurugram District

On-site
INR 800,000 - 1,200,000
SRE Lead
SRE Lead

3across • Bengaluru

Hybrid
INR 1,500,000 - 2,300,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Zorba AI • Chennai District

On-site
INR 1,200,000 - 2,400,000
SRE Developer
SRE Developer

Cloudxtreme • Hyderabad

On-site
INR 1,500,000 - 2,400,000
Site Reliability Engineer
Site Reliability Engineer

Spot Your Leaders & Consulting • Pune District

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer
Site Reliability Engineer

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 2,600,000