Site Reliability Engineer (SRE)

Tata Consultancy Services

Bengaluru

On-site

INR 2,800,000 - 4,000,000

Full time

11 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Tata Consultancy Services in Bengaluru is seeking a Site Reliability Engineer (SRE) with 8–10 years of experience. The role focuses on reliability engineering, DevOps, cloud infrastructure, and incident resolution in a high-demand eCommerce environment.

The candidate will work across product teams, implement IaC with Terraform/Ansible, and maintain monitoring with Splunk and Grafana. Shift-based on-call duties and ongoing process improvements are expected.

Qualifications

  • Cross-functional collaboration with reliability as a core capability.
  • Apply reliability practices with support from governance teams.
  • 5+ years in SRE, maintenance, operations, or development.
  • Strong experience in eCommerce environments.
  • Hands-on in DevOps, CI/CD and automated testing.
  • Experience with API-based frameworks (e.g., Commerce tools).
  • CI/CD workflows using GitHub Actions.
  • Proficient in cloud-based infra (Azure, GCP).
  • Infra as Code using Terraform/Ansible.
  • Experience with ITIL/ITSM in microservices.
  • Monitoring with Splunk, Grafana.
  • SRE metrics: SLI/SLO/Error Budget.
  • Managed cloud Kubernetes services (AKS, GKE).
  • Collaboration with product teams for stable production.
  • Technical analysis and incident troubleshooting.
  • Preventive monitoring improvements.
  • Automation to reduce toil.
  • On-call support for critical incidents.

Responsibilities

  • Analyze critical issues and resolve them promptly.
  • Lead and collaborate within teams; handle development challenges.
  • Collaborate with product teams to ensure predictable operations.
  • Participate in on-call rotations for production incidents.

Skills

Reliability engineering
DevOps practices
CI/CD workflows
Cloud infrastructure
Infrastructure as Code
Monitoring & incident analysis
Kubernetes (AKS/GKE)
On-call experience
Collaboration
Root cause analysis
API-based frameworks

Tools

GitHub Actions
Terraform
Ansible
Splunk
Grafana
ServiceNow
Kubernetes (AKS/GKE)

Job description

Role- Site Reliability Engineer (SRE)

Desired Experience Range- 8-10 Years

Location- Bangalore


Must-Have**
  • Work in a cross functional team working with Reliability as Expertise in a product or a product area.
  • Apply Reliability engineering practices with support from SRE governance teams.
  • 5+ years of experience in Site Reliability Engineering, maintenance & operations and/or development.
  • Strong working experience eCommerce.
  • Strong working experience in DevOps practices (automated testing, CI/CD etc.).
  • Experience within solutions architecture and how to fast pinpoint causes of issues.
  • Experience from working with API-based frameworks (e.g., Commerce tools or Fabric is ideal).
  • Experience in building CI/CD workflows using GitHub Actions.
  • Experience of maintaining/supporting and/or developing desktop and mobile applications.
  • Experience working on cloud-based infrastructure e.g., Azure and GCP.
  • Experience in provisioning Infra resources leveraging Infra as Code (Terraform / Ansible).
  • Experience from ITIL support processes and ITSM tools (e.g., ServiceNow) in a microservices context.
  • Experience in monitoring tools (Splunk, Grafana etc.).
  • Experience working through SRE Metrics such as SLI, SLO and Error Budget.
  • Experience with managed cloud Kubernetes services (e.g. AKS, GKE).
  • Ensure delivery quality and supply KPI reporting.
  • Collaborate closely within product teams to ensure predictable operations and minimal disruptions to Production.
  • Collaborate closely within your Capability, share best practices as well as discuss and improve on operations ways of working.
  • Technical analysis, troubleshooting of complex issues/Incidents in production.
  • Improve monitoring performance by focusing on preventive measures.
  • Product Improvements (code & log analysis).
  • Continuous improvement on proactive monitoring, housekeeping automation to proactively detect and avoid incidents.
  • Automate processes impacting development and production leveraging tools and building scripted solutions.
  • Participate in On-Call technical support to resolve business critical incidents.
Good to have
  • Familiarity with common tech stacks in Headless Ecommerce is a nice to have.
  • Knowledge of design principles and fundamentals of solutions architecture is a plus.
  • Understanding of performance engineering (Application Reliability).
  • Knowledge of multiple front-end languages and libraries (ReactJS, React Native, NodeJS).
  • Knowledge of Azure DevOps and/or other cloud environments is nice to have.
  • A passion for problem solving with strong analytical capabilities.
  • Stay current on technical trends to suggest innovative tools and approaches to interesting problems.
  • Knowledge on at least one of {Python, Ruby, Java, C#, Go} at an intermediate level.
Roles & Responsibilities-
  • Should be able to analyze the critical issue and resolve in time
  • Willingness to accept development challenges and work towards resolving them
  • Should be a good team player and should be able to lead the team
  • Should be willing to work in shifts

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Zorba AI • Chennai District

On-site
INR 4,000,000 - 7,500,000
Site Reliability Engineer
Site Reliability Engineer

Zorba AI • Hyderabad

On-site
INR 4,000,000 - 7,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hucon Solutions • Chennai District, Bengaluru, Hyderabad

On-site
INR 1,000,000 - 1,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Spot Your Leaders & Consulting • Pune District

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer
Site Reliability Engineer

S P A Enterprise Info Services • Chennai District

On-site
INR 1,800,000 - 2,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GSPANN • Hyderabad

On-site
INR 1,500,000 - 3,000,000
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Mumbai

On-site
INR 800,000 - 1,200,000
Resilience and Reliability Engineer
Resilience and Reliability Engineer

EY • Pune District, Gurugram District, Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Site Reliability Engineer
Site Reliability Engineer

Snapmint • Gurugram District

On-site
INR 800,000 - 1,200,000