SRE Professional

BT Group

Bengaluru

Hybrid

INR 1,500,000 - 2,100,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

BT Group is seeking an experienced Site Reliability Engineer to ensure reliability, performance, and operational excellence of critical customer-facing applications across Siebel, Java, and AWS platforms. The role focuses on understanding end-to-end processes, architecture, and data flows to drive continuous service improvement and automated resilience.

Responsibilities include monitoring health with Dynatrace, investigating complex production issues, and driving incident, problem, and change

Qualifications

  • Experience supporting or operating enterprise applications within complex production environments.
  • Understanding of application architecture, system integrations, data flows and customer-facing business processes.
  • Experience troubleshooting production application issues using SQL/Oracle databases.
  • Hands-on with monitoring and observability platforms such as Dynatrace; dashboard creation and alert analysis.
  • Exposure to scripting or automation technologies such as Shell/Python.

Responsibilities

  • Develop understanding of customer journeys, business processes, apps, and integrations to support service outcomes.
  • Monitor service health and performance using observability platforms; identify risks and opportunities.
  • Investigate production issues via logs, queries, and monitoring data to support rapid resolution.
  • Provide technical oversight on investigations and ensure high-quality incident resolution.
  • Lead incident, problem, and change management with root cause analysis and continuous improvement.
  • Collaborate with engineering, business, and supplier teams for improvements aligned to objectives.
  • Develop and maintain service dashboards, operational reporting, and performance insights.

Skills

Enterprise apps
Production environments
Application architecture
Monitoring observability
SQL/Oracle
Dynatrace
Java
AWS
Incident management
Automation scripting
AI/GenAI for reliability
SRE fundamentals

Tools

Dynatrace

Job description

About the role

We are looking for an experienced Site Reliability Engineer (SRE) to ensure the reliability, performance, and operational excellence of critical customer-facing applications across Siebel, Java, and AWS platforms. The role is responsible for understanding end-to-end business processes, application architecture, system integrations, and customer journeys to effectively manage service reliability and drive continuous improvement. Working with support partners, engineering teams, and business stakeholders, the role uses observability, operational data, and technical analysis to identify issues, improve application resilience, increase automation, and enhance customer and business outcomes. The role also supports the adoption of AIOps and self-healing capabilities to improve operational efficiency and service quality.


Role & responsibilities
  • Develop a strong understanding of customer journeys, business processes, application architecture, and integrations across CRM platform, AWS, and supporting platforms to effectively support service outcomes.
  • Monitor service health, performance, and availability using observability platforms and operational metrics, proactively identifying risks and improvement opportunities.
  • Investigate complex production issues through analysis of application logs, database queries, transaction flows, and monitoring data to support rapid resolution and root cause identification.
  • Provide technical oversight on investigations, validating root causes, challenging assumptions, and ensuring high-quality incident resolution.
  • Driving incident, problem, and change management activities, leveraging root cause analysis, preventive actions, and continuous service improvement to enhance reliability and prevent recurring issues.
  • Partner with engineering, business, and supplier teams to deliver application, operational, and customer experience improvements aligned to business objectives.
  • Develop and maintain service dashboards, operational reporting, and performance insights to support governance, decision-making, and continuous improvement.
  • Identify and implement opportunities for automation, self-service, and AI-driven operational capabilities to improve efficiency, resilience, and customer outcomes

Skills and Experience Required
  • Experience supporting or operating enterprise applications within complex production environments.
  • Strong understanding of application architecture, system integrations, data flows and customer-facing business processes.
  • Ability to analyse application, middleware and infrastructure logs to identify issues and support root cause investigations.
  • Knowledge of SQL/Oracle databases with experience troubleshooting production application issues.
  • Hands‑on experience with monitoring and observability platforms such as Dynatrace, including dashboard creation, alert analysis and trend identification.
  • Understanding of Java-based applications, APIs, integration services and AWS-hosted platforms.
  • Experience working with incident, problem, change and release management processes.
  • Strong analytical and stakeholder management skills with the ability to challenge technical investigations, drive service improvements and influence engineering teams.
  • Ability to identify automation, resilience and customer experience improvement opportunities through operational insights.
  • Exposure to scripting or automation technologies such as Shell script, Python etc.
  • Experience leveraging AI/ML/GenAI capabilities to improve service reliability, operational efficiency, observability, incident response, root cause analysis, and self-healing automation.
  • Understanding of application and platform performance management, capacity planning, trend analysis, demand forecasting, and resilience engineering for business-critical production services.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Peoplefy • Pune District

On-site
INR 1,200,000 - 1,800,000
SRE Reliability Engineer
SRE Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Bengaluru

On-site
INR 2,500,000 - 4,000,000
SRE Developer
SRE Developer

Cloudxtreme • Hyderabad

On-site
INR 1,500,000 - 2,400,000
Application SRE
Application SRE

Cloudxtreme • Pune District

On-site
INR 1,200,000 - 1,800,000
Application Support Engineer
Application Support Engineer

Orcapod Consulting Services • Bengaluru

On-site
INR 1,500,000 - 2,200,000
Application SRE Engineer
Application SRE Engineer

ADROITENT • Hyderabad, Pune District, Bengaluru

Hybrid
INR 1,500,000 - 2,100,000
VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. • Bengaluru

On-site
INR 1,800,000 - 2,400,000
VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. • India

On-site
INR 2,000,000 - 4,000,000
Site Reliability Engineer - Lead
Site Reliability Engineer - Lead

Cloudxtreme • Hyderabad

On-site
INR 4,000,000 - 6,000,000
SRE
SRE

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 2,100,000