Senior Site Reliability Engineer

Boston Consulting Group (BCG)

Gurgaon

Hybrid

INR 1,200,000 - 1,500,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Boston Consulting Group is seeking a Senior Site Reliability Engineer to own reliability across a defined area, spanning infrastructure, cloud, observability, automation, identity, security, and network operations. The role emphasizes reducing toil, embedding governance, mentoring engineers, and communicating trade-offs to both technical and non-technical audiences.

Ideal candidates bring 5–8 years in SRE/Platform roles, hands-on cloud experience (AWS/Azure), IaC (Terraform), and strong

Qualifications

  • 5–8 years in Site Reliability Engineering, Platform Engineering, or related roles.
  • Hands-on experience across cloud, automation, observability, CI/CD.
  • Designing and implementing automation and reliability solutions at scale.
  • Deep knowledge of one cloud platform (AWS or Azure).
  • Experience with Infrastructure-as-Code (Terraform) and CI/CD pipelines.
  • Strong scripting experience (Python).
  • Leading incident response and driving systemic improvements.
  • Experience with enterprise observability platforms like Splunk or Datadog.

Responsibilities

  • Run and continuously improve reliability engineering systems within scope.
  • Design and implement solutions to eliminate toil at scale.
  • Shape engineering standards and reusable frameworks across SRE practices.
  • Lead incident response and post-incident learning.
  • Mentor senior engineers in reliability, automation, and observability.
  • Drive cross-team collaboration to embed reliability and governance.
  • Communicate status, risks, and recommendations to leadership forums.
  • Contribute to monthly operational reviews with metrics on service health and performance.

Skills

SRE domains
Cloud platforms
CI/CD pipelines
Automation
Observability
Security and identity
Terraform
Python scripting
Incident management
Network security

Tools

Terraform
CI/CD
Splunk
Datadog
Kubernetes
Docker
Entra ID
HashiCorp Vault
OWASP / security tooling

Job description

Job Description:

Who We Are

Boston Consulting Group partners with leaders in business and society to tackle their most important challenges and capture their greatest opportunities. BCG was the pioneer in business strategy when it was founded in 1963. Today, we help clients with total transformation-inspiring complex change, enabling organizations to grow, building competitive advantage, and driving bottom-line impact. To succeed, organizations must blend digital and human capabilities. Our diverse, global teams bring deep industry and functional expertise and a range of perspectives to spark change. BCG delivers solutions through leading-edge management consulting along with technology and design, corporate and digital ventures—and business purpose. We work in a uniquely collaborative model across the firm and throughout all levels of the client organization, generating results that allow our clients to thrive.

What You'll Do

The Senior Site Reliability Engineeris responsible forrunning the engineering capability behind a defined area of reliability across the organisation. The role works across multiple SRE disciplines including infrastructure, cloud, observability, automation, identity, security, and network operations, applying engineering thinking to reduce operational toil, improve resilience, and embed reliability and governance into delivery and operational workflows. The role drives engineering quality and consistency within its scope of responsibility, contributes to wider engineering standards, and helps shape how reliability is delivered across the organisation. It builds reusable patterns, mentors engineers, and provides senior engineering input across a wider set of stakeholders. The ideal candidate is a senior practitioner who is comfortable operating across multiple domains, balances delivery with mentorship, and can articulate engineering trade-offs clearly to both technical and non-technical audiences.

Core Responsibilities
  • Run and continuously improve the reliability engineering systems within scope, including automation, pipelines, observability, and operational tooling.
  • Design and implement engineering solutions thateliminateoperational toil at scale and embed reliability into delivery workflows.
  • Help shape engineering standards, patterns, and reusable frameworks across the SRE practice.
  • Lead the engineering response to complex incidents within scope, drive systemic remediation, and contribute to post-incident learning.
  • Mentor and coachlesssenior engineers across reliability engineering, automation, observability, and SRE principles.
  • Drive cross-team collaboration with engineering, platform, and operations functions to embed reliability and governance through engineering controls.
  • Communicate engineering status, risks, and recommendations clearly to senior stakeholders and leadership forums.
  • Contribute to monthly operational reviews with structured metrics on service health, ingestion or pipeline performance, automation coverage, and improvement progress.
What You'll Bring
  • 5–8years of experience in Site Reliability Engineering, Platform Engineering, or related operational engineering disciplines.
  • Strong hands-on experience across multiple SRE domains, including cloud, automation, observability, and CI/CD.
  • Demonstrated experience designing and implementing automation and reliability solutions at scale.
  • Deep knowledge of at least one cloud platform (AWS or Azure), including networking, identity, and observability primitives.
  • Experience with Infrastructure-as-Code (e.g. Terraform) and CI/CD pipelines.
  • Strong scripting experience (e.g. Python).
  • Experience leading incident response and driving systemic improvement.
  • Strong stakeholder engagement and technical communication skills.
  • Deep hands-on experience with one or more enterprise observability platforms (e.g. Splunk, Datadog).
  • Proven experience designing and operating telemetry pipelines, ingestion controls, and observability cost management.
  • Proven experience designing signals (SLIs, SLOs, synthetic checks, alerts) and ops automation triggered from those signals.
  • Experience driving SLO/SLI practices across multiple teams.
  • Deep hands-on experience operating cloud infrastructure across at least two of AWS, Azure, GCP, or Alibaba Cloud.
  • Proven experience designing reusableIaCpatterns and landing zone components across cloud providers.
  • Strong working knowledge of cloud networking, account management, identity primitives, and policy enforcement across providers.
  • Experience driving cloud platform engineering standards and governance across multiple teams.
  • Deep hands-on experience with identity platforms (e.g. Entra ID) and secrets management (e.HashiCorpVault).
  • Proven experience designing OIDC, workload identity, and dynamic credential patterns.
  • Experience driving Zero Trust and least-privilege adoption across multiple teams.
  • Deep hands-on experience with security tooling embedded in CI/CD pipelines.
  • Proven experience designing policy-as-code controls and secure-by-default patterns.
  • Experience driving secure engineering adoption across multiple teams.
  • Deep hands-on experience with hybrid and cloud network architectures.
  • Proven experience designing automated network controls throughIaC.
  • Experience driving Zero Trust segmentation and network observability adoption.
Preferred Qualifications
  • Experience working within a federated, multi-cloud, or large enterprise environment.
  • Familiarity with containerisation (Docker) and orchestration (Kubernetes).
  • Experience with secrets management tooling (e.HashiCorpVault).
  • Cloud certification at professional level.
  • Experience with policy-as-code tooling (e.g. OPA, Sentinel).
  • Experience contributing to engineering communities of practice.
  • Experience with AIOps, noise reduction, and event correlation.
  • Experience with event-driven ops automation platforms (e.g. ServiceNow, PagerDuty, custom workflows).
  • Ability to lead complex observability platform incidents and capacity reviews.
  • Experience with cloud FinOps, cost engineering, and chargeback tooling.
  • Hands-on experience with Alibaba Cloud platform architecture.
  • Experience with cloud policy-as-code tools (e.g. AWS Service Control Policies, Azure Policy, OPA).
  • Strong understanding of identity-related security risks and mitigations.
  • Strong understanding of common security risks and mitigations across the SDLC.
  • Strong understanding of network reliability, observability, and security patterns.
Who You'll Work With
  • Hybrid or on-site work model.
  • Operates as a senior individual contributor with mentorship and cross-team influence.
  • Expected toparticipatein on-call rotation and lead incident response.
  • Occasional travel may be requiredfor team or stakeholder engagement.

Boston Consulting Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, age, religion, sex, sexual orientation, gender identity / expression, national origin, disability, protected veteran status, or any other characteristic protected under national, provincial, or local law, where applicable, and those with criminal histories will be considered in a manner consistent with applicable state and local laws.

BCG is an E - Verify Employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Boston Consulting Group (BCG) • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Senior Associate Site Reliability Engineer
Senior Associate Site Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Hyderabad

On-site
INR 1,400,000 - 2,000,000
Lead Engineer - Reliability Engineering
Lead Engineer - Reliability Engineering

StoneX Group Inc. • Bengaluru

Hybrid
INR 3,500,000 - 6,000,000
Site Reliability Engineer - Vice President
Site Reliability Engineer - Vice President

Citi • Maharashtra

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Hyderabad

On-site
INR 1,400,000 - 2,200,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Chennai District

On-site
INR 2,500,000 - 4,000,000
Lead SRE
Lead SRE

UST • Bengaluru

On-site
INR 3,000,000 - 5,500,000