Lead Engineer – Site Reliability Engineering

CBTS

Chennai District

On-site

INR 5,000,000 - 7,500,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

CBTS in India is seeking a Lead Engineer for Site Reliability Engineering to ensure high availability, performance, scalability, and resilience of cloud and infrastructure platforms through SRE principles, automation, observability, and continual reliability improvements across services.

You will drive automation for provisioning and deployment, build observability dashboards, coordinate incident response, and mentor engineers in SRE best practices while partnering with DevOps and platform teams.

Responsibilities

  • Implement SRE frameworks, SLIs/SLOs/SLAs, error budgets, performance engineering, and reliability guardrails across cloud platforms & services.
  • Drive automation for provisioning, deployment, configuration management, drift control, patching, recovery, and operations workflows.
  • Build observability stack, dashboards, anomaly detection, synthetic tests, runbooks, incident readiness and RCA automation.
  • Partner with DevOps, Platform Engineering, Cloud Engineering & application squads to define reliability patterns, capacity planning & scalable workload landing models.
  • Lead incident response, major incident coordination, postmortem improvement actions, resiliency testing, fault injection, chaos engineering initiatives.
  • Ensure infra security alignment, vulnerability remediation, compliance & secure configuration baselines in cloud infrastructure.
  • Mentor engineers in SRE best practices, automation pipelines, tooling standardization and operational excellence.

Job description

About this position

CBTS serves enterprise and midmarket clients in all industries across the United States and Canada. CBTS combines deep technical expertise with a full suite of flexible technology solutions- including Application Modernization, Managed Hybrid Cloud, Cybersecurity, Unified Communications, and Infrastructure solutions. From developing and deploying modern applications and the secure, scalable platforms on which they run, to managing, monitoring, and optimizing their operations, CBTS delivers comprehensive technology solutions for its clients' transformative business initiatives. For more information, please visit www.cbts.com .

OnX is a leading technology solution provider that serves businesses, healthcare organizations, and government agencies across Canada. OnX combines deep technical expertise with a full suite of flexible technology solutions- including Generative AI, Application Modernization, Managed Hybrid Cloud, Cybersecurity, Unified Communications, and Infrastructure solutions. From developing and deploying modern applications and the secure, scalable platforms on which they run, to managing, monitoring, and optimizing their operations, OnX delivers comprehensive technology solutions for its clients' transformative business initiatives. For more information, please visit www.onx.com .

Lead Engineer – Site Reliability Engineering
Role Purpose

Ensures high availability, performance, scalability, and resilience of cloud and infrastructure platforms by applying SRE engineering principles, automation-first practices, observability, and continual reliability improvements across services and platforms.

Key Responsibilities
  • Implement SRE frameworks, SLIs/SLOs/SLAs, error budgets, performance engineering, and reliability guardrails across cloud platforms & services.
  • Drive automation for provisioning, deployment, configuration management, drift control, patching, recovery, and operations workflows.
  • Build observability stack, dashboards, anomaly detection, synthetic tests, runbooks, incident readiness and RCA automation.
  • Partner with DevOps, Platform Engineering, Cloud Engineering & application squads to define reliability patterns, capacity planning & scalable workload landing models.
  • Lead incident response, major incident coordination, postmortem improvement actions, resiliency testing, fault injection, chaos engineering initiatives.
  • Ensure infra security alignment, vulnerability remediation, compliance & secure configuration baselines in cloud infrastructure.
  • Mentor engineers in SRE best practices, automation pipelines, tooling standardization and operational excellence.

Share this Job:

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Engineer Site Reliability Engineering
Lead Engineer Site Reliability Engineering

CBTS • Chennai District

On-site
INR 3,000,000 - 6,000,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
Lead SRE
Lead SRE

Cvent, Inc. • India

On-site
INR 2,500,000 - 4,500,000
Lead SRE
Lead SRE

Cvent • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineer - Career
Site Reliability Engineer - Career

Equifax • Pune District

On-site
INR 2,400,000 - 4,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Hyderabad

On-site
INR 1,400,000 - 2,200,000
Site Reliability Architect
Site Reliability Architect

Good co India • India

Remote
INR 2,400,000 - 4,200,000
Site Reliability Engineer
Site Reliability Engineer

Spot Your Leaders & Consulting • Pune District

On-site
INR 2,500,000 - 4,000,000
Senior Consultant - Site Reliability Engineer
Senior Consultant - Site Reliability Engineer

Darwinbox Digital Solutions Pvt. Ltd. • Hyderabad

On-site
INR 3,000,000 - 5,200,000