Site Reliability Engineering Manager

Calance

United States

Hybrid

USD 150,000 - 200,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

A global technology company is looking for a Senior Site Reliability Engineering (SRE) Leader to manage a distributed SRE team. The role requires strong leadership, extensive experience in cloud services, and the ability to enhance automation. The ideal candidate will have over 15 years of expertise in relevant engineering roles and a proven track record in mentoring teams. This position offers the opportunity to drive key initiatives in resilient systems and compliance standards.

Qualifications

  • 15+ years of experience in infrastructure, SRE, or platform engineering roles.
  • 5+ years of team management or technical leadership in SRE or DevOps.
  • Strong communication and collaboration skills.

Responsibilities

  • Lead the global SRE team in delivering resilient, cost-efficient environments.
  • Mentor SREs in developing secure API integrations and orchestration pipelines.
  • Champion platform telemetry and continuous observability improvements.

Skills

Leadership
Cloud-native patterns
Team management
Automation (Python, Bash, Go)
CI/CD pipelines

Tools

AWS managed services (EKS, RDS)
Azure managed services (AKS, Storage)
Kubernetes
GitHub Actions
Terraform Cloud

Job description

-----WILL NEED TO OBTAIN SECURITY CLEARANCE-----

WHAT YOU WILL DO
  • Lead the global SRE team in delivering resilient, cost-efficient, and highly automated runtime environments.
  • Drive SRE architecture and roadmap execution across AWS, Azure, and on-prem extensions (e.g., Outposts).
  • Ensure alignment with compliance goals (FedRAMP, IL4/IL5, CMMC 2.0, Cyber Essentials+, SecNumCloud).
  • Mentor SREs in developing orchestration pipelines, IaC modules, and secure API integrations.
  • Partner with InfoSec, Infrastructure Engineering, and Dev teams to enforce Zero Trust patterns.
  • Define and implement regional resiliency standards for global SaaS availability and continuity.
  • Identify toil and lead initiatives to eliminate it through engineering solutions.
  • Serve as the senior escalation point for complex incident triage and root cause analysis.
  • Champion platform telemetry, customer-facing audit logs, and continuous observability improvements.
WHAT IT TAKES
  • Strong leadership skills with the ability to mentor and coach senior-level engineers.
  • Experience managing distributed teams across US, Canada, EU, and global time zones.
  • Excellent communication and cross-functional collaboration capabilities.
  • Strong understanding of cloud-native patterns, service-level objectives (SLOs), and error budgets.
  • Proven ability to translate operational pain points into engineering deliverables.
  • 15+ years of experience in infrastructure, SRE, or platform engineering roles.
  • 5+ years of team management or technical leadership in SRE or DevOps.
  • 3+ years of experience building and scaling CI/CD pipelines with tools such as GitHub Actions, ArgoCD, or Terraform Cloud.
  • 2+ years of software development experience.
  • Extensive experience with AWS and Azure managed services (EKS, AKS, RDS, Storage, ALB/NLB).
  • Hands‑on experience with Kubernetes, GitOps, and service mesh (e.g., Istio, Linkerd).
  • Strong automation experience using Python, Bash, or Go.
  • Experience implementing FedRAMP, CMMC, or comparable regulatory requirements.
  • Familiarity with secrets management, workload identity federation, and boundaryless access control (e.g., SPIFFE, IRSA).
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Staffing Science • Arizona

On-site
USD 180,000 - 240,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GovCIO • Arlington (VA)

On-site
USD 210,000 - 230,000
Lead SRE
Lead SRE

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 150,000 - 190,000
Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 210,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Cloud Site Reliability Engineer (SRE)
Cloud Site Reliability Engineer (SRE)

ecsfederal • Virginia (MN)

Hybrid
USD 130,000 - 180,000
DevOps / Site Reliability Engineer (SRE)
DevOps / Site Reliability Engineer (SRE)

Zoho • United States

On-site
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Good co India • United States

Remote
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

VantageScore® • San Francisco (CA)

On-site
USD 135,000 - 165,000
Medical insurance
Dental insurance
401(k) plan
+1