Site Reliability Engineering Manager

Calance

United States

Hybrid

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A global technology company is looking for a Senior Site Reliability Engineering (SRE) Leader to manage a distributed SRE team. The role requires strong leadership, extensive experience in cloud services, and the ability to enhance automation. The ideal candidate will have over 15 years of expertise in relevant engineering roles and a proven track record in mentoring teams. This position offers the opportunity to drive key initiatives in resilient systems and compliance standards.

Qualifications

  • 15+ years of experience in infrastructure, SRE, or platform engineering roles.
  • 5+ years of team management or technical leadership in SRE or DevOps.
  • Strong communication and collaboration skills.

Responsibilities

  • Lead the global SRE team in delivering resilient, cost-efficient environments.
  • Mentor SREs in developing secure API integrations and orchestration pipelines.
  • Champion platform telemetry and continuous observability improvements.

Skills

Leadership
Cloud-native patterns
Team management
Automation (Python, Bash, Go)
CI/CD pipelines

Tools

AWS managed services (EKS, RDS)
Azure managed services (AKS, Storage)
Kubernetes
GitHub Actions
Terraform Cloud

Job description

-----WILL NEED TO OBTAIN SECURITY CLEARANCE-----

WHAT YOU WILL DO
  • Lead the global SRE team in delivering resilient, cost-efficient, and highly automated runtime environments.
  • Drive SRE architecture and roadmap execution across AWS, Azure, and on-prem extensions (e.g., Outposts).
  • Ensure alignment with compliance goals (FedRAMP, IL4/IL5, CMMC 2.0, Cyber Essentials+, SecNumCloud).
  • Mentor SREs in developing orchestration pipelines, IaC modules, and secure API integrations.
  • Partner with InfoSec, Infrastructure Engineering, and Dev teams to enforce Zero Trust patterns.
  • Define and implement regional resiliency standards for global SaaS availability and continuity.
  • Identify toil and lead initiatives to eliminate it through engineering solutions.
  • Serve as the senior escalation point for complex incident triage and root cause analysis.
  • Champion platform telemetry, customer-facing audit logs, and continuous observability improvements.
WHAT IT TAKES
  • Strong leadership skills with the ability to mentor and coach senior-level engineers.
  • Experience managing distributed teams across US, Canada, EU, and global time zones.
  • Excellent communication and cross-functional collaboration capabilities.
  • Strong understanding of cloud-native patterns, service-level objectives (SLOs), and error budgets.
  • Proven ability to translate operational pain points into engineering deliverables.
  • 15+ years of experience in infrastructure, SRE, or platform engineering roles.
  • 5+ years of team management or technical leadership in SRE or DevOps.
  • 3+ years of experience building and scaling CI/CD pipelines with tools such as GitHub Actions, ArgoCD, or Terraform Cloud.
  • 2+ years of software development experience.
  • Extensive experience with AWS and Azure managed services (EKS, AKS, RDS, Storage, ALB/NLB).
  • Hands‑on experience with Kubernetes, GitOps, and service mesh (e.g., Istio, Linkerd).
  • Strong automation experience using Python, Bash, or Go.
  • Experience implementing FedRAMP, CMMC, or comparable regulatory requirements.
  • Familiarity with secrets management, workload identity federation, and boundaryless access control (e.g., SPIFFE, IRSA).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

GovCIO • Arlington (VA)

Hybrid
USD 210,000 - 230,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Govcio LLC • United States

Hybrid
USD 210,000 - 230,000
Cloud Security Engineer – SRE
Cloud Security Engineer – SRE

TechDigital Group • Alpharetta (GA)

On-site
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
SRE Lead engineer
SRE Lead engineer

TechDigital Group • Bellevue (WA)

On-site
USD 100,000 - 130,000
Site Reliability Engineer – Lead
Site Reliability Engineer – Lead

Jobtailor • Arizona

On-site
USD 140,000 - 230,000
SRE Cloud Foundations SRE, G10
SRE Cloud Foundations SRE, G10

TechDigital Group • Georgia

Hybrid
USD 140,000 - 200,000
Site Reliability Engineer
Site Reliability Engineer

Brooksource • Hapeville (GA)

On-site
USD 110,000 - 170,000
SRE - Security Infrastructure - C2C - TS/SCI FSP
SRE - Security Infrastructure - C2C - TS/SCI FSP

Cyrad Solutions LLC • Washington

On-site
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

VantageScore® • San Francisco (CA)

On-site
USD 150,000
Medical insurance
Dental insurance
401(k) plan
+1