SRE Manager: Lead Reliability & Observability at Scale

Iac/interactivecorp

Sacramento (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Collaborative work environment
Commitment to carbon‑reduction mission
Flex schedule
Onsite gym & parking
Best-in-class health benefits

Job summary

Shell Recharge Solutions seeks a Manager of Site Reliability Engineering in Los Angeles to lead a world-class SRE team focused on reliability, monitoring and incident response. You will shape architecture and drive observability across products and services.

You will mentor engineers, own on-call rotations, run blameless post-mortems, and apply an everything-as-code approach to ensure fault tolerance. Strong leadership and cloud/container/networking expertise are essential.

Qualifications

  • Bachelor’s degree in CS or engineering required.
  • 5+ years of direct reports management experience.
  • Understanding of golden signals and how to measure them.
  • AWS certification or equivalent cloud experience.

Responsibilities

  • Lead and grow an SRE team to improve monitoring and observability.
  • Design and plan infrastructure for future products and services.
  • Apply everything-as-code to ensure fault tolerance and resilience.
  • Implement Test Driven Development for the SRE org.
  • Collaborate with product teams to define monitoring, logging, and SLOs.
  • Participate in on-call rotations and act as incident commander.
  • Conduct blameless post-mortems to share learnings and improve reliability.
  • Triage incidents and tune resources based on past incidents.

Skills

People management
Golden Signals
Incident management

Education

Bachelor’s Degree in Computer Science or Engineering

Tools

Terraform
CloudFormation
Ansible
Chef
Puppet
Salt
Docker
Podman
Swarm
Kubernetes
Linux

Job description

Shell Recharge Solutions seeks a Manager of Site Reliability Engineering in Los Angeles to lead a world-class SRE team focused on reliability, monitoring and incident response. You will shape architecture and drive observability across products and services.

You will mentor engineers, own on-call rotations, run blameless post-mortems, and apply an everything-as-code approach to ensure fault tolerance. Strong leadership and cloud/container/networking expertise are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Manager – Reliability Leader (Remote)
SRE Manager – Reliability Leader (Remote)

Triwill Group • New York (NY), Northern (KY)

Hybrid
USD 182,000 - 251,000
Health,Dental,Vision insurance
401(k) plan
Equity where applicable
+1
Senior SRE Engineering Manager – Reliability at Scale
Senior SRE Engineering Manager – Reliability at Scale

Upstart • United States

On-site
USD 180,000 - 260,000
Annual equity grants
401(k) retirement match
ESPP (US only)
+4
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Iac/interactivecorp • Sacramento (CA)

On-site
USD 150,000 - 210,000
Collaborative work environment
Commitment to carbon‑reduction mission
Flex schedule
+2
Remote SRE Manager: Lead Reliability & Automation
Remote SRE Manager: Lead Reliability & Automation

NationsBenefits, LLC • Plantation (FL)

Remote
USD 140,000 - 180,000
Fully remote
Unlimited PTO
Competitive compensation
+1
SRE Lead: Reliability, Incident Command & Automation
SRE Lead: Reliability, Incident Command & Automation

Relha LLC • Atlanta (GA), Northern (KY)

Hybrid
USD 112,000 - 131,000
Life insurance
Disability
Parental leave
+5
Senior SRE Leader: Reliability & Observability
Senior SRE Leader: Reliability & Observability

Expedite Talent Solutions • United States

Hybrid
USD 130,000 - 160,000
Senior SRE - Observability & Reliability (Remote)
Senior SRE - Observability & Reliability (Remote)

AppFolio, Inc • Santa Barbara (CA)

Hybrid
USD 138,000 - 173,000
Health insurance
Professional development opportunities
Flexible work environment
Site Reliability Engineering Manager
Site Reliability Engineering Manager

Harvey Nash • Charlotte (NC)

On-site
USD 100,000 - 130,000
SRE Lead: Reliability & Cloud Observability Architect
SRE Lead: Reliability & Cloud Observability Architect

BlackCube Labs • San Diego (CA)

On-site
USD 190,000 - 280,000
Global SRE Manager: Reliability, Automation & Observability
Global SRE Manager: Reliability, Automation & Observability

Harvey Nash • Charlotte (NC)

On-site
USD 100,000 - 130,000