SRE Enablement Engineer: Reliability & Observability Lead

AutoZone

Memphis (TN)

Hybrid

USD 120,000 - 160,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Hybrid work model

Job summary

AutoZone's Site Reliability Engineering team seeks a Systems Engineer focused on SRE Enablement to promote reliability across the engineering organization. The role emphasizes standards, shared tools, and guidance for teams, with a hybrid GCP and on‑prem footprint.

The successful candidate will collaborate with application, infrastructure, and architecture teams to embed SRE practices early in development and ensure production readiness. Strong communication and consulting skills are essential.

Qualifications

  • Bachelor's degree in computer science, MIS, information technology, or related field, or equivalent practical experience.
  • 4 to 7 years of experience in Systems Engineering, DevOps, or SRE-related roles.
  • Strong understanding of SRE principles, including SLOs, SLIs, error budgets, and TOIL reduction.
  • Hands-on experience building, administering, and optimizing observability and APM pipelines with Dynatrace.
  • Experience deploying and supporting workloads in Google Cloud Platform (GCP) and on-premises environments.
  • Experience with container orchestration platforms (Kubernetes) and IaC tools (Terraform, Ansible).
  • Excellent communication and consulting skills to influence architecture decisions.

Responsibilities

  • Define enterprise-wide reliability standards, SLO frameworks, and error budget policies.
  • Establish production readiness criteria and conduct readiness reviews across teams.
  • Own and maintain the internal SRE handbook and reliability playbooks.
  • Build, maintain, and standardize shared observability platforms, leveraging Dynatrace.
  • Provide templates for alerting, dashboards, and runbooks for cloud and on-prem workloads.
  • Participate in incident management, including post-mortems to improve reliability.
  • Run SRE training programs and reliability workshops for engineering teams.
  • Coach teams on SLO-based thinking and error budget management.
  • Embed proactive SRE practices and a continuous improvement mindset.
  • Track and report reliability metrics across the enterprise.
  • Identify systemic reliability gaps and report health to leadership.
  • Act as an internal consultant during architecture and system design reviews.
  • Advise on reliability design patterns for hybrid GCP/on-prem environments.
  • Engage early in new product development to influence system reliability.

Skills

SRE principles
Dynatrace
GCP
Kubernetes
Python
Golang
Java
Observability
Incident management
SLOs/SLIs/Error budgets
TOIL reduction
IaC (Terraform/Ansible)

Education

Bachelor's degree in CS / IT / MIS or related field

Tools

Kubernetes
Terraform
Ansible
Dynatrace

Job description

AutoZone's Site Reliability Engineering team seeks a Systems Engineer focused on SRE Enablement to promote reliability across the engineering organization. The role emphasizes standards, shared tools, and guidance for teams, with a hybrid GCP and on‑prem footprint.

The successful candidate will collaborate with application, infrastructure, and architecture teams to embed SRE practices early in development and ensure production readiness. Strong communication and consulting skills are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Engineer – SRE Enablement
Systems Engineer – SRE Enablement

AutoZone • Memphis (TN)

Hybrid
USD 120,000 - 160,000
Hybrid work model
Platform SRE Engineer for Scalable, Observable Systems
Platform SRE Engineer for Scalable, Observable Systems

AutoZone • Memphis (TN)

On-site
USD 110,000 - 150,000
Onsite Senior SRE: Reliability & Observability Leader
Onsite Senior SRE: Reliability & Observability Leader

O'Reilly Auto Parts • Springfield (MO)

On-site
USD 120,000 - 180,000
Competitive wages
401k with employer contributions
Medical, Dental, Vision insurance
+3
SRE Manager: Lead Reliability & Observability at Scale
SRE Manager: Lead Reliability & Observability at Scale

Iac/interactivecorp • Sacramento (CA)

On-site
USD 150,000 - 210,000
Collaborative work environment
Commitment to carbon‑reduction mission
Flex schedule
+2
Senior Site Reliability Engineer: Build Resilient Systems
Senior Site Reliability Engineer: Build Resilient Systems

IRB USA Inspire Resources • Atlanta (GA)

On-site
USD 130,000 - 180,000
Senior SRE: Reliability & Observability Lead
Senior SRE: Reliability & Observability Lead

Inspire • Atlanta (GA)

On-site
USD 140,000 - 200,000
Senior Site Reliability Engineer – Scale & Observability
Senior Site Reliability Engineer – Scale & Observability

Inspire Brands, Inc. • Atlanta (GA)

On-site
USD 120,000 - 180,000
Global SRE Manager: Reliability, Automation & Observability
Global SRE Manager: Reliability, Automation & Observability

Harvey Nash • Charlotte (NC)

On-site
USD 100,000 - 130,000
Senior SRE Leader: Reliability & Observability
Senior SRE Leader: Reliability & Observability

Expedite Talent Solutions • United States

Hybrid
USD 130,000 - 160,000
SRE Manager – Reliability Leader (Remote)
SRE Manager – Reliability Leader (Remote)

Triwill Group • New York (NY), Northern (KY)

Hybrid
USD 182,000 - 251,000
Health,Dental,Vision insurance
401(k) plan
Equity where applicable
+1