Senior SRE Lead: Reliability, Automation & Observability

Truist

Charlotte (NC)

On-site

USD 140,000 - 190,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision
401k plan
Paid time off

Job summary

Truist is seeking a Site Reliability Engineering Lead to enhance reliability across hybrid cloud and on-prem environments. You will lead major incidents, drive problem management, and implement automation to reduce downtime.

Mentoring, architecture input, and enterprise reliability framework contributions are expected. The role requires 7+ years in SRE/DevOps, distributed systems, Kubernetes, and strong leadership in incident management.

Qualifications

  • Bachelor’s degree in CS, SE, or related field.
  • At least 7 years of professional software development experience.
  • Deep knowledge of programming languages, software architecture, and design.
  • Understanding of SDLC, testing, deployment, and security practices.

Responsibilities

  • Lead major and high-severity incident responses and drive multi-team resolution.
  • Architect scalable, secure, and highly available software solutions.
  • Standardize incident playbooks, escalation paths, and comms frameworks.
  • Develop automation to reduce toil and MTTR across platforms.
  • Mentor engineers and drive design reviews to raise technical maturity.

Skills

Distributed systems
Kubernetes
Automation scripting
Incident management leadership
Python
Go
PowerShell
Ansible
Observability tools
Networking
Linux/Unix internals

Education

Bachelor’s degree in Computer Science, Software Engineering, or related field

Job description

Truist is seeking a Site Reliability Engineering Lead to enhance reliability across hybrid cloud and on-prem environments. You will lead major incidents, drive problem management, and implement automation to reduce downtime.

Mentoring, architecture input, and enterprise reliability framework contributions are expected. The role requires 7+ years in SRE/DevOps, distributed systems, Kubernetes, and strong leadership in incident management.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE Lead: Reliability, Automation & Incidents
Senior SRE Lead: Reliability, Automation & Incidents

Truist • Atlanta (GA)

On-site
USD 150,000 - 190,000
Senior SRE Lead: Reliability & Automation
Senior SRE Lead: Reliability & Automation

Truist • Raleigh (NC)

On-site
USD 130,000 - 180,000
SRE Lead: Reliability, Automation & Incident Leadership
SRE Lead: Reliability, Automation & Incident Leadership

Habitat For Humanity Of Durham • Raleigh (NC)

On-site
USD 120,000 - 180,000
SRE Lead: Reliability, Automation & Incidents
SRE Lead: Reliability, Automation & Incidents

Fayette Chamber of Commerce • Atlanta (GA)

On-site
USD 140,000 - 190,000
Senior SRE - Hybrid, Platform Reliability Lead
Senior SRE - Hybrid, Platform Reliability Lead

TransUnion LLC • Reston (VA)

Hybrid
USD 112,000 - 188,000
Day-one medical, dental, vision
Company-paid basic life/AD&D
12 weeks paid parental leave
+2
Site Reliability Engineering Lead
Site Reliability Engineering Lead

Fayette Chamber of Commerce • Atlanta (GA)

On-site
USD 140,000 - 190,000
Site Reliability Engineering Lead
Site Reliability Engineering Lead

Habitat For Humanity Of Durham • Raleigh (NC)

On-site
USD 120,000 - 180,000
Senior SRE Lead: Reliability, Observability & Automation
Senior SRE Lead: Reliability, Observability & Automation

Bank of America • Chandler (AZ)

On-site
USD 180,000 - 240,000
Senior SRE Leader: Hybrid, GCP & Kubernetes Expert
Senior SRE Leader: Hybrid, GCP & Kubernetes Expert

TransUnion LLC • Crum Lynne (PA)

Hybrid
USD 112,000 - 188,000
Health insurance
401(k) with match
Employee stock purchase plan
+3
Senior SRE Lead: Reliability, Architecture & AI-Driven Ops
Senior SRE Lead: Reliability, Architecture & AI-Driven Ops

Fairygodboss • Jersey City (NJ)

On-site
USD 150,000 - 210,000