Site Reliability Engineering Manager

Confidential

United States

On-site

USD 190,000 - 920,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading company is seeking a Site Reliability Engineer Practice Leader to drive sales and revenue growth by managing a global team and developing an SRE practice. This role entails overseeing cloud operations, automation initiatives, and fostering innovation in site reliability engineering, providing a significant opportunity for career growth.

Qualifications

  • 7+ years in SRE, DevOps, or infrastructure roles, with at least 3 years in a leadership capacity.
  • Strong knowledge of cloud platforms (AWS, Azure, GCP).
  • Proficiency in scripting languages (Python, Bash, Go, etc.).

Responsibilities

  • Develop and implement an SRE practice aligned with business and customer engineering goals.
  • Lead and mentor a team of SREs, fostering a culture of reliability and automation.
  • Define and drive best practices for site reliability, observability, and incident management.

Skills

Leadership
Problem-solving
Cross-functional collaboration
Cloud platforms
Automation
Incident management

Education

Bachelor’s or Master’s in Computer Science, Engineering, or a related field

Tools

Kubernetes
Docker
Terraform
AWS CloudFormation
CI/CD

Job description

Job Title: Site Reliability Engineer (SRE) Practice Leader

About the Role:

We're is seeking a dynamic Practice Leader for our Site Reliability Engineer (SRE) team. This role is pivotal in driving sales and revenue growth by leading enterprise customers through transformative projects. The ideal candidate will have a proven track record in IT Strategy, consulting, and hybrid cloud operations, all with a strong emphasis on sales and business development.

This role brings an exciting opportunity to run and grow our growing SRE Practice. The Practice Leader runs a global operations team of SRE Engineers at many different experience levels. We are looking for someone who not only has worked with offshore operations but is able to speak to technologies in the AWS, Azure and OCI cloud. You will be working with multiple cloud partners such as AWS, Azure, Google Cloud, and Oracle Cloud.

This position is a great career opportunity where you get to run your practice, work with other technology leaders and grow your base business.

Key Responsibilities:

  • Develop and implement an SRE practice aligned with business and customer engineering goals.
  • Lead and mentor a team of SREs, fostering a culture of reliability, automation, and continuous improvement.
  • Define and drive best practices for site reliability, observability, and incident management.
  • Partner with product and engineering teams to ensure scalable and highly available systems.
  • Collaborate with sales to develop offerings that resonate with the market

Customer Reliability & Performance Optimization:

  • Develop incident response strategies, including on-call rotations and post-mortem reviews.
  • Improve system observability through logging, monitoring, and alerting strategies.
  • Lead efforts in performance tuning, capacity planning, and load balancing.

Automation & DevOps Enablement:

  • Drive automation efforts for deployments, scaling, and self-healing infrastructure.
  • Implement CI/CD pipelines to ensure smooth and reliable software releases.
  • Work with DevOps teams to optimize cloud and on-prem infrastructure for reliability.
  • Ensure security and compliance standards are met across infrastructure and operations.

Qualifications & Experience:

Education: Bachelor’s or Master’s in Computer Science, Engineering, or a related field.

Experience: 7+ years in SRE, DevOps, or infrastructure roles, with at least 3 years in a leadership or managerial capacity.

Technical Expertise:

  • Strong knowledge of cloud platforms (AWS, Azure, GCP).
  • Experience with Kubernetes, Docker, and microservices architecture.
  • Proficiency in scripting languages (Python, Bash, Go, etc.).
  • Deep understanding of networking, security, and performance optimization.
  • Expert knowledge of Terraform, AWS CloudFormation or Pulumi
  • Experience with CO/CD pipelines and incident management

Soft Skills: Strong leadership, problem-solving, and cross-functional collaboration abilities.

Why Join Us?

Opportunity to build and shape an SRE practice.

Work with cutting-edge cloud and automation technologies.

Seniority level
  • Seniority level
    Not Applicable
Employment type
  • Employment type
    Full-time
Job function
  • Industries
    IT System Data Services

Referrals increase your chances of interviewing at Confidential by 2x

Get notified about new Site Reliability Engineering Manager jobs in United States.

Engineering Manager, Page Scale & Performance Team

United States $190,000.00-$920,000.00 1 week ago

Engineering Manager - Growth Experiences

Washington, DC $190,000.00-$920,000.00 1 week ago

Engineering Manager (Remote - Multiple locations)
Site Reliability Engineering (SRE) Manager, 1LMX MES COE

United States $216,750.00-$351,900.00 1 day ago

Engineering Manager (Consumer - Trading)

United States $218,025.00-$256,500.00 1 day ago

United States $181,000.00-$237,000.00 4 days ago

Engineering Manager, Detection Engineering

United States $149,450.00-$242,765.00 1 week ago

United States $180,000.00-$210,000.00 2 months ago

Senior Engineering Manager- Site Reliability
Engineering Manager / Head of Engineering - Dragonfly Portfolio

New York City Metropolitan Area 2 days ago

United States $140,000.00-$180,000.00 2 days ago

Senior Manager, Engineering - Auth Infrastructure (Core Services)

United States $225,000.00-$250,000.00 3 weeks ago

Engineering Manager, Rider Core Experience

United States $184,000.00-$253,000.00 2 weeks ago

We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Interactive Resources - iR • Austin (TX)

On-site
USD 126,000 - 223,000
Medical insurance
Vision insurance
401(k)
Site Reliability Engineering Manager
Site Reliability Engineering Manager

LexisNexis Risk Solutions • Alpharetta (GA)

On-site
USD 120,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Signature IT World Inc • Washington

On-site
USD 120,000 - 200,000
Medical insurance
Vision insurance
Child care support
+3
Site Reliability Engineer
Site Reliability Engineer

Quantitative Systems • New York (NY)

On-site
USD 250,000 - 300,000
Full medical, dental, and vision for you and your dependents
401(k) match
Unlimited sick days
+3
Site Reliability Engineer
Site Reliability Engineer

Motion Recruitment • Atlanta (GA)

On-site
USD 93,000 - 158,000
Experienced SRE | Diversified Strategies Hedge Fund
Experienced SRE | Diversified Strategies Hedge Fund

Techfellow Limited • New York (NY)

Hybrid
USD 117,000 - 250,000
SRE Engineer - Production support (Local to Washington))
SRE Engineer - Production support (Local to Washington))

Ampstek • Bellevue (WA)

On-site
USD 147,000 - 208,000
Site Reliability Engineer (10+ Years) (Need Local only)
Site Reliability Engineer (10+ Years) (Need Local only)

Ampstek • Bellevue (WA)

On-site
USD 130,000 - 150,000
Sr. Site Reliability Engineer (Python)
Sr. Site Reliability Engineer (Python)

Ledgent Technology • Irvine (CA)

On-site
USD 140,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Insight Global • United States

On-site
Medical insurance
Vision insurance
401(k)