Engineering – SRE Platforms – Software Engineer – Vice President – Dallas | Dallas, TX, USA

Goldman Sachs, Inc.

Dallas (TX)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Goldman Sachs in Dallas is seeking a talented Site Reliability Engineering Manager to lead a team of SRE engineers responsible for the reliability, availability, and performance of critical applications and infrastructure.

You will develop and implement best practices for incident management, monitoring, automation, and capacity planning, collaborating with development and operations teams to design highly available systems.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field.
  • 8+ years of experience in Site Reliability Engineering, with at least 3 years in a management role.
  • Strong leadership and people management capabilities.
  • Experience with cloud platforms such as AWS or Azure.
  • Experience with infrastructure as code tools such as Terraform or CloudFormation.
  • Experience with containerization technologies such as Docker and Kubernetes.
  • Strong problem-solving skills.

Responsibilities

  • Manage a team of Site Reliability Engineers responsible for ensuring the reliability, availability, and performance of critical applications and infrastructure.
  • Develop and implement best practices for Site Reliability Engineering, including incident management, monitoring, automation, and capacity planning.
  • Collaborate with development teams to design and build highly available and scalable systems.
  • Work with infrastructure teams to ensure that critical infrastructure components are operating optimally and are able to support the needs of the business.
  • Develop and maintain Service Level Agreements (SLAs) and Service Level Objectives (SLOs) to ensure that critical systems meet the needs of the business.
  • Manage and prioritize workload for the SRE team, ensuring alignment with business priorities.
  • Develop and maintain relationships with key stakeholders across the organization to ensure alignment of the SRE function with business goals.

Skills

Leadership
Team management
Incident management
Cloud platforms (AWS/Azure)
Infrastructure as code (Terraform/CFN)
Containerization (Docker/Kubernetes)
Problem solving

Education

Bachelor's degree in Computer Science, Engineering, or related field

Tools

AWS
Azure
Terraform
CloudFormation
Docker
Kubernetes

Job description

Engineering - SRE Platforms - Software Engineer - Vice President - Dallas Job Description Goldman Sachs is seeking a talented and motivated Site Reliability Engineering Manager to join our team. As a leader within the firm's Technology division, you will be responsible for overseeing the Site Reliability Engineering (SRE) function, ensuring the stability and reliability of critical applications and infrastructure. You will manage a team of SRE engineers who work closely with developers, infrastructure engineers, and operations teams to build and maintain highly available systems.

Key Responsibilities
  • Manage a team of Site Reliability Engineers responsible for ensuring the reliability, availability, and performance of critical applications and infrastructure
  • Develop and implement best practices for Site Reliability Engineering, including incident management, monitoring, automation, and capacity planning
  • Collaborate with development teams to design and build highly available and scalable systems
  • Work with infrastructure teams to ensure that critical infrastructure components are operating optimally and are able to support the needs of the business
  • Develop and maintain Service Level Agreements (SLAs) and Service Level Objectives (SLOs) to ensure that critical systems meet the needs of the business
  • Manage and prioritize workload for the SRE team, ensuring that they are aligned with business priorities
  • Develop and maintain relationships with key stakeholders across the organization to ensure that the SRE function is aligned with business goals
Qualifications
  • Bachelor's degree in Computer Science, Engineering, or related field
  • 8+ years of experience in Site Reliability Engineering, with at least 3 years in a management role
  • Strong leadership skills with the ability to manage a team of engineers
  • Experience with cloud computing platforms such as AWS or Azure
  • Experience with infrastructure as code (IaC) tools such as Terraform or CloudFormation
  • Experience with containerization technologies such as Docker and Kubernetes
  • Strong problem-solving skills with the ability to troubleshoot complex issues
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas
Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas

The Goldman Sachs Group • Dallas (TX)

On-site
USD 180,000 - 280,000
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas

Goldman Sachs • Dallas (TX)

On-site
USD 120,000 - 160,000
None
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas

The Goldman Sachs Group • Dallas (TX)

On-site
USD 110,000 - 140,000
VP of SRE Platforms and Reliability
VP of SRE Platforms and Reliability

Goldman Sachs, Inc. • Dallas (TX)

On-site
USD 180,000 - 240,000
VP, SRE Platforms — Scale, Reliability & Automation
VP, SRE Platforms — Scale, Reliability & Automation

The Goldman Sachs Group • Dallas (TX)

On-site
USD 180,000 - 280,000
SRE - Site Reliability Engineer - Senior
SRE - Site Reliability Engineer - Senior

ManpowerGroup Global, Inc. • Austin (TX)

On-site
USD 66,000 - 90,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

State of Wisconsin Investment Board • Madison (WI)

On-site
USD 140,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • Austin (TX)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

On-site
USD 140,000 - 200,000
Site Reliability Engineer -Jersey City, NJ & Dallas, TX
Site Reliability Engineer -Jersey City, NJ & Dallas, TX

Stradit LLC • Dallas (TX), Northern (KY)

On-site
USD 140,000 - 190,000