Engineering - SRE Platforms - SRE Engineer - Associate - Dallas

Goldman Sachs

Dallas (WV)

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Goldman Sachs seeks a Site Reliability Engineer to join our global teams responsible for the availability and reliability of critical platform services. You will help build, run, and evolve highly distributed systems, and operate observability platforms used across the firm.

You will collaborate with engineering and business units to meet stringent SLOs, automate 운영, and support incident response while driving continuous improvements in reliability and performance.

Qualifications

  • Minimum of 2+ years of hands-on experience in Site Reliability Engineering, with a proven track record in building and maintaining highly available, scalable, fault-tolerant systems at an enterprise level.
  • BS degree in Computer Science or related technical field involving coding and/or systems engineering.
  • Proficiency in one or more: Go, Python, C, C++, Java, Perl, Ruby or shell scripting.
  • Experience with UNIX operating systems internals and/or networking.

Responsibilities

  • Balance feature development velocity and reliability with well-defined SLOs.
  • Run the Production environment by monitoring availability and system health.
  • Drive incident management process and support blameless post-mortems culture.
  • Partner with development teams to improve services via testing and release procedures.
  • Participate in system design consulting, platform management, and capacity planning.
  • Create sustainable systems and services through automation.

Skills

SRE fundamentals
Go
Python
C
Java

Education

BS in Computer Science or related technical field

Tools

UNIX/Linux
Networking design
Cloud platforms

Job description

Site Reliability Engineering (SRE) is an engineering discipline that combines software development and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. At Goldman Sachs, SRE is responsible for the availability and reliability of our firm's most critical platform services and ensures they meet the requirements of our internal and external users. We also develop and operate the observability platforms that all other engineering teams use to make their services reliable. We look for engineers who are motivated to collaborate with other engineering teams and our businesses to build and run sustainable production systems, which can evolve and adapt to changes in our fast-paced, global business and regulatory environment.

How will you fulfil your potential?
  • Balance feature development velocity and reliability with well-defined SLOs.
  • Run the Production environment by monitoring availability and taking a holistic view of system health.
  • Drive incident management process and support a blameless post-mortems culture.
  • Partner with development teams to improve services via rigorous testing and release procedures.
  • Participate in system design consulting, platform management, and capacity planning.
  • Create sustainable systems and services through automation and uplifts.
  • Champion reliability and resilience engineering practices and knowledge across the firm.
Basic Qualifications
  • Minimum of 2+ years of hands-on experience in Site Reliability Engineering, with a proven track record in building, and maintaining highly available, scalable, and fault-tolerant systems at an enterprise level.
  • BS degree in Computer Science or related technical field involving coding and / or systems engineering.
  • Proficiency in one or more of the following: Go, Python, C, C++, Java, Perl, Ruby or shell scripting.
  • Experience with product engineering practices, algorithms, data structures, software design and/or Experience with UNIX operating systems internals and / or networking.
Preferred Qualifications
  • Experience in developing AI tools and working with Cloud Platforms
  • Experience with distributed systems design, maintenance, and troubleshooting.
  • Hands-on experience with debugging and optimizing code, as well as automation.
  • Strong interpersonal skills, drive, and ownership.
  • Coding beyond simple scripts.
  • Solving novel problems from first principles.
  • Experience working in highly regulated, financial services firms.
  • Excellent people leadership skills either as an engineering manager or individual contributor
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering - SRE Platforms - SRE Engineer - Associate - Dallas
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas

Goldman Sachs • Dallas (TX)

On-site
USD 120,000 - 160,000
None
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas
Engineering - SRE Platforms - SRE Engineer - Associate - Dallas

The Goldman Sachs Group • Dallas (TX)

On-site
USD 110,000 - 140,000
Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas
Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas

The Goldman Sachs Group • Dallas (TX)

On-site
USD 180,000 - 280,000
Engineering – SRE Platforms – Software Engineer – Vice President – Dallas | Dallas, TX, USA
Engineering – SRE Platforms – Software Engineer – Vice President – Dallas | Dallas, TX, USA

Goldman Sachs, Inc. • Dallas (TX)

On-site
USD 210,000 - 260,000
Site Reliability Engineering (SRE), The Core Engineering, Vice President, Dallas
Site Reliability Engineering (SRE), The Core Engineering, Vice President, Dallas

Goldman Sachs • Dallas (TX)

On-site
USD 120,000 - 160,000
Senior SRE, Compliance Engineering & DevOps
Senior SRE, Compliance Engineering & DevOps

Goldman Sachs • Dallas (TX)

On-site
USD 120,000 - 160,000
Site Reliability Engineer, Global Banking & Markets, Vice President
Site Reliability Engineer, Global Banking & Markets, Vice President

Goldman Sachs • New York (NY)

On-site
USD 150,000 - 250,000
Site Reliability Engineer, Global Banking & Markets, Vice President
Site Reliability Engineer, Global Banking & Markets, Vice President

Socket.dev • New York (NY)

On-site
USD 150,000 - 300,000
Site Reliability Engineer, Global Banking & Markets, Frontline Production Engineering
Site Reliability Engineer, Global Banking & Markets, Frontline Production Engineering

Goldman Sachs • New York (NY)

On-site
USD 130,000 - 250,000
VP, SRE Platforms & Reliability
VP, SRE Platforms & Reliability

Goldman Sachs, Inc. • Dallas (TX)

On-site
USD 210,000 - 260,000