Senior Production Engineer – Reliability & Scale

GitHub, Inc.

Northern (KY)

Hybrid

USD 160,000 - 425,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

GitHub, Inc. is building a new Production Engineering organization to elevate reliability and performance. You will embed with product and infrastructure teams, design and operate production software and systems, and mentor engineers to scale operational excellence.

You will own complex incidents end-to-end, improve observability and capacity planning, and automate critical workflows. This role emphasizes impact across GitHub's production platform and developers worldwide.

Qualifications

  • 11+ years' experience in Software Engineering or related field with production software delivery in multiple languages.

Responsibilities

  • Embed with engineering teams to improve reliability, scalability, and operability of production systems.
  • Own complex production problems end-to-end, from debugging incidents to long-term solutions.
  • Design, build, and operate software and infrastructure that improves GitHub production runs.
  • Partner with product and infrastructure teams to improve observability, capacity planning, and resiliency.
  • Write and review production-quality code, automate workflows, and reduce manual work.
  • Participate in incident response and drive learning through reviews and investments.
  • Influence technical direction and mentor engineers in Production Engineering.

Skills

Go
Rust
C++
Java
Python
Distributed systems
Linux
Observability

Education

Associate's Degree in CS/EE/Related
Bachelor's Degree in CS/Related
Master's Degree in CS/Related
Doctorate in CS/Related
Equivalent experience

Tools

Kubernetes
Cloud platforms

Job description

GitHub, Inc. is building a new Production Engineering organization to elevate reliability and performance. You will embed with product and infrastructure teams, design and operate production software and systems, and mentor engineers to scale operational excellence.

You will own complex incidents end-to-end, improve observability and capacity planning, and automate critical workflows. This role emphasizes impact across GitHub's production platform and developers worldwide.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Production Engineer Platform Reliability (Remote)
Senior Production Engineer Platform Reliability (Remote)

GitHub • United States

Remote
USD 255,000 - 425,000
Remote Staff Software Engineer: Planning & Tracking
Remote Staff Software Engineer: Planning & Tracking

GitHub, Inc. • United States

On-site
USD 140,000 - 372,000
Staff Production Engineer - Scalable Data Platform (Remote)
Staff Production Engineer - Scalable Data Platform (Remote)

GitHub • United States

Remote
USD 140,000 - 372,000
Principal Production Engineer
Principal Production Engineer

GitHub, Inc. • Northern (KY)

Hybrid
USD 160,000 - 425,000
Principal Production Engineer, Database Infrastructure
Principal Production Engineer, Database Infrastructure

GitHub • United States

Remote
USD 255,000 - 425,000
Principal Production Engineer, Database Infrastructure
Principal Production Engineer, Database Infrastructure

GitHub, Inc. • Northern (KY)

Hybrid
USD 160,000 - 425,000
Staff Production Engineer, Database Infrastructure
Staff Production Engineer, Database Infrastructure

GitHub • United States

Remote
USD 140,000 - 372,000
Senior Software Engineer
Senior Software Engineer

TechDigital Group • Austin (TX)

On-site
USD 120,000 - 140,000
Senior Production Engineer: Reliability Platform & SRE
Senior Production Engineer: Reliability Platform & SRE

Weights & Biases • Livingston (NJ)

On-site
USD 139,000 - 185,000
Medical, dental, and vision insurance
401(k) with employer match
Paid parental leave
+1
Senior Production Reliability Engineer
Senior Production Reliability Engineer

Nubeero Limited • San Francisco (CA), Northern (KY)

Hybrid
USD 153,000 - 376,000