Site Reliability Engineer

Gamma

San Francisco (CA)

On-site

USD 230,000 - 310,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A tech company is seeking an experienced Site Reliability Engineer to ensure the reliability and performance of its production systems across AWS infrastructure. You will build observability tools, lead incident responses, and collaborate on architectural improvements. The ideal candidate should have over 5 years of hands-on AWS experience and proficiency in programming languages such as Python and Go. The role offers a competitive salary range of $230K - $310K plus benefits and equity with a commitment to applicant growth.

Qualifications

  • 5+ years in site reliability engineering, DevOps, or systems engineering with AWS expertise.
  • Strong programming skills in Python, Go, or TypeScript/Node.js.
  • Experience with infrastructure-as-code and observability solutions.

Responsibilities

  • Own the reliability and performance of production systems across AWS.
  • Build observability infrastructure including metrics, logging, and alerting.
  • Lead incident response and systemic fixes.

Skills

AWS expertise
Python
Go
TypeScript/Node.js
Infrastructure-as-code (Terraform, CloudFormation)
Containerization (Docker, Kubernetes)
Incident management
Distributed systems

Job description

About the role

Gamma's infrastructure needs to be rock-solid for millions of daily users while enabling our engineering teams to ship fast. You'll own the operational health of our full backend platform, building automation and tooling that improves reliability and partnering with engineering to design systems that are observable, resilient, and easy to operate. Your work directly impacts every Gamma user's experience.

This is a high-impact role where you'll balance reliability with velocity, knowing when to move fast and when to prioritize stability. You'll lead incident response, drive systemic improvements, and help shape how Gamma scales to serve its next 100 million users.

Our team has a strong in-office culture and works in person 4–5 days per week in San Francisco. We love working together to stay creative and connected, with flexibility to work from home when focus matters most.

What you'll do
  • Own the reliability, availability, and performance of Gamma's production systems across our AWS infrastructure

  • Build observability infrastructure from the ground up: metrics, logging, tracing, and alerting that give the team genuine visibility into system health before users feel the impact

  • Design and ship automation that reduces toil, makes deployments safer, and gets us back on our feet faster when things go wrong

  • Lead incident response and blameless post-mortems, then follow through on the systemic fixes that keep the same issues from coming back

  • Partner with engineering teams on architecture reviews, SLO and SLI design, and reliability best practices that scale with the product

  • Manage and optimize our compute, networking, databases, and managed services

What you'll bring
  • 5+ years in site reliability engineering, DevOps, or systems engineering with deep, hands-on AWS expertise

  • Strong programming skills in Python, Go, or TypeScript/Node.js, applied to building real tools and automation

  • Solid experience with infrastructure-as-code (Terraform, CloudFormation) and end-to-end observability solutions

  • Track record of making systems meaningfully more reliable through automation, smarter monitoring, and architectural improvements

  • Deep understanding of networking, distributed systems, containerization (Docker, Kubernetes), and database performance at scale

  • Sharp incident management instincts and the debugging skills to navigate complex production failures

  • AWS certifications, or experience with security and compliance frameworks like SOC 2 or ISO 27001 (Nice to have)

  • Experience scaling SaaS products to millions of users, or background with Kafka, chaos engineering, or service mesh technologies (Nice to have)

Compensation range:

The base salary for this full-time position, which spans multiple internal levels depending on qualifications, ranges between $230K - $310K plus benefits & equity.

Final offer amounts are determined by multiple factors, including but not limited to experience and expertise in the requirements listed above.

If you're interested in this role but you don't meet every requirement, we encourage you to apply anyway! We're always excited about meeting great people.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

gamma capital markets • San Francisco (CA)

On-site
USD 230,000 - 310,000
Benefits
Equity
Site Reliability Engineer
Site Reliability Engineer

gamma.app • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible work-from-home options
Creative and collaborative team environment
Software Engineer, Platform
Software Engineer, Platform

Gamma • San Francisco (CA)

On-site
USD 180,000 - 275,000
Security Engineer: Cloud Security & Automation
Security Engineer: Cloud Security & Automation

Gamma • San Francisco (CA)

On-site
USD 180,000 - 310,000
Security Engineer
Security Engineer

gamma capital markets • San Francisco (CA)

On-site
USD 180,000 - 310,000
Software Engineer, Distributed Systems - Infra
Software Engineer, Distributed Systems - Infra

Gamma • San Francisco (CA)

On-site
USD 180,000 - 275,000
Software Engineer, Distributed Systems - Infra
Software Engineer, Distributed Systems - Infra

gamma capital markets • San Francisco (CA)

On-site
USD 180,000 - 310,000
Senior Site Reliability Engineer: Scale & Resilience
Senior Site Reliability Engineer: Scale & Resilience

gamma capital markets • San Francisco (CA)

On-site
USD 230,000 - 310,000
Benefits
Equity
Software Engineer, Backend
Software Engineer, Backend

Gamma • San Francisco (CA)

On-site
USD 180,000 - 275,000
Senior IT Engineer
Senior IT Engineer

gamma capital markets • San Francisco (CA)

On-site
USD 140,000 - 190,000