Site Reliability Engineer - Scale & Observability

gamma.app

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible work-from-home options
Creative and collaborative team environment

Job summary

A dynamic tech firm located in San Francisco is seeking a Site Reliability Engineer to enhance operational health across their production systems. This high-impact role demands expertise in AWS and strong programming skills. You will manage production systems' reliability and lead incident response efforts to prevent issues, all while contributing to the scalability and efficiency of their services. Ideal candidates will have 5+ years of relevant experience and a passion for leveraging technology to drive outcomes.

Qualifications

  • 5+ years in Site Reliability Engineering, DevOps, or systems engineering roles.
  • Strong programming skills in Python, Go, or TypeScript/Node.js.
  • Experience with infrastructure-as-code and observability solutions.

Responsibilities

  • Own reliability, availability, and performance of Gamma's production systems.
  • Build observability infrastructure with metrics and logging.
  • Lead incident response and drive systemic improvements.

Skills

Site Reliability Engineering
AWS expertise
Python
Go
TypeScript/Node.js
Terraform
CloudFormation
Docker
Kubernetes

Education

Bachelor's degree in relevant field

Tools

CloudFormation
Docker
Kubernetes
Kafka

Job description

A dynamic tech firm located in San Francisco is seeking a Site Reliability Engineer to enhance operational health across their production systems. This high-impact role demands expertise in AWS and strong programming skills. You will manage production systems' reliability and lead incident response efforts to prevent issues, all while contributing to the scalability and efficiency of their services. Ideal candidates will have 5+ years of relevant experience and a passion for leveraging technology to drive outcomes.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: Scale Reliability, Observability & CI/CD
Senior SRE: Scale Reliability, Observability & CI/CD

Breakout Tools • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer - Scale Resilient Systems
Senior Site Reliability Engineer - Scale Resilient Systems

Jobgether • United States

Remote
USD 150,000 - 200,000
Site Reliability Engineer — Scale & Resilience for AI Ops
Site Reliability Engineer — Scale & Resilience for AI Ops

HappyRobot • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer — Cloud, Resilience & Automation
Senior Site Reliability Engineer — Cloud, Resilience & Automation

Compunnel Inc. • Denton (TX)

Hybrid
USD 120,000 - 150,000
Site Reliability Engineer — Scale, Observability & Automation
Site Reliability Engineer — Scale, Observability & Automation

airapps • San Francisco (CA)

On-site
USD 130,000 - 160,000
Apple hardware ecosystem for work
Annual Bonus
Medical Insurance
+3
Site Reliability Engineer — Cloud-Scale & Automation
Site Reliability Engineer — Cloud-Scale & Automation

ByteDance • San Jose (CA)

On-site
USD 136,000 - 360,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+3
Senior SRE: Scale, Reliability & Observability Leader
Senior SRE: Scale, Reliability & Observability Leader

Brez Technology Inc. • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Private Medical, Dental and Vision Benefits
Retirement Savings plan with matching contributions
Workspace benefits for your home office
+4
Senior SRE & Software Engineer — Scalable Infra
Senior SRE & Software Engineer — Scalable Infra

Harvey • San Francisco (CA)

On-site
USD 200,000 - 260,000
Senior Site Reliability Engineer: Observability & Cloud
Senior Site Reliability Engineer: Observability & Cloud

VBeyond Corporation • Jersey City (NJ)

On-site
USD 100,000 - 260,000
SRE: Scale, Uptime & Observability (Remote)
SRE: Scale, Uptime & Observability (Remote)

WorkOS • Denver (NC)

Remote
USD 175,000 - 250,000