Site Reliability Engineer

gamma.app

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible work-from-home options
Creative and collaborative team environment

Job summary

A dynamic tech firm located in San Francisco is seeking a Site Reliability Engineer to enhance operational health across their production systems. This high-impact role demands expertise in AWS and strong programming skills. You will manage production systems' reliability and lead incident response efforts to prevent issues, all while contributing to the scalability and efficiency of their services. Ideal candidates will have 5+ years of relevant experience and a passion for leveraging technology to drive outcomes.

Qualifications

  • 5+ years in Site Reliability Engineering, DevOps, or systems engineering roles.
  • Strong programming skills in Python, Go, or TypeScript/Node.js.
  • Experience with infrastructure-as-code and observability solutions.

Responsibilities

  • Own reliability, availability, and performance of Gamma's production systems.
  • Build observability infrastructure with metrics and logging.
  • Lead incident response and drive systemic improvements.

Skills

Site Reliability Engineering
AWS expertise
Python
Go
TypeScript/Node.js
Terraform
CloudFormation
Docker
Kubernetes

Education

Bachelor's degree in relevant field

Tools

CloudFormation
Docker
Kubernetes
Kafka

Job description

We're building the creative layer for modern communication. Every month, over a billion people make presentations — but the tools they use to make them haven't evolved in decades. We're changing that, using AI to disrupt a massive market.

📈 Millions of people rely on Gamma to create, teach, and persuade, creating more than 1 million gammas every day.

💻 We see Gamma as the next great workplace tool, combining viral B2C love with a massive B2B opportunity. We believe AI can be a true creative partner: one that understands context, clarity, and taste.

💸 We’ve reached a $2.1B valuation, crossed $100M in annual recurring revenue, and have been profitable since 2023.

💙 We're an imaginative, passionate team who takes our work seriously, but not ourselves. Our culture is warm, a little quirky, and fueled by curiosity.

About the role

Gamma's infrastructure needs to be rock-solid for millions of daily users while enabling our engineering teams to ship fast. You'll own the operational health of our full backend platform, building automation and tooling that improves reliability and partnering with engineering to design systems that are observable, resilient, and easy to operate. Your work directly impacts every Gamma user's experience.

This is a high-impact role where you'll balance reliability with velocity, knowing when to move fast and when to prioritize stability. You'll lead incident response, drive systemic improvements, and help shape how Gamma scales to serve its next 100 million users.

Our team has a strong in-office culture and works in person 4–5 days per week in San Francisco. We love working together to stay creative and connected, with flexibility to work from home when focus matters most.

What you'll do
  • Own reliability, availability, and performance of Gamma's production systems across primarily AWS infrastructure
  • Build observability infrastructure with metrics, logging, tracing, and alerting that provide deep visibility into system health
  • Design automation to reduce toil, improve deployment safety, and accelerate incident resolution
  • Lead incident response, conduct blameless post-mortems, and drive systemic improvements to prevent recurring issues
  • Partner with engineering teams on architecture reviews, SLOs/SLIs, and reliability best practices
  • Manage and optimize our infrastructure including compute, networking, databases, and managed services
What you'll bring
  • 5+ years in Site Reliability Engineering, DevOps, or systems engineering roles with deep AWS expertise
  • Strong programming skills (Python, Go, or TypeScript/Node.js) for building tools and automation
  • Experience with infrastructure-as-code (Terraform, CloudFormation) and comprehensive observability solutions
  • Track record improving system reliability through automation, monitoring, and architectural improvements
  • Solid understanding of networking, distributed systems, containerization (Docker, Kubernetes), and database performance
  • Strong incident management and debugging skills for complex production issues
  • (Nice to have) Experience scaling SaaS applications to millions of users
  • (Nice to have) Background with real-time collaborative systems, Kafka, chaos engineering, or service mesh technologies
  • (Nice to have) AWS certifications or experience with security/compliance requirements (SOC 2, ISO 27001)
Compensation range

Final offer amounts are determined by multiple factors, including but not limited to experience and expertise in the requirements listed above.

If you're interested in this role but you don't meet every requirement, we encourage you to apply anyway! We're always excited about meeting great people.

We're building on a full Typescript stack centered around some of the most modern and popular technologies.

We use our own custom, open-source AI prompting framework, AIJSX. We have a lot of custom tools built in-house, but also new ones like Vercel AI SDK.

Our tiny team operates at massive scale:

1M+

70M users around the world

6M+ AI images generated daily

1 trillion LLM tokens processed per month

Life at Gamma

You get energy from small teams doing big things.

You love when design, code, and storytelling overlap.

You default to action, even when the answer isn’t clear yet.

You value details, but know when to ship and move on.

You bring both the spreadsheets and the sparkle, equal parts workhorse and unicorn.

You believe AI should amplify creativity, not replace it.

You know kindness and intensity are not opposites.

You like working with people who care deeply: about their craft, their teammates, and the users on the other side of the screen.

Who we are

Gamma is full of imaginative, passionate people who take their work seriously but not themselves. The culture is warm, a little quirky, and fueled by curiosity. It’s the kind of place where you’ll debate a pixel on Monday, laugh over someone’s keyboard setup on Tuesday, and ship something remarkable by Friday.

We care about craft, move with intention, and don’t mind getting a little scrappy. It’s fast, creative, and occasionally chaotic — but that’s what makes it interesting.

Here’s a bit about what it’s like to work here, from people on the inside:

“quirky, inspiring, fun, a little wild in the best way”

“You can have an idea and just run with it.”

“Everyone’s talented and humble — the mix keeps you sharp.”

“We ship cool stuff, learn a ton, and laugh a lot doing it.”

Meet the team

We're a team of dreamers and doers building in beautiful San Francisco 🌉

We're kabbadi enthusiasts, pickleballers, dog herders, woodworkers, keyboard nerds, potters, and more — and we can't wait to meet you!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Distributed Systems (Infra)
Software Engineer, Distributed Systems (Infra)

gamma.app • San Francisco (CA)

Hybrid
USD 180,000 - 310,000
Software Engineer, Platform
Software Engineer, Platform

gamma.app • San Francisco (CA)

On-site
USD 120,000 - 160,000
Collaborative team environment
Creative workspace
Software Engineer, Backend
Software Engineer, Backend

gamma.app • San Francisco (CA)

On-site
USD 120,000 - 150,000
Software Engineer, Data Systems
Software Engineer, Data Systems

gamma.app • San Francisco (CA)

On-site
USD 230,000 - 310,000
Software Engineer, Trust & Safety
Software Engineer, Trust & Safety

gamma.app • San Francisco (CA)

On-site
USD 180,000 - 310,000
Engineering Manager, Growth
Engineering Manager, Growth

gamma.app • San Francisco (CA)

Hybrid
USD 250,000 - 325,000
UI Engineer
UI Engineer

gamma.app • San Francisco (CA)

On-site
USD 120,000 - 160,000
Software Engineer, Frontend
Software Engineer, Frontend

gamma.app • San Francisco (CA)

Hybrid
USD 150,000 - 180,000
In-office in San Francisco
Hybrid work flexibility
Software Engineer, Distributed Systems (Trust & Safety)
Software Engineer, Distributed Systems (Trust & Safety)

gamma.app • San Francisco (CA)

On-site
USD 180,000 - 310,000
Software Engineer, Growth
Software Engineer, Growth

gamma.app • San Francisco (CA)

On-site
USD 120,000 - 180,000