Senior Site Reliability Engineer

Satsuma AI, Inc.

Austin, Northern (TX, KY)

Hybrid

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Unlimited PTO
401(K)
Healthcare Stipend
Gym stipend

Job summary

Satsuma AI, Inc. is seeking a Senior Site Reliability Engineer to own the reliability, scalability, and operational posture of our multi-cloud infrastructure. You will build resilient systems, prevent fires, and improve on-call experiences.

This infra-first role embraces AI-assisted development (Claude Code) for tooling, runbooks, and automation, and you will partner with engineering on reliability reviews and architecture decisions.

Qualifications

  • 5–8 years in SRE, DevOps, or infrastructure engineering.
  • Hands-on across at least two major cloud providers.
  • Kubernetes, Terraform, and observability tooling (Datadog, Grafana).
  • Comfortable reading/editing code and shipping scripts/tools.
  • Experience with AI-assisted development (e.g., Claude Code).
  • On-call maturity; incident ownership and postmortems.
  • Prior startup or high-growth SaaS experience.
  • Familiarity with API gateway infrastructure or commerce tech stacks.

Responsibilities

  • Own infrastructure across AWS, GCP, and Azure environments.
  • Build and maintain CI/CD pipelines, observability stacks, and incident response workflows.
  • Define and enforce SLOs/SLIs; lead postmortems.
  • Author and maintain IaC (Terraform preferred).
  • Write internal tooling and automation using AI-assisted development workflows.
  • Partner closely with engineering on reliability reviews and architecture decisions.

Skills

SRE mindset
Cloud-native
On-call

Tools

Kubernetes
Terraform
Datadog
Grafana
CI/CD

Job description

About Satsuma

Satsuma is a commerce iPaaS that builds merchant-specific APIs, MCP Servers, and MCP Apps, enabling retailers to connect their full commerce stack once and deploy branded shopping experiences across every AI channel. We work with enterprise retailers and move fast. Our infra has to match.

The role

We're looking for a Senior SRE to own the reliability, scalability, and operational posture of Satsuma's multi-cloud infrastructure. You'll be the person who keeps things running, builds the systems that prevent fires, and makes on-call not terrible.

This is an infra-first role. But we're an AI-native company, and we expect you to use AI-assisted development (Claude Code) as a core part of your workflow — writing tooling, automating runbooks, building internal utilities.

What you'll do
  • Own infrastructure across AWS, GCP, and Azure environments
  • Build and maintain CI/CD pipelines, observability stacks, and incident response workflows
  • Define and enforce SLOs/SLIs; lead postmortems
  • Author and maintain IaC (Terraform preferred)
  • Write internal tooling and automation using AI-assisted development workflows
  • Partner closely with engineering on reliability reviews and architecture decisions
  • 5-8 years in SRE, DevOps, or infrastructure engineering
  • Hands-on experience across at least two major cloud providers
  • Strong Kubernetes, Terraform, and observability tooling (Datadog, Grafana, or equivalent)
  • Comfortable reading and editing code; able to ship scripts and internal tools
  • Experience with AI-assisted development (Copilot, Cursor, Claude Code)
  • On-call maturity -- you've owned incidents end-to-end and made systems better afterward
  • Prior experience at a startup or high-growth SaaS company
  • Familiarity with API gateway infrastructure or commerce tech stacks
  • Hands-on experience with MCP or agentic AI infrastructure
  • Unlimited PTO
  • 401(K)
  • Healthcare Stipend
  • Gym stipend
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Satsuma • Austin (TX)

On-site
USD 150,000 - 210,000
Unlimited PTO
401(K)
Healthcare Stipend
+1
Senior SRE - Multi-Cloud Reliability & AI-Driven Ops
Senior SRE - Multi-Cloud Reliability & AI-Driven Ops

Satsuma AI, Inc. • Austin (TX), Northern (KY)

Hybrid
USD 140,000 - 210,000
Unlimited PTO
401(K)
Healthcare Stipend
+1
Senior SRE — AI-Driven, Multi-Cloud Reliability Leader
Senior SRE — AI-Driven, Multi-Cloud Reliability Leader

Satsuma • Austin (TX)

On-site
USD 150,000 - 210,000
Unlimited PTO
401(K)
Healthcare Stipend
+1
Senior Site Reliability Engineer II
Senior Site Reliability Engineer II

Juniper Square • United States

On-site
USD 165,000 - 195,000
Health, dental, and vision care
Life insurance
Mental wellness coverage
+3
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cvent • Tysons (VA)

Hybrid
USD 110,000 - 140,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kovoro • Denver (CO), Northern (KY)

Hybrid
USD 150,000 - 190,000
Senior Backend Engineer for AI-First Platform - Equity & PTO
Senior Backend Engineer for AI-First Platform - Equity & PTO

Satsuma • Austin (TX)

On-site
USD 100,000 - 130,000
Retirement Plan (401k, IRA)
Paid Time Off (Vacation, Sick & Public Holidays)
Free Food & Snacks
+2
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cvent, Inc. • Tysons (VA)

Hybrid
USD 100,000 - 130,000