Site Reliability Engineer: Scale, Automate & Resilient Systems

United States Digital Space LLC

United States

Remote

USD 120,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

United States Digital Space LLC is seeking Site Reliability Engineers for our Infrastructure Platforms group. We hire across Intermediate through Senior Staff levels, matching you to the best opportunity based on your experience and needs.

We value growth mindset and rapid learning, with high-performance culture and AI-driven workflows. You’ll join a team focused on reliability, automation, and scalable production systems.

Qualifications

  • Experience keeping production systems reliable, with an ops mindset and software engineering practice.
  • Proficiency with Kubernetes, CI/CD, and infrastructure as code.
  • Production experience with Go or Ruby in software systems.
  • Hands-on experience with Terraform and cloud providers (AWS/GCP).
  • Strong debugging, incident management, and observability skills.

Responsibilities

  • Keep user-facing services and production systems reliable, scalable, and efficient.
  • Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code workflows.
  • Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling.
  • Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps.
  • Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately.
  • Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early.
  • Take part in incident response and post-incident reviews, turning learnings into changes in automation and process.
  • Document runbooks, architecture decisions, and reviews so findings become repeatable practices.

Skills

Reliability engineering
Automation
Go
Ruby
Kubernetes
Incident response
Observability
CI/CD
GitOps
On-call

Education

Bachelor's degree in Computer Science or related field

Tools

Terraform
Kubernetes operators
Production automation
AWS
GCP
Go tooling

Job description

United States Digital Space LLC is seeking Site Reliability Engineers for our Infrastructure Platforms group. We hire across Intermediate through Senior Staff levels, matching you to the best opportunity based on your experience and needs.

We value growth mindset and rapid learning, with high-performance culture and AI-driven workflows. You’ll join a team focused on reliability, automation, and scalable production systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer - Automate & Scale
Senior Site Reliability Engineer - Automate & Scale

United States Digital Space LLC • New York (NY)

Hybrid
USD 179,000 - 226,000
Unlimited PTO
Employee stock options
Medical, dental, vision with HSA
+6
Senior Platform SRE: Scale & Reliability
Senior Platform SRE: Scale & Reliability

United States Digital Space LLC • United States

Hybrid
USD 103,000 - 162,000
Health insurance
Vacation and RTT
Mental health and coaching
+7
Senior SRE - CI/CD Platform & AI-Driven Reliability
Senior SRE - CI/CD Platform & AI-Driven Reliability

United States Digital Space LLC • United States

Remote
USD 140,000 - 215,000
Market leader compensation
Comprehensive wellness programs
Paid vacation and holidays
+5
Remote SRE & Infra SWE Lead — Scale, Secure, Automate
Remote SRE & Infra SWE Lead — Scale, Secure, Automate

United States Digital Space LLC • United States

Hybrid
USD 100,000 - 140,000
Senior Site Reliability Engineer - Scale Resilient Systems
Senior Site Reliability Engineer - Scale Resilient Systems

Jobgether • United States

Remote
USD 150,000 - 200,000
Staff SRE — Platform Reliability Lead
Staff SRE — Platform Reliability Lead

United States Digital Space LLC • United States

Hybrid
USD 127,000 - 161,000
Deutschlandticket (Germany-wide public
28 vacation days
Work from abroad up to 10 days/year
+9
Senior Kubernetes SRE: Scale, Automation & Reliability
Senior Kubernetes SRE: Scale, Automation & Reliability

United States Digital Space LLC • Bastrop (TX)

On-site
USD 100,000 - 130,000
Senior SRE: Secure Cloud Platform & Automation Leader
Senior SRE: Secure Cloud Platform & Automation Leader

United States Digital Space LLC • Bellevue (CA)

On-site
USD 180,000 - 230,000
Amazing Benefits
Making Social Impact
Fostering Diversity, Equity, Inclusion
Senior DB Reliability Engineer: Scale & Automate DB Platform
Senior DB Reliability Engineer: Scale & Automate DB Platform

United States Digital Space LLC • United States

Remote
USD 56,000 - 106,000
SRE Production Engineer: Mission-Critical Systems
SRE Production Engineer: Mission-Critical Systems

United States Digital Space LLC • Hawthorne (CA)

On-site
USD 125,000 - 195,000
Stock options
401(k) retirement plan
Medical, vision, dental coverage
+3