Site Reliability Engineer — Mission-Critical Systems

Iac/interactivecorp

Sacramento (CA)

On-site

USD 120,000 - 170,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Stock options
Health insurance
Paid vacation

Job summary

SpaceX is hiring a Site Reliability Engineer to operate and scale mission-critical software products used in launch, test, and Starlink operations. You’ll own the software delivery pace and work with a team of engineers to keep systems reliable.

You will join a collaborative culture with on-call rotations, strong emphasis on automation, observability, and performance tuning in a fast-paced aerospace environment.

Qualifications

  • Bachelor's degree in computer science, information systems, or engineering; OR 2+ years of professional experience with site reliability or DevOps in lieu of a degree.
  • Experience with Linux operating systems.
  • 5+ years of DevOps, site reliability engineering, or system administration experience.
  • 3+ years of experience with Python and Python-based development frameworks.
  • Experience with source code and version control tools such as Git or Subversion.
  • Experience with infrastructure as code (IaC) products for automatically managing fleets of servers.
  • Experience with build systems (Make, Bazel/Pants/Buck, Gradle, etc.) and package management tools (pip, npm, etc.).
  • Experience with container and virtualization technologies (Docker, Kubernetes, etc.).
  • Experience with Terraform, Ansible, Puppet, or other automation frameworks.
  • Knowledge of TCP/IP networking.
  • Experience with databases and data modeling.
  • Experience with JIRA and workflow/issue management.
  • Ability to work on mission-critical and sensitive systems with urgency.

Responsibilities

  • Deploy, upgrade, operate/maintain, and scale our suite of mission-critical products and services.
  • Manage our infrastructure as code and use observability tools to monitor application health.
  • Collaborate with software engineers to create highly operable and maintainable products.
  • Engage in and improve the full software development lifecycle from design to deployment.
  • Practice sustainable incident response and blameless post-mortems.
  • Provide end-user support to vehicle software engineers for products.
  • Participate in the team’s on-call rotation periodically.
  • Focus on performance bottlenecks and improvement techniques.

Skills

DevOps
Site Reliability
Python
Git
IaC
Build Systems
Docker
Kubernetes
Terraform
Ansible
Networking
Databases
JIRA
Communication

Education

Bachelor's degree in CS/IS/Engineering

Tools

Git
Subversion
Docker
Kubernetes
Terraform
Ansible
Puppet
Bazel/Pants/Buck
Gradle
Pip
NPM

Job description

SpaceX is hiring a Site Reliability Engineer to operate and scale mission-critical software products used in launch, test, and Starlink operations. You’ll own the software delivery pace and work with a team of engineers to keep systems reliable.

You will join a collaborative culture with on-call rotations, strong emphasis on automation, observability, and performance tuning in a fast-paced aerospace environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer — Mission-Critical Infra & Automation
Site Reliability Engineer — Mission-Critical Infra & Automation

SpaceX • Hawthorne (CA)

On-site
USD 125,000 - 195,000
Stock options
401(k) retirement plan
Medical, vision, dental coverage
+1
Production Site Reliability Engineer—Mission-Critical Infra
Production Site Reliability Engineer—Mission-Critical Infra

InvestedintheMission • Hawthorne (CA)

On-site
USD 125,000 - 195,000
Stock options
401(k) plan
Medical, vision, and dental insurance
Software Engineer - Build Mission-Critical Systems On-Site
Software Engineer - Build Mission-Critical Systems On-Site

Future Ventures • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer — Multi-Region Systems
Senior Site Reliability Engineer — Multi-Region Systems

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 265,000
Senior Site Reliability Engineer – Starlink Global Infra
Senior Site Reliability Engineer – Starlink Global Infra

SpaceX • Palo Alto (CA)

On-site
USD 165,000 - 280,000
Stock options
401(k)
Medical, vision, dental
+3
Senior Site Reliability Engineer – Global, Multi-Region Infra
Senior Site Reliability Engineer – Global, Multi-Region Infra

InvestedintheMission • Palo Alto (CA)

On-site
USD 165,000 - 280,000
Stock options
401(k) plan
Medical, vision, dental coverage
+2
Senior Software Engineer — Mission-Critical Space Systems
Senior Software Engineer — Mission-Critical Space Systems

SPACE EXPLORATION TECHNOLOGIES CORP • Hawthorne (CA), Northern (KY)

Hybrid
USD 165,000 - 230,000
Stock options
Medical, vision & dental
401(k)
+2
Senior Software Engineer - Build Mission-Critical Systems
Senior Software Engineer - Build Mission-Critical Systems

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Medical coverage
Paid vacation
+1
Site Reliability Engineer (Application Software)
Site Reliability Engineer (Application Software)

Iac/interactivecorp • Sacramento (CA)

On-site
USD 120,000 - 170,000
Stock options
Health insurance
Paid vacation
Site Reliability Engineer: HPC & Automation
Site Reliability Engineer: HPC & Automation

SpaceX • Redmond (WA)

On-site
USD 125,000 - 150,000
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
Paid parental leave
+2