Senior Site Reliability Engineer

LeoLabs

San Francisco (CA)

Hybrid

USD 190,000 - 210,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible remote
Challenging missions
Unlimited PTO
Salary + equity
Health coverage
Space ops leadership

Job summary

LeoLabs is seeking a Senior Site Reliability Engineer to bridge development and operations. You will design scalable systems, automate deployments, set up monitoring, and lead incident response to boost reliability across a global radar network.

The role emphasizes capacity planning, security, and collaboration with product teams to improve availability and efficiency. Hybrid remote work and competitive compensation are offered.

Qualifications

  • 5+ years in Site Reliability Engineering, DevOps, or related role.
  • Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent work experience.
  • Proficiency in scripting/programming (Python, Go).
  • Experience with cloud services (AWS, Azure).
  • Containerization proficiency (Docker, Kubernetes, ECS).
  • Configuration management tools (Terraform, Atlantis, Terragrunt).
  • CI/CD tools (GitHub Actions, AWS CodeBuild, CircleCI).
  • Monitoring tools (Grafana, Datadog).
  • Databases (RDS, Aurora, PostgreSQL).
  • Experience with large-scale distributed systems and microservices.

Responsibilities

  • Design, implement, and maintain scalable, reliable systems.
  • Set up monitoring and incident response plans to identify and resolve issues.
  • Develop and maintain automation tools for deployment, monitoring, and health checks.
  • Analyze capacity and performance to forecast future needs and scale accordingly.
  • Collaborate with development teams to improve product reliability and deployment processes.
  • Create and maintain architecture and incident documentation.
  • Participate in on-call rotations for 24/7 support.
  • Implement security best practices across systems and ensure compliance.

Skills

Python
Go
Communication
Problem solving
Security clearance
Site Reliability

Education

Bachelor’s degree in CS/Engineering

Tools

Docker
Kubernetes
ECS
Terraform
Atlantis
Terragrunt
GitHub Actions
AWS CodeBuild
CircleCI
Grafana
Datadog
RDS
Aurora
PostgreSQL
Distributed systems

Job description

At LeoLabs, we’re building the living map of activity in space. Through our proprietary global radar network and AI-enabled analytics platform, we collect millions of measurements daily on more than 25,000 objects in low Earth orbit (LEO). Our radar-powered intelligence protects billions in assets, monitors adversarial behavior, and ensures safe operations for commercial and government missions.

We’re not just building technology, we are redefining global security, safety, and transparency in space. As orbital activity accelerates and threats grow more complex, LeoLabs is a trusted partner for Space Domain Awareness, Space Traffic Management, and Satellite Operations for top-tier space operators and allied defense organizations.

If you're looking to work on mission-critical challenges at the forefront of aerospace, national security, and AI, your impact starts here.

Role Overview:

LeoLabs is seeking a skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will bridge the gap between development and operations, ensuring that our systems are scalable, reliable, and efficient. You will be responsible for automating processes, monitoring system performance, and resolving incidents to enhance our service reliability.

Key Responsibilities:

  • System Reliability: Design, implement, and maintain scalable and reliable systems.
  • Monitoring and Incident Response: Set up monitoring tools and create incident response plans to quickly identify and resolve issues, as well as implementing preventative measures.
  • Automation: Develop and maintain scripts and automation tools for deployment, monitoring, and system health checks.
  • Capacity Planning: Analyze system capacity and performance metrics to forecast future needs and implement scaling solutions.
  • Collaboration: Work closely with development teams to enhance product reliability and streamline the deployment process.
  • Documentation: Create and maintain documentation for system architecture, processes, and incident reports.
  • On-Call Support: Participate in on-call rotations to provide 24/7 support for critical systems.

Security: Implement and enforce security best practices across all systems, ensuring compliance with industry standards.

Qualifications:

  • Education: Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience.
  • Experience: 5+ years of experience in a Site Reliability Engineering, DevOps, or related role.
  • Technical Skills:
  • Proficiency in scripting or programming language (e.g., Python, Go)
  • Experience with cloud services (AWS, Azure)
  • Proficiency with containerization (Docker, Kubernetes, ECS)
  • Proficiency in configuration management tools (Terraform, Atlantis, Terragrunt)
  • Familiarity with CI/CD tools (GitHub Actions, AWS CodeBuild, CircleCI)
  • Experience with monitoring tools (Grafana, Datadog)
  • Familiarity with database technologies (RDS, Aurora, PostgreSQL)
  • Experience with large-scale distributed systems and microservices architecture.
  • Problem-Solving: Strong analytical and problem-solving skills with the ability to troubleshoot complex systems.
  • Communication: Excellent verbal and written communication skills, with the ability to collaborate effectively across teams.
  • Ability to obtain a U.S. personnel security clearance.

Preferred qualifications

  • Active TS/SCI clearance

What Success Looks Like:

What Success Looks Like:

Within 1 month, you’ll:

  • Complete onboarding to understand our business, vision, and team structure.
  • Get familiar with LeoLabs' engineering stack, security posture, and key initiatives.
  • Gain an understanding about how your role fits into LeoLabs broader organization.

Within 3 months, you’ll:

  • Independently deploy infrastructure changes using Infrastructure as Code.
  • Identify key reliability risks and recommend improvements.
  • Improve dashboards, alerts, and operational runbooks.

Within 6 months, you’ll:

  • Optimize infrastructure utilization and cloud costs without compromising reliability.
  • Drive automation that reduces operational toil and improves deployment reliability.

Within 12 months, you’ll:

  • Lead cross-functional initiatives to improve availability, scalability, and operational efficiency.
  • Be a key advisor for site reliability in new product developments and platform evolution.
  • Mentor junior engineers and foster the development culture.

Perks and Benefits

  • Global workforce: flexible remote/hybrid opportunities
  • Work on complex, meaningful missions with real-world impact
  • Unlimited paid time off for most roles
  • Competitive salary and equity packages
  • Comprehensive health, dental, and vision coverage
  • Access to the forefront of commercial space operations and defense innovation

Compensation for this role is based on the San Francisco Bay Area market and may be adjusted based on the candidate’s final location. The estimated base salary is $192,000, with additional compensation opportunities to include bonus and equity.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identify, national origin, disability, or status as a protected veteran.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Program Manager, National Security Programs
Senior Program Manager, National Security Programs

LeoLabs, Inc. • Chantilly (VA)

Hybrid
USD 140,000 - 190,000
Remote/hybrid options
Meaningful missions impact
Unlimited PTO
+3
Full Stack Engineer at LeoLabs
Full Stack Engineer at LeoLabs

Feedinkoo • United States

Hybrid
USD 90,000 - 130,000
Unlimited paid time off
Competitive salary and equity packages
Comprehensive health, dental, and vision coverage
Site Acquisition Lead
Site Acquisition Lead

LeoLabs • United States

Hybrid
USD 140,000 - 190,000
Global remote/hybrid opportunities
Competitive salary and equity packages
Comprehensive health, dental, and VIsA
Senior SRE — Mission-Critical Cloud & Automation (Remote)
Senior SRE — Mission-Critical Cloud & Automation (Remote)

LeoLabs • San Francisco (CA)

Hybrid
USD 190,000 - 210,000
Flexible remote
Challenging missions
Unlimited PTO
+3
Front End Developer at LeoLabs
Front End Developer at LeoLabs

Feedinkoo • United States

Hybrid
USD 80,000 - 120,000
Unlimited paid time off
Flexible remote/hybrid opportunities
Competitive salary and equity packages
+1
Senior Program Manager, National Security Programs
Senior Program Manager, National Security Programs

LeoLabs, Inc. • Washington

Hybrid
USD 120,000 - 160,000
Unlimited paid time off
Competitive salary and equity packages
Comprehensive health, dental, and vision coverage
+1
Technical Recruiter
Technical Recruiter

Socket.dev • Washington

On-site
USD 90,000 - 130,000
Program Manager
Program Manager

InvestedintheMission • Denver (CO)

Hybrid
USD 120,000 - 190,000
Global remote/hybrid opportunities
Work on complex, meaningful missions
Unlimited paid time off
+3
Sr. Site Reliability Engineer (Application Software)
Sr. Site Reliability Engineer (Application Software)

InvestedintheMission • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Medical coverage
Vision coverage
Dental coverage
+6
Sr. Site Reliability Engineer (Application Software)
Sr. Site Reliability Engineer (Application Software)

United States Digital Space LLC • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
Paid parental leave
+2