Staff SRE — Cloud Reliability & Kubernetes Leader

Socket.dev

San Mateo (CA)

On-site

USD 240,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Stock options
Health insurance
401K savings plan

Job summary

Skydio is seeking a hands-on Staff Site Reliability Engineer to build, operate, and scale the cloud infrastructure powering our products. You’ll own production infrastructure including Kubernetes, AWS, infrastructure as code, CI/CD, observability, networking, and reliability.

You don't need to be an expert in every area, but you should have strong Kubernetes and cloud fundamentals with meaningful depth in at least one infrastructure domain.

Qualifications

  • 8+ years of experience as a Site Reliability Engineer, Platform Engineer, DevOps, Production Engineer or equivalent infrastructure role.
  • Strong hands-on experience operating Kubernetes, not simply deploying applications to existing clusters.
  • Experience managing Kubernetes/EKS upgrades and production clusters.
  • Strong AWS fundamentals, including VPCs, public/private subnets, networking, load balancers, EKS, IAM, and databases.
  • Production experience with Terraform or similar infrastructure-as-code tooling.
  • Experience owning or maintaining CI/CD and deployment systems such as Argo CD, Spinnaker, GitHub Actions, GitLab CI/CD, or Jenkins.
  • Experience diagnosing production infrastructure and networking problems.
  • Experience solving meaningful scaling or reliability challenges.
  • This position requires access to export-controlled data and U.S. government security requirements.

Responsibilities

  • Build, operate, and troubleshoot production Kubernetes/EKS clusters.
  • Perform Kubernetes upgrades, node rollouts, and cluster maintenance.
  • Build and manage AWS infrastructure including VPCs, networking, subnets, load balancers, IAM, EKS, databases, and storage.
  • Define and maintain infrastructure using Terraform.
  • Build and operate CI/CD and deployment infrastructure.
  • Troubleshoot production issues across Kubernetes, AWS, Linux, networking, and databases.
  • Build monitoring, alerting, and observability for critical infrastructure.
  • Participate in on-call rotations and respond to production incidents.
  • Identify and solve infrastructure scaling and reliability problems.
  • Automate operational work using Python, Go, or similar languages.
  • Help expand infrastructure across new regions and deployment environments.

Skills

Kubernetes
AWS
Terraform
CI/CD
Python
Go
Observability
On-call support

Tools

Argo CD
Spinnaker
GitHub Actions
GitLab CI/CD
Jenkins
Datadog

Job description

Skydio is seeking a hands-on Staff Site Reliability Engineer to build, operate, and scale the cloud infrastructure powering our products. You’ll own production infrastructure including Kubernetes, AWS, infrastructure as code, CI/CD, observability, networking, and reliability.

You don't need to be an expert in every area, but you should have strong Kubernetes and cloud fundamentals with meaningful depth in at least one infrastructure domain.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff SRE — Kubernetes & Cloud Reliability
Staff SRE — Kubernetes & Cloud Reliability

Skydio, Inc. • San Mateo (CA)

On-site
USD 240,000 - 300,000
Equity
Relocation aid
Health insurance
+1
Senior SRE: Cloud, Kubernetes & Scale Leader
Senior SRE: Cloud, Kubernetes & Scale Leader

Skydio • San Mateo (CA)

On-site
USD 240,000 - 300,000
Equity
Health insurance
Paid time off
+1
Staff Cloud SRE: Kubernetes, AWS & Reliability Lead
Staff Cloud SRE: Kubernetes, AWS & Reliability Lead

Skydio • United States

On-site
USD 240,000 - 300,000
Stock options
Health insurance
Paid vacation
+4
Site Reliability Engineer - Kubernetes & Cloud
Site Reliability Engineer - Kubernetes & Cloud

Skydio • United States

On-site
USD 180,000 - 240,000
Stock options
Health insurance
Paid vacation
+3
Cloud SRE: Kubernetes & Platform Engineer
Cloud SRE: Kubernetes & Platform Engineer

Booster • United States

Hybrid
USD 140,000 - 210,000
Staff Cloud & Kubernetes Infrastructure Engineer
Staff Cloud & Kubernetes Infrastructure Engineer

Skydio • San Mateo (CA)

Hybrid
USD 230,000 - 275,000
Competitive base salaries
Stock options
Comprehensive benefits packages
+2
Senior Cloud & Kubernetes Infrastructure Engineer
Senior Cloud & Kubernetes Infrastructure Engineer

Skydio, Inc. • San Mateo (CA)

On-site
USD 190,000 - 250,000
Equity (stock options)
Health insurance
401K savings plan
+1
Staff Cloud & Kubernetes Infrastructure Engineer
Staff Cloud & Kubernetes Infrastructure Engineer

Skydio • United States

Hybrid
USD 230,000 - 275,000
Equity in the form of stock options
Comprehensive health insurance
Paid vacation time
+1
Staff Infrastructure Engineer - Kubernetes, Cloud & Security
Staff Infrastructure Engineer - Kubernetes, Cloud & Security

Skydio Inc. • San Mateo (CA)

Hybrid
USD 230,000 - 275,000
Health insurance
Paid vacation
401(k) savings plan
+1
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Socket.dev • San Mateo (CA)

On-site
USD 240,000 - 300,000
Stock options
Health insurance
401K savings plan