Edge & Cloud SRE: Fleet Reliability & Observability

Specter

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Specter is seeking a Site Reliability Engineer to manage the operational health of our sensor platform, ensuring the reliability of both edge hardware deployed at customer sites and cloud infrastructure.

This role requires strong Linux systems administration skills, as well as experience with networking and cloud services. You will debug issues, build fleet management systems, and design observability strategies. Join us in automating the physical world with cutting-edge AI technology.

Qualifications

  • Strong Linux systems administration, comfortable working over SSH in production.
  • Experience with edge or on-prem hardware alongside cloud infrastructure.
  • Solid networking fundamentals including DNS, firewalls, and VPNs.

Responsibilities

  • Debug and triage issues across diverse Linux-based sensor nodes.
  • Build and maintain fleet management systems for updates and diagnostics.
  • Design and implement observability across edge devices and cloud infrastructure.

Skills

Linux systems administration
Scripting in Python, Go, or Bash
Networking fundamentals
Experience with containerization
Embedded systems experience
AWS experience

Tools

Docker
Kubernetes

Job description

Specter is seeking a Site Reliability Engineer to manage the operational health of our sensor platform, ensuring the reliability of both edge hardware deployed at customer sites and cloud infrastructure.

This role requires strong Linux systems administration skills, as well as experience with networking and cloud services. You will debug issues, build fleet management systems, and design observability strategies. Join us in automating the physical world with cutting-edge AI technology.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Specter • San Francisco (CA)

On-site
USD 120,000 - 160,000
Platform SRE: Scalable Cloud & Kubernetes Engineer
Platform SRE: Scalable Cloud & Kubernetes Engineer

Specter • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 250,000
Edge & Cloud SRE — Reliability & Observability
Edge & Cloud SRE — Reliability & Observability

Claryo, Inc. • San Francisco (CA)

On-site
USD 150,000 - 170,000
Medical coverage
Dental coverage
Vision coverage
+4
Senior Observability SRE – Fleet Telemetry & Infra
Senior Observability SRE – Fleet Telemetry & Infra

Anduril • United States

On-site
USD 166,000 - 220,000
Competitive salary
Comprehensive benefits
Equity
Edge & Cloud SRE: Reliability & Real-World Deployments
Edge & Cloud SRE: Reliability & Real-World Deployments

Claryo • San Francisco (CA)

On-site
USD 140,000 - 190,000
Senior SRE - Cloud & Observability
Senior SRE - Cloud & Observability

Ridgeline • Reno (NV)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Education reimbursement
Wellness reimbursement
+1
Senior SRE & Reliability Engineer - Edge & Observability
Senior SRE & Reliability Engineer - Edge & Observability

Hispanic Alliance for Career Enhancement • Scottsdale (AZ)

On-site
USD 93,000 - 204,000
Medical coverage
Dental coverage
Vision coverage
+1
Senior SRE - AI and Edge Platform Reliability
Senior SRE - AI and Edge Platform Reliability

Cognativ • United States

Remote
USD 180,000 - 260,000
Site Reliability Engineer (Edge Services), Infrastructure Services
Site Reliability Engineer (Edge Services), Infrastructure Services

Apple Inc. • Austin (TX)

On-site
USD 110,000 - 150,000
Software Engineer - SRE & Observability for Edge Fleet
Software Engineer - SRE & Observability for Edge Fleet

Hispanic Alliance for Career Enhancement • Richardson (TX)

On-site
USD 72,000 - 159,000