Remote SRE — Observability Platform Lead

Aalyria

United States

On-site

USD 125,000 - 150,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid/remote flexibility
401(k) plan, health/dental/vision/life

Job summary

Aalyria seeks an experienced SRE/Platform Engineer to build its centralized observability stack for satellite and aerospace networks. You will shape the strategy, standardize tooling (Prometheus, OpenTelemetry, Tempo/Jaeger), and scale across GCP and AWS. This remote US role includes on-call duties.

You will collaborate with SWE and infra teams to define SLOs/SLIs, automate deployment with Terraform and ArgoCD, and drive reliability for high-availability distributed systems.

Qualifications

  • Active TS/SCI security clearance required.
  • 4+ years in an SRE or platform engineering role with observability focus.
  • Hands-on experience building and scaling observability platforms (Prometheus, Grafana, Loki, OpenTelemetry, Tempo/Jaeger, Honeycomb).
  • Strong production-level experience with GCP and Kubernetes.
  • Experience using IaC and GitOps (Terraform, ArgoCD).
  • Proficiency in Go and Python for debugging and tooling.

Responsibilities

  • Design and build Aalyria's centralized observability platform, integrating and scaling metrics, logging, and tracing.
  • Define, implement, and manage SLOs/SLIs and error budgets for core products.
  • Partner with SWEs to implement observability best practices, templates, and documentation.
  • Automate deployment, scaling, and management of the observability stack with IaC and GitOps.
  • Collaborate with core infra to ensure visibility into Kubernetes clusters and cloud environments (GCP/AWS).
  • Develop and lead monitoring, alerting, and incident-response strategy with blameless post-mortems.

Skills

SRE/Platform engineering experience
Go/Python tooling
Security clearance TS/SCI

Tools

Prometheus
Grafana
Loki
OpenTelemetry
Tempo/Jaeger
Honeycomb
Terraform
ArgoCD
Kubernetes
GCP
AWS
GitLab CI

Job description

Aalyria seeks an experienced SRE/Platform Engineer to build its centralized observability stack for satellite and aerospace networks. You will shape the strategy, standardize tooling (Prometheus, OpenTelemetry, Tempo/Jaeger), and scale across GCP and AWS. This remote US role includes on-call duties.

You will collaborate with SWE and infra teams to define SLOs/SLIs, automate deployment with Terraform and ArgoCD, and drive reliability for high-availability distributed systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE - Observability Platform for Satellite Networks(Hybrid)
SRE - Observability Platform for Satellite Networks(Hybrid)

Aalyria Technologies, Inc. • United States

Remote
USD 125,000 - 150,000
Equity options
Flexible hybrid remote/in‑office
401(k) benefits
+2
Site Reliability Engineer, USG
Site Reliability Engineer, USG

Aalyria • United States

On-site
USD 125,000 - 150,000
Hybrid/remote flexibility
401(k) plan, health/dental/vision/life
Remote SRE Lead - Incident, Reliability & Observability
Remote SRE Lead - Incident, Reliability & Observability

NightDragon Acquisition Corp. • United States

On-site
USD 260,000 - 280,000
Hybrid & Remote Work
Competitive Compensation
Equity package
Senior Observability SRE — Telemetry Platform & Cloud Infra
Senior Observability SRE — Telemetry Platform & Cloud Infra

Anduril Industries • Boston (MA)

On-site
USD 166,000 - 220,000
Equity grants
Top-tier benefits
Observability Engineer / Site Reliability Engineer
Observability Engineer / Site Reliability Engineer

Ontrac Solutions • New York (NY)

On-site
USD 120,000 - 190,000
Remote Observability Platform Engineer — Scale Telemetry
Remote Observability Platform Engineer — Scale Telemetry

BairesDev • Peru (IL)

On-site
USD 120,000 - 170,000
100% remote work
USD or local currency pay
Home office setup
+3
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000
Senior Observability & SRE Engineer (GCP/Kubernetes)
Senior Observability & SRE Engineer (GCP/Kubernetes)

Ontrac Solutions • New York (NY)

On-site
USD 120,000 - 190,000
Observability Engineer / Site Reliability Engineer
Observability Engineer / Site Reliability Engineer

Ontrac Solutions • Chicago (IL)

On-site
USD 120,000 - 180,000
Remote Senior DevOps & SRE — Observability & Resilience
Remote Senior DevOps & SRE — Observability & Resilience

adhoc • McLean (VA)

On-site
USD 120,000 - 150,000
Company-subsidized health, dental, and
vision insurance
Flexible PTO
+3