Remote Site Reliability Engineer — Observability Platform

Aalyria

United States

On-site

USD 115,000 - 135,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

401(k) plan
Equity
Hybrid remote work

Job summary

Aalyria, a leading aerospace technology company, seeks a senior SRE/Platform Engineer to build the core observability stack for satellite and space systems. You will own metrics, logs, and tracing, defining SLOs/SLIs and error budgets, and collaborating with engineers to implement scalable, reliable tooling.

Role involves on-call duties and guiding the roadmap from cloud-native tools to a robust production-grade platform using Prometheus, OpenTelemetry, Tempo, Terraform, and multi-cloud

Qualifications

  • 4+ years in SRE or platform engineering focusing on observability for large-scale systems.
  • Hands-on expertise building and scaling observability platforms (Prometheus, Grafana, Loki/ELK, OpenTelemetry, Tempo/Jaeger).
  • Strong production experience with GCP and Kubernetes.
  • Experience with IaC and GitOps (ArgoCD).
  • Proficiency in Go and Python for tooling.
  • Experience defining and managing SLOs/SLIs and error budgets.

Responsibilities

  • Design and build Aalyria's centralized observability platform, scaling metrics, logging, and tracing.
  • Define and manage SLOs/SLIs and error budgets for core products.
  • Collaborate with SWE teams to implement observability best practices and templates.
  • Automate deployment and management of the observability stack using Terraform and GitOps.

Skills

Prometheus
Grafana
OpenTelemetry
Tempo/Jaeger
Kubernetes
Go/Python tooling

Tools

GCP
AWS
Terraform
ArgoCD
GitLab CI

Job description

Aalyria, a leading aerospace technology company, seeks a senior SRE/Platform Engineer to build the core observability stack for satellite and space systems. You will own metrics, logs, and tracing, defining SLOs/SLIs and error budgets, and collaborating with engineers to implement scalable, reliable tooling.

Role involves on-call duties and guiding the roadmap from cloud-native tools to a robust production-grade platform using Prometheus, OpenTelemetry, Tempo, Terraform, and multi-cloud

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE - Observability Platform for Satellite Networks(Hybrid)
SRE - Observability Platform for Satellite Networks(Hybrid)

Aalyria Technologies, Inc. • United States

Remote
USD 125,000 - 150,000
Equity options
Flexible hybrid remote/in‑office
401(k) benefits
+2
Observability SRE for Satellite Network Platform (Remote)
Observability SRE for Satellite Network Platform (Remote)

aalyria-careers • United States

Hybrid
USD 140,000 - 210,000
401(k)
Health insurance
Equity options
+1
Remote SRE — Observability Platform Lead
Remote SRE — Observability Platform Lead

Aalyria • United States

On-site
USD 125,000 - 150,000
Hybrid/remote flexibility
401(k) plan, health/dental/vision/life
Site Reliability Engineer, Commercial
Site Reliability Engineer, Commercial

Aalyria • United States

On-site
USD 115,000 - 135,000
401(k) plan
Equity
Hybrid remote work
Site Reliability Engineer, Commercial
Site Reliability Engineer, Commercial

aalyria-careers • United States

Hybrid
USD 140,000 - 210,000
401(k)
Health insurance
Equity options
+1
Site Reliability Engineer, USG
Site Reliability Engineer, USG

Aalyria • United States

On-site
USD 125,000 - 150,000
Hybrid/remote flexibility
401(k) plan, health/dental/vision/life
Remote Senior Site Reliability Engineer — Observability
Remote Senior Site Reliability Engineer — Observability

Cribl • Des Moines (IA)

Remote
USD 142,000 - 195,000
Health insurance
Dental insurance
Vision insurance
+7
Remote Senior SRE — Observability & Cloud Reliability
Remote Senior SRE — Observability & Cloud Reliability

Cribl • Des Moines (IA)

Remote
USD 142,000 - 195,000
Health insurance
Dental insurance
Vision insurance
+7
Senior Site Reliability Engineer — Remote, AWS & Observability
Senior Site Reliability Engineer — Remote, AWS & Observability

Prove • United States

Hybrid
USD 140,000 - 190,000
Wellbeing reimbursement
401k Match
Parental Leave Policy
+5
Senior SRE: Automate Reliability & Observability
Senior SRE: Automate Reliability & Observability

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package