Site Reliability Engineer- Spacetime UK

Aalyria Technologies, Inc.

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity participation
Private health insurance
Generous leave

Job summary

Aalyria Technologies in London is seeking a senior SRE to design and build a centralized observability platform for satellite‑level systems. You will shape the strategy, implement best practices, and automate the stack using Terraform and ArgoCD.

You will own SLOs/SLIs, contribute to incident response, and partner with SWE teams to ensure reliability across Kubernetes and cloud environments, with on‑call responsibilities.

Qualifications

  • 4+ years in SRE/platform engineering focusing on observability for large-scale distributed systems.
  • Hands-on with Prometheus, Grafana, Loki/OpenTelemetry Tempo/Jaeger and related tooling.
  • Strong production experience with GCP and Kubernetes.
  • Experience with IaC and GitOps (ArgoCD).
  • Proficiency in Go and Python for tooling and debugging.
  • Defined, implemented, and managed SLOs/SLIs for production services.

Responsibilities

  • Design and build Aalyria's centralized observability platform.
  • Define and manage SLOs/SLIs and error budgets for core products.
  • Collaborate with SWE teams to implement best practices and templates.
  • Automate deployment, scaling, and management of the observability stack (IaC, GitOps).
  • Ensure visibility into Kubernetes, GCP, and AWS environments.
  • Lead monitoring, alerting, and incident response strategies.

Skills

Observability
Go
Python
Kubernetes
GCP
GitOps
Prometheus
OpenTelemetry

Tools

Grafana
Loki
Tempo/Jaeger
ArgoCD
Terraform

Job description

Role Overview

This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that ensures the reliability of systems critical to the operation of satellite megaconstellations and missions to deep space.

This is a greenfield/brownfield opportunity. You will be a trusted expert, helping to define and implement the strategy and building the tools that empower our engineers. You will support the roadmap to mature our observability stack, moving from cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role is for you.

Note: this role includes on-call responsibilities.

Key Responsibilities
  • Help design and build Aalyria's centralized observability platform, integrating and scaling tools for metrics (e.g. Prometheus), logging (e.g. Loki), and distributed tracing (e.g. Tempo/OpenTelemetry).
  • Define, implement, and manage a robust framework of Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets for our core products, ensuring we are launch-ready.
  • Partner with SWEs to implement observability best practices, develop standard templates and documentation, and configure tooling (e.g., OpenTelemetry libraries).
  • Automate the deployment, scaling, and management of the entire observability stack using Infrastructure as Code (e.g. Terraform) and GitOps principles (e.g. ArgoCD).
  • Partner closely with the core infrastructure team to ensure deep visibility into our Kubernetes clusters and underlying GCP and AWS environments.
  • Develop and lead the company's monitoring, alerting, and incident response strategy, driving a culture of proactive reliability and blameless post-mortems.
Required Qualifications
  • 4+ years of experience in an SRE or platform engineering role, with a focus on observability for large-scale, distributed compute or network systems.
  • Deep, hands-on expertise building, scaling, and managing observability platforms (e.g., Prometheus, Grafana, Loki/ELK, OpenTelemetry, Tempo/Jaeger, Honeycomb, etc.). You have proven experience using these tools to support performance analysis and debugging of complex distributed systems.
  • Strong production-level experience with Google Cloud Platform (GCP) and Kubernetes.
  • Experience using Infrastructure as Code (IaC) and GitOps principles (e.g., ArgoCD).
  • Proficiency in a systems programming language, with a strong preference for Go and Python for debugging and writing tooling.
  • Demonstrable experience defining, implementing, and managing SLOs, SLIs, and error budgets for production services for high availability distributed systems.
Preferred Qualifications
  • Experience operating a multi-cloud environment, specifically GCP and AWS.
  • Hands-on experience with GitLab CI for CI/CD pipelines.
  • Working knowledge of service mesh technologies such as Istio or Linkerd.
  • Familiarity with instrumenting applications written in Go and C++.
  • An active Secret clearance, or higher, is preferred for this position.
  • Experience with JVM observability (tuning, monitoring) for Java-based applications.
What We Offer
  • Competitive salary benchmarked to UK aerospace and defence technology market rates
  • Equity participation — share in Aalyria's growth at an early stage
  • Comprehensive benefits including pension, private health insurance, and generous annual leave
  • Flexible and hybrid working arrangements
  • The opportunity to work on genuinely novel technology with real-world operational impact across national security, commercial satellite, and deep-space programmes
  • A collaborative, low-hierarchy team environment with direct exposure to technical leadership and customers
Equal Opportunity Employer Statement

Aalyria Technologies is an equal opportunities employer. We are committed to building an inclusive workplace and welcome applicants from all backgrounds. We do not discriminate on the basis of race, religion, gender, sexual orientation, age, disability, or any other protected characteristic under UK law.

Aalyria Technologies operates in sectors subject to UK export control regulations. Candidates may be asked to confirm their eligibility to access export-controlled technology as part of the hiring process.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer- Spacetime UK
Site Reliability Engineer- Spacetime UK

Aalyria • Greater London

Hybrid
GBP 110,000 - 150,000
Equity participation
Pension plan
Private health insurance
+1
Mission Engineer - Spacetime UK
Mission Engineer - Spacetime UK

Aalyria • Greater London

Hybrid
GBP 90,000 - 130,000
Competitive salary
Equity participation
Private health insurance
+2
Senior Backend Software Engineer - Spacetime UK
Senior Backend Software Engineer - Spacetime UK

Aalyria Technologies, Inc. • Greater London

Hybrid
GBP 90,000 - 130,000
Equity
Health insurance
Pension
+5
Senior Software Engineer, Backend Software
Senior Software Engineer, Backend Software

Front Door Defense • Greater London

Hybrid
GBP 110,000 - 150,000
Innovative environment
Impactful work
Growth opportunities
+3
Senior Software Engineer, Backend
Senior Software Engineer, Backend

Aalyria Technologies, Inc. • Greater London

Hybrid
GBP 90,000 - 120,000
Hybrid remote work arrangement
Health insurance
Pension plan
+2
Solutions Architect, 5G NTN
Solutions Architect, 5G NTN

Aalyria Technologies, Inc. • Greater London

Hybrid
GBP 90,000 - 150,000
Health insurance
Equity
Pension
+1
UK Project Manager
UK Project Manager

Aalyria • Greater London

Hybrid
GBP 55,000 - 75,000
Competitive salary
Pension
Health insurance
+1
Software Engineer, DTN
Software Engineer, DTN

Aalyria • Greater London

Hybrid
GBP 85,000 - 120,000
Equity
Hybrid remote work
Competitive salary
+1
Observability Platform SRE - Greenfield, Cloud-Native
Observability Platform SRE - Greenfield, Cloud-Native

Aalyria • Greater London

Hybrid
GBP 110,000 - 150,000
Equity participation
Pension plan
Private health insurance
+1
Senior Software Engineer - Space Reliability
Senior Software Engineer - Space Reliability

spire • Glasgow

Hybrid
GBP 70,000 - 100,000
Name Your Satellite Program (NYSP)
Launch Attendance
Generous Time Off Policy
+7