Senior Platform Engineer, Reliability & Scale

Apply

Greater London

Hybrid

GBP 110,000 - 150,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Astronomer empowers data teams to bring mission-critical software, analytics, and AI to life. We are redefining how Apache Airflow runs at scale and building a high-scale PaaS platform.

Join us to influence the production systems, establish standards, and own critical reliability initiatives across Astro, Observe and our IDE product. As a Senior Software Engineer in Platform Engineering, you’ll help drive incident management, testing, deployment, and documentation while collaborating with a

Qualifications

  • Strong experience in Non-Abstract Systems design and implementation.
  • Strong proficiency in Python, Golang and Kubernetes (CKA or equivalent).
  • Experience with observability principles and technologies, including SLI/SLO definition and tracking.
  • Strong communication skills, both written and verbal, with experience in working with a globally distributed team in delivery.
  • A passion for reliability and operational excellence. A low tolerance for toil and other nonsense.
  • Ability to estimate the scope of work accurately and coordinate with stakeholders to address risks and ensure successful project delivery.
  • Experience with (and ideally strong opinions on) software development best practices, such as code review, testing, CI/CD, version control, automation and debugging.
  • Proactive approach to identifying and addressing issues, with a focus on ownership and accountability.

Responsibilities

  • Make high-quality, data-driven and experience-driven decisions on how we build this and the next generation of our production platform, then deliver the results.
  • Own and build how we test, build and deploy code in a high-scale PaaS environment.
  • Collaborate across the whole company on how we design production systems, set standards and make technology choices for new and existing products, and how these fit together.
  • Deliver results - we routinely "change the wheels on the bus while it's moving", in a predictable, safe and reliable way.
  • Be at the forefront of how we work together as a Platform Engineering team.
  • Blaze a Trail: Work on a small but growing team on building out the Platform/Reliability practice for the company - this role reports directly to the VP of Reliability.
  • Be an Owner: Be directly involved in decision-making on what we work on, as well as how we work on it. Make promises, and keep them.
  • Do Sensible Things: Be directly involved in determining how our platform works. Participate in incident management and determine sensible practices as the platform evolves.
  • Garage Door Open: Create and maintain comprehensive internal documentation for systems and processes, ensuring clarity and accessibility.

Skills

Non-Abstract Systems design
Python
Golang
Kubernetes
Observability (SLI/SLO)
Incident management
Communication
Reliability & operational excellence
Estimation & planning
CI/CD
Automation & debugging

Tools

CircleCI
Chronosphere
Splunk
Bazel
Istio
Playwright
Karpenter
GitHub Actions

Job description

Astronomer empowers data teams to bring mission-critical software, analytics, and AI to life. We are redefining how Apache Airflow runs at scale and building a high-scale PaaS platform.

Join us to influence the production systems, establish standards, and own critical reliability initiatives across Astro, Observe and our IDE product. As a Senior Software Engineer in Platform Engineering, you’ll help drive incident management, testing, deployment, and documentation while collaborating with a

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Platform Engineer – Reliability at Scale
Senior Platform Engineer – Reliability at Scale

Astronomer • Greater London

Hybrid
GBP 90,000 - 120,000
Senior Software Engineer, Platform
Senior Software Engineer, Platform

United States Digital Space LLC • Greater London

On-site
GBP 90,000 - 150,000
Senior Software Engineer, Platform
Senior Software Engineer, Platform

Astronomer • Greater London

Hybrid
GBP 90,000 - 120,000
Senior Software Engineer, Platform
Senior Software Engineer, Platform

Apply • Greater London

Hybrid
GBP 110,000 - 150,000
Senior Platform Engineer: DataOps & Cloud Reliability
Senior Platform Engineer: DataOps & Cloud Reliability

United States Digital Space LLC • Greater London

On-site
GBP 90,000 - 150,000
Staff Software Engineer, Platform Infrastructure
Staff Software Engineer, Platform Infrastructure

Sierra Ventures • Greater London

On-site
GBP 85,000 - 120,000
Senior Platform Engineer: Scale Real-Time Infra & SRE
Senior Platform Engineer: Scale Real-Time Infra & SRE

Matcha • United Kingdom

Remote
GBP 105,000 - 125,000
£105,000–£125,000 salary
Equity in an early-stage tech company
25 days holiday plus local publicholid
+5
Platform Engineer for Scalable AI Systems
Platform Engineer for Scalable AI Systems

scaleai • Greater London

On-site
GBP 120,000 - 180,000
Senior Platform Engineer — Scale & Observability
Senior Platform Engineer — Scale & Observability

Understanding Recruitment • Greater London

On-site
GBP 90,000 - 120,000
Lucrative Performance-based bonus
Equity package with significant long‑m
Platform Engineer: Reliability & Scale Dynamo
Platform Engineer: Reliability & Scale Dynamo

Mat Vin • Greater London

Hybrid
GBP 90,000 - 130,000
Private medical insurance
Competitive annual leave
First Friday of every month off
+5