Senior Cloud Infrastructure Engineer - Scale & Observability

Thought Machine

Greater London

On-site

GBP 120,000 - 180,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive salary
Pension plan
Life insurance
Parental leave
Annual leave
Flexible hours
Cycle-to-work
Tech learning budget
Health and snack perks

Job summary

Thought Machine in London seeks a Senior Software Engineer, Infrastructure to scale our cloud-agnostic platform infrastructure and improve reliability at massive scale. You will own tooling in Python/Go, extend Kubernetes and Docker usage, and help advance multi-cloud readiness.

You will contribute to observability through Prometheus/OpenTelemetry/Grafana, strengthen DR/BCP, and collaborate with a talented team to deliver robust, scalable infrastructure across customers.

Qualifications

  • Degree in Computer Science, Engineering, or a similar technical field.
  • 5+ years of hands-on software engineering experience building and scaling infrastructure platforms.
  • Strong proficiency in building production-ready platform tooling using Python or Golang.
  • Deep understanding of container and orchestration internals (e.g., Kubernetes, Docker).
  • Hands-on experience extending and integrating open-source tools like Prometheus, OpenTelemetry, and Grafana.

Responsibilities

  • Scale the Core Platform: Evolve, scale, and optimise the cloud-agnostic platform infrastructure powering Vault globally.
  • Simplify Orchestration: Build intuitive control planes and automation to reduce infrastructure complexity.
  • Design for Resiliency: Engineer robust, fault-tolerant architectures and DR/BCP systems.
  • Advance Observability: Enhance observability suites with deep monitoring and insights for clients.
  • Optimise Data & Streaming: Scale databases and event-streaming for performance and cost.

Skills

Python
Go
Observability tooling
Container orchestration
Kubernetes
Docker
Prometheus
OpenTelemetry
Grafana

Education

Degree in Computer Science

Tools

Kubernetes
Docker
Prometheus
OpenTelemetry
Grafana

Job description

Thought Machine in London seeks a Senior Software Engineer, Infrastructure to scale our cloud-agnostic platform infrastructure and improve reliability at massive scale. You will own tooling in Python/Go, extend Kubernetes and Docker usage, and help advance multi-cloud readiness.

You will contribute to observability through Prometheus/OpenTelemetry/Grafana, strengthen DR/BCP, and collaborate with a talented team to deliver robust, scalable infrastructure across customers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Platform Engineer – Multi-Cloud & Observability
Senior Cloud Platform Engineer – Multi-Cloud & Observability

Whereby • Greater London

On-site
GBP 110,000 - 160,000
Competitive salary
Pension plan
Life insurance
+5
Senior Cloud Infrastructure Engineer | Scale a Global Platform
Senior Cloud Infrastructure Engineer | Scale a Global Platform

Whereby • Greater London

On-site
GBP 90,000 - 150,000
Competitive salary
Pension plan
Life insurance
+2
Senior SRE: Cloud-Scale, Fault-Tolerant SaaS Engineer
Senior SRE: Cloud-Scale, Fault-Tolerant SaaS Engineer

Thought Machine • Greater London

On-site
GBP 90,000 - 130,000
Employee share package
Cloud-Native Infrastructure Engineer – Automation & Observability
Cloud-Native Infrastructure Engineer – Automation & Observability

Thought Machine • Greater London

On-site
GBP 90,000 - 130,000
Pension plan (match up to 5%)
Life insurance (3x annual salary)
Maternity and paternity leave
+11
Platform Infra Engineer: Cloud Deployments & Observability
Platform Infra Engineer: Cloud Deployments & Observability

Scale AI, Inc. • Greater London

Hybrid
GBP 110,000 - 180,000
Senior Cloud & CI Infrastructure Engineer
Senior Cloud & CI Infrastructure Engineer

EngineersOfAI • Greater London

On-site
GBP 40,000 - 60,000
Senior Platform Engineer — Scale & Observability
Senior Platform Engineer — Scale & Observability

Understanding Recruitment • Greater London

On-site
GBP 90,000 - 120,000
Lucrative Performance-based bonus
Equity package with significant long‑m
Senior Platform Engineer — Scale & AI-Driven Infra
Senior Platform Engineer — Scale & AI-Driven Infra

Preply • Greater London

Hybrid
GBP 110,000 - 150,000
Monthly lessons allowance on Preply
Learning & Development budget
Equity
+1
Senior Platform & Observability Engineer
Senior Platform & Observability Engineer

Koda Tech • Greater London

Hybrid
GBP 90,000 - 120,000
Observability Engineer - Cloud Telemetry & Automation
Observability Engineer - Cloud Telemetry & Automation

develop • Greater London

On-site