Senior Infra Engineer: Observability, Cloud & ML Ops

Creandum

Greater London

Hybrid

GBP 90,000 - 120,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive salary
Career growth opportunities
Dynamic multicultural team

Job summary

H is building the next generation of AI infrastructure to empower researchers and engineers. The Infra team ensures robust, scalable, and accessible infrastructure for foundational models, agents, and public services.

The role focuses on designing and managing infra, enabling multi-tenant and on-prem deployments, and implementing comprehensive observability with modern tooling to support rapid experimentation and reliable production workloads.

Qualifications

  • Observability and monitoring with modern tooling.
  • Proficient in Python or JavaScript/TypeScript.
  • Experience with cloud architectures and distributed systems is a plus.

Responsibilities

  • Design and manage the infrastructure to support research efforts (model and agent development).
  • Support product engineering on the agent platform including client-facing APIs and runtimes across deployment scenarios.
  • Set up and maintain observability and monitoring strategies.

Skills

Observability
Programming (Python/JS)

Tools

Datadog
Prometheus
Grafana
Docker
Kubernetes
CI/CD
Terraform
CDK

Job description

H is building the next generation of AI infrastructure to empower researchers and engineers. The Infra team ensures robust, scalable, and accessible infrastructure for foundational models, agents, and public services.

The role focuses on designing and managing infra, enabling multi-tenant and on-prem deployments, and implementing comprehensive observability with modern tooling to support rapid experimentation and reliable production workloads.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Infra Systems Engineer — ML Ops & Observability (Hybrid)
Infra Systems Engineer — ML Ops & Observability (Hybrid)

hcompany • Greater London

Hybrid
GBP 70,000 - 110,000
Hybrid work model
Competitive salary
Senior MLOps Engineer: Build Scalable AI Infra
Senior MLOps Engineer: Build Scalable AI Infra

Harnham - Data and Analytics Recruitment • Greater London

Hybrid
GBP 75,000 - 85,000
Hybrid work model
Platform Infra Engineer: Cloud Deployments & Observability
Platform Infra Engineer: Cloud Deployments & Observability

Scale AI, Inc. • Greater London

Hybrid
GBP 110,000 - 180,000
Senior MLOps Engineer — Lead AI Platform Infra (Hybrid London)
Senior MLOps Engineer — Lead AI Platform Infra (Hybrid London)

Harnham - Data & Analytics Recruitment • Greater London

Hybrid
GBP 75,000 - 85,000
Senior AI Infra Engineer: Scale Training & Inference
Senior AI Infra Engineer: Scale Training & Inference

Mat Vin • Greater London

Hybrid
GBP 90,000 - 130,000
MLOps & Platform Engineer: Build Scalable AI Infra
MLOps & Platform Engineer: Build Scalable AI Infra

DGH Recruitment • England

On-site
GBP 90,000 - 120,000
Platform & Infra Engineer for Enterprise AI
Platform & Infra Engineer for Enterprise AI

Conduct • Greater London

On-site
GBP 120,000 - 180,000
Staff Software Engineer, Observability & Profiling
Staff Software Engineer, Observability & Profiling

United States Digital Space LLC • Greater London

On-site
GBP 100,000 - 140,000
AI Infrastructure Engineering Lead
AI Infrastructure Engineering Lead

Mat Vin • Greater London

Hybrid
GBP 120,000 - 190,000
Senior MLOps Engineer: Own Scalable AI Platform (Hybrid)
Senior MLOps Engineer: Own Scalable AI Platform (Hybrid)

Harnham • Greater London

Hybrid
GBP 90,000 - 130,000
Bonus up to 10%
Hybrid work
Private medical
+3