Platform Engineer – Scale GPU Infra for AI Platform

Ineffable Intelligence LTD

Greater London

Hybrid

GBP 85,000 - 120,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ineffable Intelligence LTD is seeking a Platform Engineer for our London office to shape the evolving platform that powers our AI research. You will own the infrastructure powering large-scale GPU compute, ensuring reliability, perf and fast developer feedback.

You’ll work across Kubernetes, containers, cloud providers and observability tools to keep the platform humming and accelerate experimentation. This is a hands-on role with real ownership and impact.

Qualifications

  • Experience designing and operating large-scale GPU-enabled platforms.
  • Strong background in cloud infrastructure and reliability engineering.
  • Familiarity with observability tools and developer experience tooling.

Responsibilities

  • Own and optimise platform infrastructure for reliable GPU compute at scale.
  • Design and implement resilient cluster and tooling to speed up experimentation.
  • Improve developer environments to accelerate research workflows.
  • Collaborate on Cloud infrastructure across providers and ensure observability is actionable.

Skills

GPU scheduling
Platform engineering
Observability
Dev tooling
Python
Rust

Tools

Kubernetes
containers
KAI scheduler
Kueue
Google Cloud
Grafana
Datadog
Tailscale
Workbrew
dev containers

Job description

Ineffable Intelligence LTD is seeking a Platform Engineer for our London office to shape the evolving platform that powers our AI research. You will own the infrastructure powering large-scale GPU compute, ensuring reliability, perf and fast developer feedback.

You’ll work across Kubernetes, containers, cloud providers and observability tools to keep the platform humming and accelerate experimentation. This is a hands-on role with real ownership and impact.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer – Scale AI Infra & GPU Orchestration
Platform Engineer – Scale AI Infra & GPU Orchestration

Ineffable Intelligence • Greater London

On-site
GBP 70,000 - 110,000
Platform Engineer - AI Infra & Large-Scale GPU Systems
Platform Engineer - AI Infra & Large-Scale GPU Systems

Ineffable • Greater London

On-site
GBP 90,000 - 110,000
Infrastructure Engineering Manager, AI Platforms
Infrastructure Engineering Manager, AI Platforms

Scale AI • Greater London

On-site
GBP 110,000 - 180,000
Scale-Focused AI Infra & MLOps Engineer
Scale-Focused AI Infra & MLOps Engineer

EngineersOfAI • Greater London

On-site
GBP 90,000 - 120,000
AI Infrastructure Engineering Manager – Platform & Automation
AI Infrastructure Engineering Manager – Platform & Automation

scaleai • Greater London

On-site
GBP 120,000 - 180,000
Senior AI Infra & DS Engineer — London
Senior AI Infra & DS Engineer — London

LinuxRecruit • Greater London

On-site
GBP 100,000 - 140,000
Senior Data Platform Engineer for Large-Scale AI Data
Senior Data Platform Engineer for Large-Scale AI Data

Scale • Greater London

On-site
GBP 120,000 - 170,000
GPU Infrastructure Engineer — Scale & Reliability for Production
GPU Infrastructure Engineer — Scale & Reliability for Production

Triwill Group • Greater London

Hybrid
GBP 90,000 - 120,000
GPU Infra Engineer — Scale, Automation & AI Compute
GPU Infra Engineer — Scale, Automation & AI Compute

OpenAI • Greater London

On-site
GBP 120,000 - 190,000
Staff Inference Platform Engineer — Low-Latency GPU, Kubernetes
Staff Inference Platform Engineer — Low-Latency GPU, Kubernetes

CoreWeave Europe • Greater London

On-site
GBP 120,000 - 190,000
Medical Insurance
Pension Plan
Life Insurance
+1