Platform Engineer

Ineffable Intelligence LTD

Greater London

Hybrid

GBP 85,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ineffable Intelligence LTD is seeking a Platform Engineer for our London office to shape the evolving platform that powers our AI research. You will own the infrastructure powering large-scale GPU compute, ensuring reliability, perf and fast developer feedback.

You’ll work across Kubernetes, containers, cloud providers and observability tools to keep the platform humming and accelerate experimentation. This is a hands-on role with real ownership and impact.

Qualifications

  • Experience designing and operating large-scale GPU-enabled platforms.
  • Strong background in cloud infrastructure and reliability engineering.
  • Familiarity with observability tools and developer experience tooling.

Responsibilities

  • Own and optimise platform infrastructure for reliable GPU compute at scale.
  • Design and implement resilient cluster and tooling to speed up experimentation.
  • Improve developer environments to accelerate research workflows.
  • Collaborate on Cloud infrastructure across providers and ensure observability is actionable.

Skills

GPU scheduling
Platform engineering
Observability
Dev tooling
Python
Rust

Tools

Kubernetes
containers
KAI scheduler
Kueue
Google Cloud
Grafana
Datadog
Tailscale
Workbrew
dev containers

Job description

We are recruiting a Platform Engineer to join us in our London office.

Our mission is make first contact with superintelligence.

We are creating a superlearner that discovers all knowledge from its own experience, from elementary motor skills through to profound intellectual breakthroughs. This superlearning capability - the ability to endlessly discover knowledge and skills, without relying on human data - will be driven by the world’s most powerful reinforcement learning algorithms.

The superlearner is expected to rediscover and then transcend the greatest inventions in human history, such as language, science, mathematics and technology. If successful, this will represent a scientific breakthrough of comparable magnitude to Darwin: where his law explained all life, our law will explain and build all intelligence.

Role brief

As a foundational hire on the platform team, you'll be joining at a point where individual decisions shape how the platform is designed, not just maintained. This isn't a role where you inherit someone else's architecture, you'll be one of the people defining it.

We're pushing these learning methods to a scale that hasn’t been tried before and we care deeply about the infrastructure that gets us there. You'll own the infrastructure that the mission depends on; keeping large-scale GPU compute reliable, developer environments fast and the whole platform humming.

This is an opportunity for someone who wants to be a part of defining a new paradigm in AI and wants their work to directly speed up the experimentation and progress of our ambitious research.

This is a broad role, and the list below is non-exhaustive. You’ll be empowered to shape the role how you’d like in order to best enable our mission.

  • Kubernetes and containers: manage clusters and deploy internal tools

  • GPU scheduling at scale: hands-on with tools like KAI scheduler and Kueue to keep thousands of GPUs busy and productive

  • Making large-scale infrastructure resilient: get to the root of hardware failures and cluster-scale chaos, building the systems that prevent them from recurring

  • Cloud infrastructure: experience on Google Cloud and other major providers

  • Observability that actually helps: log management and monitoring with tools like QuickWit, Grafana or Datadog

  • Developer environments people love: tooling like Tailscale, Workbrew and dev containers that make everyone's day-to-day faster

  • Clean, well-crafted code: e.g. Python and Rust for tools and infrastructure that are a joy to build on

We move fast, give people real ownership and trust engineers to make good judgment calls. We are a team that cares as much about doing this well as doing it fast.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer
Platform Engineer

Ineffable Intelligence • Greater London

On-site
GBP 70,000 - 110,000
Member of Technical Staff (DevOps / Platform Engineer)
Member of Technical Staff (DevOps / Platform Engineer)

Ineffable • Greater London

On-site
GBP 90,000 - 110,000
Platform Engineer
Platform Engineer

Hiverge Ltd • Cambridge

On-site
GBP 70,000 - 115,000
Principal ML Platform Engineer
Principal ML Platform Engineer

Synthesia • Greater London

On-site
GBP 70,000 - 100,000
Site Reliability - Member of Technical Staff
Site Reliability - Member of Technical Staff

Callosum • Greater London

On-site
GBP 90,000 - 150,000
Equity & Ownership
Private healthcare
Visa sponsorship & relocation
+1
Engineering Manager, ML Infrastructure,
Engineering Manager, ML Infrastructure,

United States Digital Space LLC • Greater London

On-site
GBP 120,000 - 180,000
Staff Software Engineer
Staff Software Engineer

Unlikely AI • Greater London

Hybrid
GBP 120,000 - 180,000
Staff Engineer
Staff Engineer

Arrows • Greater London

Hybrid
GBP 90,000 - 150,000
AI Platform Engineer
AI Platform Engineer

Tempest Vane Partners • Greater London

Hybrid
GBP 120,000 - 180,000
Competitive compensation
Discretionary bonus
Benefits package
Staff Platform Engineer
Staff Platform Engineer

AI Startups UK • Greater London

Hybrid
GBP 90,000 - 120,000
Hybrid work schedule
Virtual Shares
30 days annual leave
+1