Senior AI Inference Platform Engineer

Lilasciences

Cambridge

On-site

GBP 144,480 - 204,680

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
International benefits

Job summary

Lila Sciences is seeking a Staff/Principal DevOps Engineer for AI Inference to design and optimize infrastructure for serving ML models at scale. You will bridge platform engineering, SRE, and ML infra to enable low-latency, high-throughput inference across GPU clusters and cloud accelerators.

You will collaborate with ML engineers, researchers, and software engineers to build reliable inference platforms, while maximizing compute efficiency and supporting multi-region deployments.

Qualifications

  • Expertise in DevOps, SRE, or Platform Engineering operating GPU/accelerator infra at scale.
  • Strong Kubernetes proficiency for ML workloads, including scheduling and device management.
  • Proficient in Python for automation and tooling in ML environments.

Responsibilities

  • Design and optimize GPU/accelerator infrastructure on Kubernetes for low-latency inference.
  • Build and maintain model serving platforms with scalable batching and routing.

Skills

DevOps/SRE/Platform Engineering
Kubernetes for ML workloads
Python automation

Tools

Terraform
Helm
Triton Inference Server

Job description

Lila Sciences is seeking a Staff/Principal DevOps Engineer for AI Inference to design and optimize infrastructure for serving ML models at scale. You will bridge platform engineering, SRE, and ML infra to enable low-latency, high-throughput inference across GPU clusters and cloud accelerators.

You will collaborate with ML engineers, researchers, and software engineers to build reliable inference platforms, while maximizing compute efficiency and supporting multi-region deployments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Inference Platform Engineer (Hybrid)
Senior AI Inference Platform Engineer (Hybrid)

Isomorphiclabs • Greater London

Hybrid
GBP 110,000 - 150,000
Staff ML & Physics Infrastructure Architect
Staff ML & Physics Infrastructure Architect

Lila Sciences • Greater London

On-site
GBP 166,000 - 218,000
Equity
Bonus potential
Senior Compute Infrastructure Engineer – AI Training & LLMs
Senior Compute Infrastructure Engineer – AI Training & LLMs

Inherentlabs • Greater London

On-site
GBP 70,000 - 90,000
Good lunch and dinner
Collaborative work culture
No bureaucracy
Lead AI/ML Platform Engineer — LLM Inference & Scale
Lead AI/ML Platform Engineer — LLM Inference & Scale

JPMorgan Chase & Co. • Auchentibber

On-site
GBP 90,000 - 140,000
ML Ops Engineer: AI Platform & GPU Infra
ML Ops Engineer: AI Platform & GPU Infra

Anaplan • Greater London

On-site
GBP 90,000 - 150,000
Senior AI Platform Engineer: Scale Infra & LLM Gateway
Senior AI Platform Engineer: Scale Infra & LLM Gateway

Wave Group • Greater London

Hybrid
GBP 69,000 - 115,000
Private medical
Share schemes
Learning allowance
+1
Senior ML Systems Engineer — LLM Inference & Serving
Senior ML Systems Engineer — LLM Inference & Serving

Google DeepMind • Greater London

Hybrid
GBP 153,000 - 222,000
Equity
Bonus target 20%
Comprehensive benefits
Lead ML Platform Engineer - Production AI
Lead ML Platform Engineer - Production AI

Accelerant • United Kingdom

Hybrid
GBP 90,000 - 150,000
Senior ML Platform Engineer - Scalable AI Deployment
Senior ML Platform Engineer - Scalable AI Deployment

Scale AI, Inc. • Greater London

On-site
GBP 100,000 - 150,000
Senior AI/ML Platform Engineer – Scale & Inference
Senior AI/ML Platform Engineer – Scale & Inference

JPMorganChase • Greater London

On-site
GBP 90,000 - 115,000