AI Infrastructure Engineer: GPU & Performance Optimisation

twentyAI

Greater London

On-site

GBP 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

twentyAI is seeking an AI Infrastructure Engineer to build autonomous systems that optimise modern AI workloads. You will work across GPU programming, ML infrastructure, and automated optimisation to tackle challenging performance problems with real production impact.

The role involves designing, deploying and tuning production AI infrastructure, measuring end-to-end efficiency, and contributing to open-source projects. Lead candidates may mentor a small team.

Qualifications

  • Experience building performance-critical software for GPU-accelerated environments.
  • Deep understanding of GPU performance optimisation and memory usage.
  • Experience with mixed-precision and quantised AI workloads (INT4, INT8, FP8).
  • Understanding of distributed AI systems and performance techniques.
  • Ability to diagnose bottlenecks using modern profiling tools.

Responsibilities

  • Develop, optimise and deploy performance-critical software for modern AI workloads, including low-level GPU acceleration.
  • Own production ML infrastructure and AI systems from design to deployment.
  • Measure how performance improvements affect end-to-end efficiency.
  • Build infrastructure for large-scale optimisation experiments with tracking and benchmarking.
  • Design automated optimisation and search strategies to improve system performance.
  • Investigate bottlenecks and translate findings into engineering improvements.
  • Share knowledge via documentation, talks and open-source contributions.
  • For lead candidates, hire and mentor a small team.

Skills

GPU programming
Performance optimization
Profiling tools
Distributed AI systems
Transformer architectures
Open-source contributions
Research publications

Tools

CUDA
Triton
CuTe
Helion
Megatron-LM
vLLM

Job description

twentyAI is seeking an AI Infrastructure Engineer to build autonomous systems that optimise modern AI workloads. You will work across GPU programming, ML infrastructure, and automated optimisation to tackle challenging performance problems with real production impact.

The role involves designing, deploying and tuning production AI infrastructure, measuring end-to-end efficiency, and contributing to open-source projects. Lead candidates may mentor a small team.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Systems Engineer | VC-backed Startup | TWE46402
AI Systems Engineer | VC-backed Startup | TWE46402

twentyAI • Greater London

On-site
GBP 90,000 - 140,000
GPU Infra Engineer — Scale, Automation & AI Compute
GPU Infra Engineer — Scale, Automation & AI Compute

OpenAI • Greater London

On-site
GBP 120,000 - 190,000
GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning
GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning

United States Digital Space LLC • Greater London

Hybrid
GBP 120,000 - 170,000
GPU Infrastructure Engineer for Scalable AI Inference
GPU Infrastructure Engineer for Scalable AI Inference

AI Startups UK • Greater London

Hybrid
GBP 120,000 - 180,000
Performance Engineer: AI Systems & Optimization
Performance Engineer: AI Systems & Optimization

CommonAI CIC • Cambridge

On-site
GBP 65,000 - 90,000
Collaborative environment
High impact in growing org
Competitive salary and pension
+3
AI Infrastructure Architect
AI Infrastructure Architect

Accenture UK & Ireland • Greater London

On-site
GBP 120,000 - 170,000
Senior AI Systems Performance Engineer
Senior AI Systems Performance Engineer

NVIDIA Development UK Limited • United Kingdom

On-site
GBP 100,000 - 180,000
Site Reliability Engineer, GPUs in AI
Site Reliability Engineer, GPUs in AI

Radley James • Greater London

On-site
GBP 20,000 - 40,000
Senior AI Infrastructure Engineer - Scale Multi-GPU Training
Senior AI Infrastructure Engineer - Scale Multi-GPU Training

LinuxRecruit • Greater London

On-site
GBP 90,000 - 120,000
Competitive salary
Equity in early-stage startup
London lab
Compute Infrastructure Engineer: Scale AI Compute
Compute Infrastructure Engineer: Scale AI Compute

OpenAI • Greater London

On-site
GBP 171,000 - 302,000
Competitive salary
Remote work flexibility
Health insurance