Staff ML Performance Engineer: Edge Inference & Systems

AI Startups UK

Greater London

Hybrid

GBP 120,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Wayve is hiring a Staff ML Performance Engineer to optimise ML inference on edge devices, enabling large transformer models to run efficiently on cost-effective hardware for our autonomous driving product.

The role spans ML systems, compilers, runtimes, and embedded deployment, with hands-on work across multiple platforms and close collaboration with model developers to shape architecture and tooling.

Qualifications

  • Proven ability to optimise production systems under tight constraints (latency, memory, power).
  • Familiar with ML stacks and rapid learning of new frameworks.
  • Able to operate across abstraction levels from model behavior to kernel execution.
  • Strong software engineering fundamentals (debugging, profiling, testing).
  • Excellent communication and cross-team collaboration.

Responsibilities

  • Identify, implement and validate optimisations in ML compilers, runtimes, and kernels (e.g. operator fusion, scheduling, quantisation).
  • Profile and pinpoint bottlenecks across the full inference stack and deliver measurable improvements.
  • Build benchmarking and regression tests to ensure performance across models and devices.
  • Develop and optimise for multiple target platforms (e.g. NVIDIA Orin/Thor, Qualcomm).
  • Collaborate with model developers to influence architecture and deployment decisions.
  • Contribute to roadmaps and tooling and raise performance standards across the team.

Skills

Performance optimization
CUDA/TensorRT
Python
C++
Profiling
Communication

Tools

TensorRT
CUDA
Qualcomm QNN
Triton
ONNX

Job description

Wayve is hiring a Staff ML Performance Engineer to optimise ML inference on edge devices, enabling large transformer models to run efficiently on cost-effective hardware for our autonomous driving product.

The role spans ML systems, compilers, runtimes, and embedded deployment, with hands-on work across multiple platforms and close collaboration with model developers to shape architecture and tooling.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff ML Performance Engineer - Edge Inference Optimizer
Staff ML Performance Engineer - Edge Inference Optimizer

Wayve • Greater London

On-site
GBP 120,000 - 180,000
Senior Edge ML Performance Engineer
Senior Edge ML Performance Engineer

Icehouseventures • Greater London

On-site
GBP 90,000 - 130,000
Staff ML Performance Engineer: Edge Inference Optimizer
Staff ML Performance Engineer: Edge Inference Optimizer

EngineersOfAI • Greater London

Hybrid
GBP 120,000 - 180,000
Staff ML Performance Engineer — Edge Inference Optimizer
Staff ML Performance Engineer — Edge Inference Optimizer

Icehouseventures • Greater London

Hybrid
GBP 70,000 - 90,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

EngineersOfAI • Greater London

Hybrid
GBP 120,000 - 180,000
Senior ML Engineer: AI Performance & Edge Deployment
Senior ML Engineer: AI Performance & Edge Deployment

AI Startups UK • Greater London

Hybrid
GBP 90,000 - 140,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

Wayve • Greater London

On-site
GBP 120,000 - 180,000
Senior ML Engineer — Production AI Performance
Senior ML Engineer — Production AI Performance

EngineersOfAI • Greater London

On-site
GBP 90,000 - 130,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

Icehouseventures • Greater London

On-site
GBP 90,000 - 130,000
Staff ML Performance Engineer (Inference Optimisation)
Staff ML Performance Engineer (Inference Optimisation)

Wayve • Greater London

On-site
GBP 70,000 - 90,000