Senior Edge ML Performance Engineer

Icehouseventures

Greater London

On-site

GBP 90,000 - 130,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Wayve in London is seeking a Staff ML Performance Engineer to optimize inference for edge accelerators and GPUs, enabling transformer models to run efficiently on low-power devices for Wayve's driving product.

You will steer technical direction for production systems, work across ML systems, compilers, runtimes, and embedded deployment on multiple targets, collaborating with model developers to influence architecture and deployment decisions.

Qualifications

  • Proven experience improving performance in production systems.
  • Strong proficiency with stack/toolchain such as TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL, MLIR or ONNX.
  • Comfort operating across abstraction levels from high-level model behavior to low-level kernel/runtime execution.
  • Strong software engineering fundamentals (debugging, profiling, testing, maintainable code).
  • Clear communicator and collaborative teammate; able to align stakeholders on trade-offs.

Responsibilities

  • Identify, implement and validate optimisations in ML compilers, runtimes, and kernels (e.g., operator fusion, scheduling, quantisation‑aware performance).
  • Profile and pinpoint bottlenecks across the inference stack and deliver measurable improvements.
  • Build robust benchmarking and regression tests to ensure improvements hold across models and devices.
  • Develop and optimise for multiple target platforms (e.g., NVIDIA Orin/Thor, Qualcomm).
  • Collaborate with model developers to influence architecture and deployment decisions for on‑device performance.
  • Contribute to roadmaps and tooling to raise performance standards across the team.

Skills

Latency optimization
Profiling
Communication
Collaborative teamwork
Software engineering fundamentals

Tools

TensorRT
CUDA
Qualcomm QNN
Triton
OpenCL
MLIR
ONNX

Job description

Wayve in London is seeking a Staff ML Performance Engineer to optimize inference for edge accelerators and GPUs, enabling transformer models to run efficiently on low-power devices for Wayve's driving product.

You will steer technical direction for production systems, work across ML systems, compilers, runtimes, and embedded deployment on multiple targets, collaborating with model developers to influence architecture and deployment decisions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff ML Performance Engineer - Edge Inference Optimizer
Staff ML Performance Engineer - Edge Inference Optimizer

Wayve • Greater London

On-site
GBP 120,000 - 180,000
Staff ML Performance Engineer: Edge Inference & Systems
Staff ML Performance Engineer: Edge Inference & Systems

AI Startups UK • Greater London

Hybrid
GBP 120,000 - 170,000
Staff ML Performance Engineer: Edge Inference Optimizer
Staff ML Performance Engineer: Edge Inference Optimizer

EngineersOfAI • Greater London

Hybrid
GBP 120,000 - 180,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

EngineersOfAI • Greater London

Hybrid
GBP 120,000 - 180,000
Staff ML Performance Engineer — Edge Inference Optimizer
Staff ML Performance Engineer — Edge Inference Optimizer

Icehouseventures • Greater London

Hybrid
GBP 70,000 - 90,000
Senior ML Engineer: AI Performance & Edge Deployment
Senior ML Engineer: AI Performance & Edge Deployment

AI Startups UK • Greater London

Hybrid
GBP 90,000 - 140,000
Staff ML Performance Engineer (Inference Optimisation)
Staff ML Performance Engineer (Inference Optimisation)

Wayve • Greater London

On-site
GBP 70,000 - 90,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

Icehouseventures • Greater London

On-site
GBP 90,000 - 130,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

Wayve • Greater London

On-site
GBP 120,000 - 180,000
Senior ML Engineer — Production AI Performance
Senior ML Engineer — Production AI Performance

EngineersOfAI • Greater London

On-site
GBP 90,000 - 130,000