Staff ML Performance Engineer — Edge Inference Optimizer

Icehouseventures

Greater London

Hybrid

GBP 70,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Icehouseventures is looking for a Staff ML Performance Engineer to optimize ML inference for edge accelerators and GPUs. You will focus on running large transformer-based models efficiently on low-cost, low-power devices, directing technical projects and ensuring reliable operation in vehicular compute environments.

This full-time role operates under a hybrid policy, balancing office collaboration with remote work. Candidates can expect a dynamic environment, fostering innovation and inclusion.

Qualifications

  • Proven experience improving performance in production systems with tight constraints.
  • Strong proficiency with at least one relevant stack/toolchain.
  • Comfort operating at multiple levels of abstraction.
  • Strong software engineering fundamentals.
  • Clear communicator and collaborative teammate.

Responsibilities

  • Profile and pinpoint bottlenecks across the inference stack.
  • Implement and validate optimisations in compilers and runtimes.
  • Build robust benchmarking and regression testing.
  • Optimise for multiple targets.
  • Collaborate with model developers.

Skills

Performance improvement in production systems
Proficiency with TensorRT, CUDA, Qualcomm QNN
Software engineering fundamentals
Collaboration and communication skills
Python proficiency
C++ proficiency

Tools

TensorRT
CUDA
Qualcomm QNN
Triton
OpenCL

Job description

Icehouseventures is looking for a Staff ML Performance Engineer to optimize ML inference for edge accelerators and GPUs. You will focus on running large transformer-based models efficiently on low-cost, low-power devices, directing technical projects and ensuring reliable operation in vehicular compute environments.

This full-time role operates under a hybrid policy, balancing office collaboration with remote work. Candidates can expect a dynamic environment, fostering innovation and inclusion.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff ML Performance Engineer: Edge Inference Optimizer
Staff ML Performance Engineer: Edge Inference Optimizer

EngineersOfAI • Greater London

Hybrid
GBP 120,000 - 180,000
Staff ML Performance Engineer: Edge Inference & Systems
Staff ML Performance Engineer: Edge Inference & Systems

AI Startups UK • Greater London

Hybrid
GBP 120,000 - 170,000
Staff ML Performance Engineer - Edge Inference Optimizer
Staff ML Performance Engineer - Edge Inference Optimizer

Wayve • Greater London

On-site
GBP 120,000 - 180,000
Senior Edge ML Performance Engineer
Senior Edge ML Performance Engineer

Icehouseventures • Greater London

On-site
GBP 90,000 - 130,000
Senior ML Engineer - Edge AI Performance & Deployment
Senior ML Engineer - Edge AI Performance & Deployment

United States Digital Space LLC • Greater London

On-site
GBP 90,000 - 125,000
Staff ML Performance Engineer (Compiler)
Staff ML Performance Engineer (Compiler)

EngineersOfAI • Greater London

Hybrid
GBP 120,000 - 180,000
Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
Inference Performance Engineer
Inference Performance Engineer

adaption • Greater London

On-site
GBP 90,000 - 130,000
Flexible work
Lunch stipend
Well-Being benefits
Senior ML Engineer: AI Performance & Edge Deployment
Senior ML Engineer: AI Performance & Edge Deployment

AI Startups UK • Greater London

Hybrid
GBP 90,000 - 140,000
Senior ML Performance Engineer - Real-Time Inference & Scale
Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey • Greater London

On-site
GBP 70,000 - 90,000