AI Runtime Engineer - High-Performance ML Platform

Modular

City of Edinburgh

On-site

GBP 83,000 - 124,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Amazing Team.
World-class Benefits.
Competitive Compensation.
Team Building Events.

Job summary

Modular in Edinburgh is seeking an AI Runtime Engineer to own a runtime across CPU, GPU and accelerator hardware, delivering cross-stack optimizations for diverse models. You will design and implement runtime improvements, optimize performance, and collaborate with multiple teams to reach state-of-the-art end-to-end performance.

The role requires 2+ years in high-performance computing and strong C++ skills, plus experience with performance analysis and profiling tools.

Qualifications

  • 2+ years of experience working on high-performance computing systems.
  • Experience in C++ programming and complex software systems.
  • Experience with CPU or GPU runtime optimizations and performance analysis on CPUs, GPUs, or AI accelerators.
  • Proficiency with profiling tools (CPU or GPU).
  • Creativity and curiosity for solving complex problems, a team-oriented attitude that enables you to work well with others, and alignment with our culture.

Responsibilities

  • Design and develop runtime and cross-stack optimizations to improve CPU, GPU, and accelerator efficiency, addressing issues such as CPU overhead, caching, and data locality across multiple devices.
  • Work with vendor-specific networking libraries to unlock high performance data transfer for multiple topologies.
  • Collaborate with the compiler, kernels, serving, and models teams to design core technologies that achieve state-of-the-art end-to-end performance on various CPU and GPU hardware.
  • Collaborate with the customer success team and engage with customers to understand their performance requirements and use cases.
  • Collaborate with tooling and infrastructure teams to design systems for automated performance analysis and benchmarking.

Skills

C++
High-performance computing
Performance analysis
Profiling tools
Team collaboration

Job description

Modular in Edinburgh is seeking an AI Runtime Engineer to own a runtime across CPU, GPU and accelerator hardware, delivering cross-stack optimizations for diverse models. You will design and implement runtime improvements, optimize performance, and collaborate with multiple teams to reach state-of-the-art end-to-end performance.

The role requires 2+ years in high-performance computing and strong C++ skills, plus experience with performance analysis and profiling tools.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Runtime Engineer
AI Runtime Engineer

Modular • City of Edinburgh

On-site
GBP 83,000 - 124,000
Amazing Team.
World-class Benefits.
Competitive Compensation.
+1
AI Compiler Optimization Engineer - Edinburgh
AI Compiler Optimization Engineer - Edinburgh

microTECH Global Limited • City of Edinburgh

Hybrid
GBP 90,000 - 120,000
AI Runtime Engineer
AI Runtime Engineer

Modular, a Qualcomm company • City of Edinburgh

On-site
GBP 83,000 - 124,000
World-class benefits
RSU grants
Team building events
C++ ML Runtime Engineer - AI Acceleration & Systems
C++ ML Runtime Engineer - AI Acceleration & Systems

Arm Limited • Cambridge

On-site
GBP 70,000 - 120,000
Senior C++ ML Framework & Runtime Engineer
Senior C++ ML Framework & Runtime Engineer

CamWebDir • United Kingdom

Hybrid
GBP 85,000 - 120,000
ML Runtime & Inference Systems Engineer – Hybrid
ML Runtime & Inference Systems Engineer – Hybrid

Arm • Cambridge

Hybrid
GBP 74,000 - 100,000
AI Hardware Acceleration Engineer for ML Performance
AI Hardware Acceleration Engineer for ML Performance

XTX Markets • Greater London

On-site
GBP 60,000 - 100,000
Onsite gym
Extensive medical benefits
Daily breakfast and lunch
+2
Senior C++ ML Inference Runtime Engineer
Senior C++ ML Inference Runtime Engineer

Semiconductor Engineering • United Kingdom

Remote
GBP 70,000 - 110,000
C++ ML Framework & Runtime Engineer
C++ ML Framework & Runtime Engineer

Arm Limited • United Kingdom

Hybrid
GBP 90,000 - 130,000
C++ ML Frameworks & Runtime Engineer
C++ ML Frameworks & Runtime Engineer

Arm Limited • Cambridge

Hybrid
GBP 65,000 - 90,000
Hybrid working
Accommodations during recruitment