AI Runtime Engineer

Modular

City of Edinburgh

On-site

GBP 83,000 - 124,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Amazing Team.
World-class Benefits.
Competitive Compensation.
Team Building Events.

Job summary

Modular in Edinburgh is seeking an AI Runtime Engineer to own a runtime across CPU, GPU and accelerator hardware, delivering cross-stack optimizations for diverse models. You will design and implement runtime improvements, optimize performance, and collaborate with multiple teams to reach state-of-the-art end-to-end performance.

The role requires 2+ years in high-performance computing and strong C++ skills, plus experience with performance analysis and profiling tools.

Qualifications

  • 2+ years of experience working on high-performance computing systems.
  • Experience in C++ programming and complex software systems.
  • Experience with CPU or GPU runtime optimizations and performance analysis on CPUs, GPUs, or AI accelerators.
  • Proficiency with profiling tools (CPU or GPU).
  • Creativity and curiosity for solving complex problems, a team-oriented attitude that enables you to work well with others, and alignment with our culture.

Responsibilities

  • Design and develop runtime and cross-stack optimizations to improve CPU, GPU, and accelerator efficiency, addressing issues such as CPU overhead, caching, and data locality across multiple devices.
  • Work with vendor-specific networking libraries to unlock high performance data transfer for multiple topologies.
  • Collaborate with the compiler, kernels, serving, and models teams to design core technologies that achieve state-of-the-art end-to-end performance on various CPU and GPU hardware.
  • Collaborate with the customer success team and engage with customers to understand their performance requirements and use cases.
  • Collaborate with tooling and infrastructure teams to design systems for automated performance analysis and benchmarking.

Skills

C++
High-performance computing
Performance analysis
Profiling tools
Team collaboration

Job description

About the role:

ML developers today face significant friction when deploying trained models. They work in a fragmented space with incomplete, patchwork solutions that require extensive performance tuning and model-specific optimizations. At Modular, we are building the next-generation AI platform that will radically improve how developers build and deploy AI models.

A core part of this offering is a platform that enables customers to achieve state-of-the-art performance across model families and frameworks. As an AI Runtime Engineer, you will own a runtime that operates on various CPU, GPU, and accelerator hardware platforms, optimizing performance for diverse customer AI models.

LOCATION:Candidates based in the United Kingdom are welcome to apply. This role will be based in our Edinburgh office (minimum 3 days per week on-site) with relocation assistance provided for eligible candidates. All new hires complete onboarding in-person.

What you will do:
  • Design and develop runtime and cross-stack optimizations to improve CPU, GPU, and accelerator efficiency, addressing issues such as CPU overhead, caching, and data locality across multiple devices.
  • Work with vendor-specific networking libraries to unlock high performance data transfer for multiple topologies.
  • Collaborate with the compiler, kernels, serving, and models teams to design core technologies that achieve state-of-the-art end-to-end performance on various CPU and GPU hardware.
  • Collaborate with the customer success team and engage with customers to understand their performance requirements and use cases.
  • Collaborate with tooling and infrastructure teams to design systems for automated performance analysis and benchmarking.
What you bring to the table:
  • 2+ years of experience working on high-performance computing systems.
  • Experience in C++ programming and complex software systems.
  • Experience with CPU or GPU runtime optimizations and performance analysis on CPUs, GPUs, or AI accelerators.
  • Proficiency with one or more profiling tools (CPU or GPU).
  • Creativity and curiosity for solving complex problems, a team-oriented attitude that enables you to work well with others, and alignment with our culture.
Helpful, but not required:
  • Experience with ML graph optimizations, parallel / distributed programming, heterogeneous ML computation, and/or code generation.
  • Exposure to MLIR, LLVM, and/or the Mojo programming language.
  • Advanced degree in Computer Science or a related area is a plus.
What Modular brings to the table:
  • Amazing Team.We are a progressive and agile team with some of the industry’s best engineering and product leaders.
  • World-class Benefits.In order to attract the best, we need to offer the best. Your benefits package may include comprehensive healthcare coverage, retirement and savings programs, employee stock purchase opportunities, paid time off, wellbeing resources, family support programs, and learning and development opportunities. Please note that specific benefit packages may vary based on your location, you can read more about benefits offered by Qualcomm here.
  • Competitive Compensation.We offer very strong compensation packages, including RSU grants. We want people to be focused on their best work and believe in tailoring compensation plans to meet the needs of our workforce.
  • Team Building Events. We organize regular team onsites and local meetups in Los Altos, CA as well as different cities. Traveling 2-4 times a year is expected for all roles.

Working at Modular will enable you to grow quickly as you work alongside incredibly motivated and talented people who have high standards, possess a growth mindset, and a purpose to truly change the world.

The estimated base salary range for this role to be performed in the United Kingdom, is £82,800.00 - £123,600.00 GBP.

The salary for the successful applicant will depend on a variety of permissible, non-discriminatory job-related factors, which include but are not limited to education, training, work experience, business needs, or market demands. This range may be modified in the future. The total compensation for a candidate will also include annual target bonus, equity, and benefits, with equity making up a significant portion of your total compensation.

For candidates who fall outside of the listed requirements, we nevertheless encourage you to apply as we may have upcoming openings that are lower/higher level than the ones advertised.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Runtime Engineer
AI Runtime Engineer

Modular, a Qualcomm company • City of Edinburgh

On-site
GBP 83,000 - 124,000
World-class benefits
RSU grants
Team building events
AI Inference Tools Engineer
AI Inference Tools Engineer

Modular • City of Edinburgh

Hybrid
GBP 83,000 - 150,000
Amazing Team
World-class Benefits
Competitive Compensation
+1
AI Runtime Engineer - High-Performance ML Platform
AI Runtime Engineer - High-Performance ML Platform

Modular • City of Edinburgh

On-site
GBP 83,000 - 124,000
Amazing Team.
World-class Benefits.
Competitive Compensation.
+1
Machine Learning Framework/Runtime Software Engineer
Machine Learning Framework/Runtime Software Engineer

Arm • Cambridge

On-site
GBP 74,000 - 100,000
Machine Learning Framework/Runtime Software Engineer (C++)
Machine Learning Framework/Runtime Software Engineer (C++)

Arm • Cambridge

Hybrid
GBP 55,000 - 75,000
Experienced Machine Learning Framework/Runtime Software Engineer
Experienced Machine Learning Framework/Runtime Software Engineer

CamWebDir • United Kingdom

Hybrid
GBP 85,000 - 120,000
Staff / Principal Machine Learning Engineer, Serving
Staff / Principal Machine Learning Engineer, Serving

Inworld AI • United Kingdom

On-site
GBP 140,000 - 200,000
Principal AI Platform Engineer
Principal AI Platform Engineer

Arm • Cambridge

Hybrid
GBP 126,000 - 171,000
Senior AI Engineer
Senior AI Engineer

MCS Group | Your Specialist Recruitment Consultancy • Belfast

Hybrid
GBP 80,000 - 95,000
Senior Performance Modelling Engineer (CPU group)
Senior Performance Modelling Engineer (CPU group)

Arm • Cambridge

Hybrid
GBP 74,000 - 100,000