ML Inference Engineer, C++/Python, Arm Acceleration

Arm Limited

Galway

Hybrid

EUR 85,000 - 125,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Arm Limited is seeking engineers to develop and optimize C++, C, and Python software that connects machine-learning platforms with Arm hardware acceleration. This role involves integrating ML frameworks and runtimes with Arm technology, including llama.cpp, and leading complex projects from design to delivery.

The ideal candidate will have strong C++, C, and Python skills, experience delivering features across teams, and knowledge of LLM inference.

Qualifications

  • Strong C++, C, and Python programming and software-design skills.
  • Experience delivering significant software features across multiple teams.
  • Familiarity with Git, GitLab, continuous integration, and automated testing.
  • Effective technical leadership and communication skills.
  • Knowledge of LLM inference, including tokenization, attention, KV caches, batching and quantisation.

Responsibilities

  • Develop and optimize C++, C, and Python software that integrates ML platforms with Arm hardware acceleration.
  • Design and implement integrations between ML frameworks, runtimes, and hardware-acceleration technologies.
  • Analyse and improve the performance of ML and LLM workloads.
  • Lead significant work from design through delivery, mentor junior engineers, and support technical decisions.

Skills

C++
Python
Software design
Git/GitLab/CI/Testing
Technical leadership
LLM inference basics

Tools

llama.cpp

Job description

Arm Limited is seeking engineers to develop and optimize C++, C, and Python software that connects machine-learning platforms with Arm hardware acceleration. This role involves integrating ML frameworks and runtimes with Arm technology, including llama.cpp, and leading complex projects from design to delivery.

The ideal candidate will have strong C++, C, and Python skills, experience delivering features across teams, and knowledge of LLM inference.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff ML Frameworks Engineer - LLM & Acceleration
Staff ML Frameworks Engineer - LLM & Acceleration

Arm Limited • Galway

Hybrid
EUR 90,000 - 130,000
Senior Software Engineer – ML Inference Runtime
Senior Software Engineer – ML Inference Runtime

Arm Limited • Galway

Hybrid
EUR 85,000 - 125,000
Staff Software Engineer – ML Inference Runtime
Staff Software Engineer – ML Inference Runtime

Arm Limited • Galway

Hybrid
EUR 90,000 - 130,000
Senior ML Systems Engineer - Hybrid, AI/Compiler Lead
Senior ML Systems Engineer - Hybrid, AI/Compiler Lead

Arm Limited • Galway

Hybrid
EUR 70,000 - 110,000
Senior ML Software Engineer - Hybrid & Impact
Senior ML Software Engineer - Hybrid & Impact

Arm • Galway

Hybrid
EUR 97,000 - 133,000
Training and professional development
Friendly working environment
Flexibility in hybrid work
Senior ML Software Engineer – Compiler & Framework
Senior ML Software Engineer – Compiler & Framework

Arm Limited • Galway

Hybrid
EUR 65,000 - 95,000
Senior Software ML Engineer
Senior Software ML Engineer

Arm Limited • Galway

Hybrid
EUR 70,000 - 110,000
Staff Software ML Engineer
Staff Software ML Engineer

Arm Limited • Galway

Hybrid
EUR 65,000 - 95,000
Staff Software ML Engineer
Staff Software ML Engineer

Arm • Galway

Hybrid
EUR 97,000 - 133,000
Training and professional development
Friendly working environment
Flexibility in hybrid work
Senior AI Inference Engineer - High-Throughput LLM Serving
Senior AI Inference Engineer - High-Throughput LLM Serving

Confidential • Ireland

On-site
EUR 120,000 - 180,000