Senior Software Engineer – ML Inference Runtime

NLP PEOPLE

Galway

Hybrid

EUR 120,000 - 160,000

Full time

14 days+
Application generator

Get a reply from this recruiter — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Arm in Ireland seeks an experienced software engineer to help shape machine learning on Arm platforms, focusing on integrating ML frameworks with hardware acceleration.

You will work across teams to deliver LLM runtimes, including llama.cpp, and optimize performance for AI workloads on next-generation hardware. This fast-paced role involves mentoring engineers and collaborating on architecture and deployment, with a strong emphasis on DEI and cross-functional teamwork.

Qualifications

  • Strong C++, C, and Python programming and software-design skills.
  • Experience delivering significant software features across multiple teams.
  • Familiarity with Git, GitLab, continuous integration, and automated testing.
  • Effective technical leadership and communication skills.
  • Knowledge of LLM inference, including tokenization, KV caches, batching and quantisation.

Responsibilities

  • Develop and optimize C++, C, and Python software that integrates ML platforms with Arm hardware acceleration.
  • Design and implement integrations between ML frameworks, runtimes, and hardware-acceleration technologies.
  • Analyse and improve the performance of ML and LLM workloads.
  • Lead significant work from design through delivery; mentor junior engineers.

Skills

C++ programming
Python programming
Software design
Git/GitLab CI
Leadership
LLM inference
Hardware acceleration
Performance profiling
Tokenization/quantisation

Tools

llama.cpp
Vulkan
GPU compute

Job description

Job Overview:We have a phenomenal opportunity to help shape the future of Machine Learning and Artificial Intelligence on Arm platforms. Arm technology is used wherever computing matters. Are you passionate about ML systems and keen to influence how generative AI and other machine-learning workloads run efficiently on next-generation hardware?

Then we should talk!

We are growing our ML framework integrations team and seek dedicated and motivated engineers to join us. You will work with teams across Arm and with leading technology companies building products based on Arm technologies. You will share ideas, learn from outstanding engineers, and make contributions that have a real impact. This role offers the opportunity to provide develop the software that connects widely used machine-learning frameworks and runtimes to Arm acceleration technologies.

Responsibilities

Responsibilities include developing and optimizing C++, C, and Python software that integrates machine-learning platforms and execution environments with Arm hardware acceleration. This work entails integration with open-source LLM software such as llama.cpp. You will design and implement integrations between ML frameworks, runtimes, and hardware-acceleration technologies, as well as analyse and improve the performance of ML and LLM workloads. The role also involves leading significant work from design through delivery, mentoring junior engineers, and supporting technical decisions.

We are looking for someone with strong analytical and problem-solving skills who enjoys finding innovative solutions to complex technical challenges. You will be comfortable working in a fast-paced environment and collaborating across complementary teams to achieve shared goals.

Required Skills and Experience

Strong C++, C, and Python programming and software-design skillsExperience delivering significant software features across multiple teamsFamiliarity with Git, GitLab, continuous integration, and automated testingEffective technical leadership and communication skillsKnowledge of LLM inference, including tokenization, attention, KV caches, batching and quantisation

Nice To Have Skills and Experience

Experience with llama.cpp or other open-source ML or LLM runtimesExperience with profiling and optimizing ML workloadsExperience with compiler or ML graph technologiesExperience with Vulkan, GPU compute, or other hardware-acceleration APIsExperience using AI-assisted software-development tools effectively

In Return

All Arm employees are provided with vital training to succeed in their respective roles, as well as a friendly and high-performance working environment.

You will be working with a bunch of enthusiastic and brilliant colleagues. We are proud to have a set of behaviours that reflect our DEI (Diversity, Equity & Inclusion) culture and guide our decisions, defining how we work together to shape extraordinary!

Please note that no relocation package is available for this role.

#LI-CM1

|Develop C++, C and Python integrations connecting ML frameworks and LLM runtimes like llama.cpp with Arm acceleration technologies, optimising AI workloads for next-generation hardware.!

Accommodations at Arm

At Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email . To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.

Hybrid Working at Arm

Arm’s approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.

Equal Opportunities at Arm

Arm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Company

ARM

Level of experience (years)

Senior (5+ years of experience)

Tagged as: Industry, Ireland, Machine Learning, NLP

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer – ML Inference Runtime
Senior Software Engineer – ML Inference Runtime

Arm Limited • Galway

Hybrid
EUR 85,000 - 125,000
Staff Software Engineer – ML Inference Runtime
Staff Software Engineer – ML Inference Runtime

Arm • Galway

Hybrid
EUR 98,000 - 132,000
Senior Software Engineer – ML Inference Runtime
Senior Software Engineer – ML Inference Runtime

Arm • Galway

Hybrid
EUR 77,000 - 104,000
Staff Software Engineer – ML Inference Runtime
Staff Software Engineer – ML Inference Runtime

Arm Limited • Galway

Hybrid
EUR 90,000 - 130,000
Senior Software ML Engineer
Senior Software ML Engineer

Arm Limited • Galway

Hybrid
EUR 70,000 - 110,000
Staff Software ML Engineer
Staff Software ML Engineer

Arm • Galway

Hybrid
EUR 97,000 - 133,000
Training and professional development
Friendly working environment
Flexibility in hybrid work
Staff Software ML Engineer
Staff Software ML Engineer

Arm Limited • Galway

Hybrid
EUR 65,000 - 95,000
Senior ML Systems Engineer: Arm Acceleration & LLM
Senior ML Systems Engineer: Arm Acceleration & LLM

NLP PEOPLE • Galway

On-site
EUR 120,000 - 160,000
Staff ML Frameworks Engineer - LLM & Acceleration
Staff ML Frameworks Engineer - LLM & Acceleration

Arm Limited • Galway

Hybrid
EUR 90,000 - 130,000
ML Inference Engineer, C++/Python, Arm Acceleration
ML Inference Engineer, C++/Python, Arm Acceleration

Arm Limited • Galway

Hybrid
EUR 85,000 - 125,000