Senior ML Inference Systems Engineer (Hybrid, GPU Fleet)

TensorX

Dublin

Hybrid

EUR 120,000 - 180,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Hybrid work
25 days paid annual leave
Central Dublin office
Free inference tokens

Job summary

TensorX, a sovereign AI infrastructure platform headquartered in Dublin, is seeking a Senior ML Systems Engineer (Inference). You will own the model serving layer, deploy and optimize LLMs, and plan GPU capacity across on-prem and AWS.

Based in Dublin with hybrid work options, you will work on vLLM, SGLang and TensorRT-LLM, benchmarking models and driving hardware roadmaps. You will collaborate with platform and product teams to raise the bar on reliability and performance.

Qualifications

  • 5+ years of professional experience in ML infrastructure, systems engineering or a closely related discipline.
  • Hands-on experience deploying and tuning large language models with modern inference frameworks.
  • Strong working knowledge of GPU architecture and NVIDIA hardware.
  • Experience with inference optimisation techniques including quantisation and batching.
  • Proficiency in Python and systems-level code.
  • Solid experience with Linux, containers and GPU workloads orchestration.

Responsibilities

  • Deploy and operate open-source LLMs in production using vLLM, SGLang and TensorRT-LLM.
  • Profile and tune inference workloads for latency, throughput and GPU utilisation.
  • Benchmark new open-weight models and inform productisation decisions.
  • Lead GPU procurement and future hardware roadmap decisions.
  • Build and maintain the model fleet infrastructure across on-prem and AWS.
  • Instrument the serving stack with metrics and drive incident response.

Skills

ML infrastructure
GPGPU knowledge
Python
Linux
Container orchestration
Benchmarking
CUDA basics
Inference frameworks
Communication skills
Ambiguity tolerance

Education

BSc/MSc in CS or related

Tools

vLLM
SGLang
TensorRT-LLM
CUDA

Job description

TensorX, a sovereign AI infrastructure platform headquartered in Dublin, is seeking a Senior ML Systems Engineer (Inference). You will own the model serving layer, deploy and optimize LLMs, and plan GPU capacity across on-prem and AWS.

Based in Dublin with hybrid work options, you will work on vLLM, SGLang and TensorRT-LLM, benchmarking models and driving hardware roadmaps. You will collaborate with platform and product teams to raise the bar on reliability and performance.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Systems Engineer (Inference) – Hybrid + Tokens
Senior ML Systems Engineer (Inference) – Hybrid + Tokens

TensorX • Dublin

Hybrid
EUR 120,000 - 180,000
Hybrid working from Dublin office
25 days paid annual leave
Free inference tokens
Senior ML Inference Systems Engineer (Hybrid, On-Prem GPU)
Senior ML Inference Systems Engineer (Hybrid, On-Prem GPU)

Uniting Holding • Dublin

Hybrid
EUR 80,000 - 120,000
25 days paid annual leave
Free inference tokens
Remote work flexibility
Senior ML Systems Engineer (Inference)
Senior ML Systems Engineer (Inference)

Uniting Holding • Dublin

Hybrid
EUR 80,000 - 120,000
25 days paid annual leave
Free inference tokens
Remote work flexibility
Senior Machine Learning Engineer (Inference)
Senior Machine Learning Engineer (Inference)

TensorX • Dublin

Hybrid
EUR 120,000 - 180,000
Hybrid working from Dublin office
25 days paid annual leave
Free inference tokens
Engineering Manager, AI Platform & Delivery (Hybrid Dublin)
Engineering Manager, AI Platform & Delivery (Hybrid Dublin)

TensorX • Dublin

Hybrid
EUR 120,000 - 180,000
Hybrid working from Dublin office
25 days paid annual leave
Free inference tokens
Senior ML Systems Engineer (Inference
Senior ML Systems Engineer (Inference

TensorX • Dublin

Hybrid
EUR 120,000 - 180,000
Hybrid work
25 days paid annual leave
Central Dublin office
+1
Hybrid/Remote Senior AI Inference Engineer
Hybrid/Remote Senior AI Inference Engineer

F5 Networks, Inc.  • Dublin

Hybrid
EUR 120,000 - 180,000
Senior Full-Stack Engineer - Hybrid Dublin
Senior Full-Stack Engineer - Hybrid Dublin

TensorX • Dublin

Hybrid
EUR 90,000 - 120,000
Hybrid working from Dublin office
25 days paid annual leave
Free inference tokens
Hybrid AI Platform Engineering Manager — Lead, Hire & Ship
Hybrid AI Platform Engineering Manager — Lead, Hire & Ship

Uniting Holding • Dublin

Hybrid
EUR 120,000 - 180,000
Hybrid Dublin office
Free inference tokens
25 days annual leave
Senior LLM Inference & GPU Performance Engineer
Senior LLM Inference & GPU Performance Engineer

Confidential • Ireland

On-site
EUR 70,000 - 90,000