Senior ML Ops Engineer - Remote GPU Inference

Pragmatike

Madrid

On-site

EUR 90,000 - 130,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Pragmatike is recruiting for an ML Ops Engineer to design, build and operate scalable inference platforms powering AI workloads in a remote-first setup across the EMEA region. You will collaborate with infrastructure, platform, and applied AI teams to ensure low latency and cost-efficient deployments.

The role emphasizes production-grade model serving, distributed GPU systems, and ownership of the full ML lifecycle from development to production in a fast-growing startup context.

Qualifications

  • 4+ years of experience in ML Ops, Platform Engineering, SRE, or similar roles.
  • Hands-on experience with model serving frameworks such as vLLM, TGI, Triton, or equivalents.
  • Strong background in container orchestration and GPU-based production workloads.
  • Experience with MLOps tooling including model registries, experiment tracking, and automated deployment pipelines.
  • Proficiency in Python and infrastructure-as-code tools (Terraform, Helm, or similar).
  • Strong understanding of distributed systems, performance tuning, and reliability engineering.
  • Ownership mindset with ability to work autonomously in a remote-first environment.

Responsibilities

  • Build and operate production-grade ML inference platforms powering real-time AI applications.
  • Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models.
  • Develop and maintain auto-scaling systems and intelligent request routing layers.
  • Optimize GPU utilization, memory efficiency, and model artifact storage performance.
  • Define observability systems for tracking latency, throughput, GPU usage, cost metrics, and system health.
  • Manage model registries and CI/CD pipelines enabling automated deployments.
  • Own the full lifecycle of ML systems from development through production, including on-call responsibilities.
  • Define engineering best practices and contribute to platform scalability in a fast-moving startup environment.

Skills

ML Ops
Model serving
Container orchestration
Python
Infrastructure as code
Distributed systems
Remote work

Tools

Terraform
Helm

Job description

Pragmatike is recruiting for an ML Ops Engineer to design, build and operate scalable inference platforms powering AI workloads in a remote-first setup across the EMEA region. You will collaborate with infrastructure, platform, and applied AI teams to ensure low latency and cost-efficient deployments.

The role emphasizes production-grade model serving, distributed GPU systems, and ownership of the full ML lifecycle from development to production in a fast-growing startup context.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote AI Infrastructure Engineer – GPU ML Ops (EMEA)
Remote AI Infrastructure Engineer – GPU ML Ops (EMEA)

Pragmatike • Madrid

On-site
EUR 60,000 - 80,000
Work from home flexibility
Inclusive recruitment process
Opportunity to influence core engineering decisions
AI Infrastructure Engineer (GPU) - Remote EMEA
AI Infrastructure Engineer (GPU) - Remote EMEA

Pragmatike • Madrid

On-site
EUR 60,000 - 80,000
Work from home flexibility
Inclusive recruitment process
Opportunity to influence core engineering decisions
Compute Platform Engineering Manager (Remote Europe)
Compute Platform Engineering Manager (Remote Europe)

Pragmatike • Madrid

Remote
EUR 75,000 - 95,000
Remote work flexibility
Equal Opportunity Employer
Inclusive hiring process
Machine Learning Engineer (Governance ML Platform)
Machine Learning Engineer (Governance ML Platform)

Lever, Inc. • Spain

Remote
EUR 85,000 - 110,000
Remote work
CuraLinc health resources
Wellness Fridays
+2
Remote Technical Lead - GPU Infrastructure Platform
Remote Technical Lead - GPU Infrastructure Platform

Tether • Barcelona

On-site
EUR 90,000 - 130,000
AI/ML Platform Engineer — Hybrid & Remote
AI/ML Platform Engineer — Hybrid & Remote

Prima • Madrid

Hybrid
EUR 40,000 - 75,000
Private healthcare
Gym discounts
Wellbeing programs
+5
Senior MLOps Engineer (Training & Inference Optimization)
Senior MLOps Engineer (Training & Inference Optimization)

multiversecomputing • Donostia/San Sebastián

On-site
EUR 70,000 - 110,000
Indefinite contract
Equal pay guaranteed
Variable performance bonus
+8
Senior ML Engineer (Applied AI)
Senior ML Engineer (Applied AI)

Ninetwothree Ai Studio • Spain

On-site
USD 150,000 - 190,000
Remote work
Senior ML Platform Engineer - Scale & GenAI
Senior ML Platform Engineer - Scale & GenAI

Preply • Barcelona

On-site
EUR 80,000 - 110,000
Generous learning allowance
Health insurance
Access to mental health support
AI/ML Engineer (Remote)
AI/ML Engineer (Remote)

Quik Hire Staffing • Spain

On-site
EUR 60,000 - 90,000