Senior ML Ops Engineer - Remote GPU Inference

Pragmatike

Madrid

Presencial

EUR 90.000 - 130.000

Jornada completa

hace 9 horas
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Pragmatike is recruiting for an ML Ops Engineer to design, build and operate scalable inference platforms powering AI workloads in a remote-first setup across the EMEA region. You will collaborate with infrastructure, platform, and applied AI teams to ensure low latency and cost-efficient deployments.

The role emphasizes production-grade model serving, distributed GPU systems, and ownership of the full ML lifecycle from development to production in a fast-growing startup context.

Formación

  • 4+ years of experience in ML Ops, Platform Engineering, SRE, or similar roles.
  • Hands-on experience with model serving frameworks such as vLLM, TGI, Triton, or equivalents.
  • Strong background in container orchestration and GPU-based production workloads.
  • Experience with MLOps tooling including model registries, experiment tracking, and automated deployment pipelines.
  • Proficiency in Python and infrastructure-as-code tools (Terraform, Helm, or similar).
  • Strong understanding of distributed systems, performance tuning, and reliability engineering.
  • Ownership mindset with ability to work autonomously in a remote-first environment.

Responsabilidades

  • Build and operate production-grade ML inference platforms powering real-time AI applications.
  • Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models.
  • Develop and maintain auto-scaling systems and intelligent request routing layers.
  • Optimize GPU utilization, memory efficiency, and model artifact storage performance.
  • Define observability systems for tracking latency, throughput, GPU usage, cost metrics, and system health.
  • Manage model registries and CI/CD pipelines enabling automated deployments.
  • Own the full lifecycle of ML systems from development through production, including on-call responsibilities.
  • Define engineering best practices and contribute to platform scalability in a fast-moving startup environment.

Conocimientos

ML Ops
Model serving
Container orchestration
Python
Infrastructure as code
Distributed systems
Remote work

Herramientas

Terraform
Helm

Descripción del empleo

Pragmatike is recruiting for an ML Ops Engineer to design, build and operate scalable inference platforms powering AI workloads in a remote-first setup across the EMEA region. You will collaborate with infrastructure, platform, and applied AI teams to ensure low latency and cost-efficient deployments.

The role emphasizes production-grade model serving, distributed GPU systems, and ownership of the full ML lifecycle from development to production in a fast-growing startup context.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Remote AI Infrastructure Engineer – GPU ML Ops (EMEA)
Remote AI Infrastructure Engineer – GPU ML Ops (EMEA)

Pragmatike • Madrid

Presencial
EUR 60.000 - 80.000
AI Infrastructure Engineer (GPU) - Remote EMEA
AI Infrastructure Engineer (GPU) - Remote EMEA

Pragmatike • Madrid

Presencial
EUR 60.000 - 80.000
Work from home flexibility
Inclusive recruitment process
Opportunity to influence core engineering decisions
Engineering Manager, Cloud Platform & Compute (Remote)
Engineering Manager, Cloud Platform & Compute (Remote)

Pragmatike • Madrid

A distancia
EUR 110.000 - 160.000
Senior Applied Research Engineer | Barcelona, Spain, Hybrid
Senior Applied Research Engineer | Barcelona, Spain, Hybrid

SGI • Barcelona

Híbrido
EUR 90.000 - 130.000
Equity
Relocation support
Hybrid work
+1
Compute Platform Engineering Manager (Remote Europe)
Compute Platform Engineering Manager (Remote Europe)

Pragmatike • Madrid

A distancia
EUR 75.000 - 95.000
Remote work flexibility
Equal Opportunity Employer
Inclusive hiring process
Senior Applied Research Engineer | Barcelona | Up to €150k
Senior Applied Research Engineer | Barcelona | Up to €150k

Source Technology Limited • Madrid

Presencial
EUR 120.000 - 150.000
Competitive compensation
Equity
Comprehensive benefits
+1
Senior ML Engineer, AI Infrastructure at Scale
Senior ML Engineer, AI Infrastructure at Scale

IFS • Barcelona

Híbrido
EUR 55.000 - 90.000
Hybrid work opportunities
Inclusive workplace culture
Senior MLOps Engineer (Training & Inference Optimization)
Senior MLOps Engineer (Training & Inference Optimization)

multiversecomputing • Donostia/San Sebastián

Presencial
EUR 70.000 - 110.000
Indefinite contract
Equal pay guaranteed
Variable performance bonus
+8
Senior SRE - AI GPU Infra & HPC Orchestrator (Remote EU)
Senior SRE - AI GPU Infra & HPC Orchestrator (Remote EU)

Hamilton Barnes ? • España

Presencial
EUR 120.000 - 190.000
Senior ML Engineer - Scalable AI Infra & Pipelines
Senior ML Engineer - Scalable AI Infra & Pipelines

IFS • Madrid

Presencial
EUR 45.000 - 65.000