Senior ML Ops Engineer — Remote, Scalable GPU Inference

Pragmatike

Italia

In loco

EUR 90.000 - 120.000

Tempo pieno

3 giorni fa
Candidati tra i primi

Ricevi più risposte dai datori di lavoro

Invia un CV specifico per questa offerta in pochi minuti.

Descrizione del lavoro

Pragmatike is seeking a skilled ML Ops Engineer to design, build and operate scalable model serving for AI workloads in a fully remote, Europe/EMEA-based role. You will collaborate with infra, platform and AI teams to ensure low latency, high availability and cost-efficient inference across distributed GPU systems.

Responsibilities include deploying robust pipelines, auto-scaling, multi-model serving, and observability.

Competenze

  • 4+ years in ML Ops or similar infrastructure roles.
  • Hands-on with model serving frameworks and production systems.
  • Production-grade GPU workloads experience and orchestration.
  • Python and infra-as-code tools (Terraform/Helm).
  • Strong distributed systems knowledge and reliability focus.
  • Ability to work independently in a remote-first environment.

Mansioni

  • Build and operate production-grade model serving infrastructure using frameworks such as vLLM, TGI, Triton, or equivalent.
  • Design and implement robust deployment pipelines with blue/green and canary rollouts for ML models.
  • Develop and maintain auto-scaling systems, multi-model serving architectures, and intelligent request routing layers.
  • Optimize GPU utilization, memory efficiency, network throughput, and model artifact storage performance.
  • Design observability systems for tracking inference latency, throughput, GPU usage, cost metrics, and system health.
  • Manage model registries and CI/CD pipelines enabling automated and reproducible model deployments.
  • Own the full lifecycle of ML systems from development through production, including operational support and on‑call responsibilities.
  • Define engineering best practices and contribute to platform scalability in a fast‑moving startup environment.

Conoscenze

ML Ops
Model Serving
Kubernetes
Terraform
Python
CI/CD for ML
Distributed Systems
Remote Ownership
Remote Work

Strumenti

Kubeflow
MLflow
KubeAI

Descrizione del lavoro

Pragmatike is seeking a skilled ML Ops Engineer to design, build and operate scalable model serving for AI workloads in a fully remote, Europe/EMEA-based role. You will collaborate with infra, platform and AI teams to ensure low latency, high availability and cost-efficient inference across distributed GPU systems.

Responsibilities include deploying robust pipelines, auto-scaling, multi-model serving, and observability.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Senior ML Engineer: Remote Work & Growth Opportunities
Senior ML Engineer: Remote Work & Growth Opportunities

Prima • Milano

Ibrido
EUR 55.000 - 85.000
Staff ML Platform Engineer — Scale GPU Pipelines
Staff ML Platform Engineer — Scale GPU Pipelines

Qualcomm • Roma

Ibrido
EUR 50.000 - 80.000
Senior ML Backend Engineer - Scalable ML Infra (Italy)
Senior ML Backend Engineer - Scalable ML Infra (Italy)

Jobgether • Italia

In loco
EUR 90.000 - 150.000
Engineering Manager, Compute Platform (Remote EU TZ)
Engineering Manager, Compute Platform (Remote EU TZ)

Pragmatike • Italia

In loco
EUR 80.000 - 100.000
ML Platform Software Engineer - Qualcomm, flexible on location anywhere in Europe
ML Platform Software Engineer - Qualcomm, flexible on location anywhere in Europe

Qualcomm • Roma

Ibrido
EUR 50.000 - 80.000
AI Platform Engineer: Build Scalable ML Pipelines
AI Platform Engineer: Build Scalable ML Pipelines

Reply • Torino

In loco
EUR 50.000 - 75.000
Career path
Continuous learning
Reply benefits
+1
GenAI Platforms Engineer — Real-Time ML Systems (Remote)
GenAI Platforms Engineer — Real-Time ML Systems (Remote)

Skillvue • Milano

Remoto
EUR 70.000 - 90.000
Senior ML Engineer — GenAI & Production ML Lead
Senior ML Engineer — GenAI & Production ML Lead

NTT DATA Europe & Latam • Bari

In loco
EUR 90.000 - 120.000
MLOps Platform Engineer: Build Scalable AI Infra (Hybrid)
MLOps Platform Engineer: Build Scalable AI Infra (Hybrid)

TXT GROUP • Cologno Monzese

Ibrido
EUR 35.000 - 40.000
Hybrid working mode
Career opportunities
Continuous training
+3
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Skillvue • Milano

Remoto
EUR 70.000 - 90.000
Competitive compensation
Flexible work
Budget for conferences and training
+1