Senior MLOps / ML Platform Engineer

Sigmasoftware2

Poland

On-site

PLN 150,000 - 230,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Sigmasoftware2 is seeking a senior MLOps engineer to build and run scalable ML platforms in production. You will design multi-tenant training and serving environments, implement robust CI/CD for ML workloads, and collaborate with DevOps and SRE teams on infrastructure automation.

Strong Python expertise and hands-on experience with Kubernetes, Docker, MLflow, Kubeflow, and Airflow are essential. You will work in a fast-paced product-driven environment focusing on reliability, observability, and

Qualifications

  • 5+ years of experience in MLOps, ML platform engineering, or infrastructure engineering supporting production ML systems
  • Strong Python skills and experience building platform-level tooling and automation
  • Hands-on experience with Kubernetes and Docker
  • Experience building CI/CD pipelines for ML workloads
  • Hands-on production experience with MLflow, Kubeflow, Airflow, Argo Workflows, Vertex Pipelines, or similar orchestration and ML lifecycle platforms
  • Experience with ML platforms and model lifecycle tools such as Vertex AI, MLflow, or Kubeflow
  • Strong understanding of ML observability including drift detection, train/serve skew monitoring, and incident response
  • Experience designing or supporting multi-tenant ML systems and isolated model environments
  • Experience working with cloud platforms, preferably GCP
  • Experience with infrastructure-as-code tools such as Terraform
  • Experience with Linux environments
  • Understanding of the ML lifecycle and productionization processes
  • Upper-Intermediate English level or higher

Responsibilities

  • Build and maintain ML training orchestration pipelines across hourly, daily, and weekly schedules
  • Implement retries, backfills, and idempotent execution mechanisms
  • Design and support model registry workflows including versioning, lineage, evaluation gates, and promotion processes
  • Develop isolated per-advertiser model environments with namespace and configuration separation
  • Build scalable refresh pipelines and publishing workflows for serving infrastructure
  • Implement shadow mode and champion/challenger deployment strategies
  • Develop monitoring and alerting for ML-specific metrics including feature drift, prediction drift, train/serve skew, and calibration decay
  • Ensure reproducibility of ML workflows using containerized environments, pinned dependencies, and data snapshots
  • Monitor training and scoring costs across tenants
  • Collaborate with DevOps and SRE engineers on CI/CD and infrastructure automation
  • Prepare operational documentation and platform handover materials

Skills

Python
Kubernetes
Docker
CI/CD
MLflow
Kubeflow
Airflow
Argo Workflows
Vertex AI
Multi-tenant ML
Terraform
Linux
GCP
Data pipelines

Job description

  • Build and maintain ML training orchestration pipelines across hourly, daily, and weekly schedules
  • Implement retries, backfills, and idempotent execution mechanisms
  • Design and support model registry workflows including versioning, lineage, evaluation gates, and promotion processes
  • Develop isolated per-advertiser model environments with namespace and configuration separation
  • Build scalable refresh pipelines and publishing workflows for serving infrastructure
  • Implement shadow mode and champion/challenger deployment strategies
  • Develop monitoring and alerting for ML-specific metrics including feature drift, prediction drift, train/serve skew, and calibration decay
  • Ensure reproducibility of ML workflows using containerized environments, pinned dependencies, and data snapshots
  • Monitor training and scoring costs across tenants
  • Collaborate with DevOps and SRE engineers on CI/CD and infrastructure automation
  • Prepare operational documentation and platform handover materials
  • 5+ years of experience in MLOps, ML platform engineering, or infrastructure engineering supporting production ML systems
  • Strong Python skills and experience building platform-level tooling and automation
  • Hands-on experience with Kubernetes and Docker
  • Experience building CI/CD pipelines for ML workloads
  • Hands-on production experience with MLflow, Kubeflow, Airflow, Argo Workflows, Vertex Pipelines, or similar orchestration and ML lifecycle platforms
  • Experience with ML platforms and model lifecycle tools such as Vertex AI, MLflow, or Kubeflow
  • Strong understanding of ML observability including drift detection, train/serve skew monitoring, and incident response
  • Experience designing or supporting multi-tenant ML systems and isolated model environments
  • Experience working with cloud platforms, preferably GCP
  • Experience with infrastructure-as-code tools such as Terraform
  • Experience with Linux environments
  • Understanding of the ML lifecycle and productionization processes
  • Upper-Intermediate English level or higher
WILL BE A PLUS
  • Experience with feature stores and feature consistency management
  • Experience with large-scale batch scoring systems operating under freshness SLAs
  • Familiarity with experiment tracking platforms and evaluation gates
  • Experience with on-premises Kubernetes or bare-metal Linux infrastructure
  • Knowledge of DVC, lakeFS, or other data versioning tools
  • Experience with Bigtable, Redis, Aerospike, or similar low-latency serving databases
  • GPU scheduling and training cost optimization experience
  • Familiarity with SOC 2, ISO 27001, or GDPR-related compliance requirements
PERSONAL PROFILE
  • Strong ownership mindset and focus on operational reliability
  • Ability to work independently in complex distributed systems environments
  • Strong collaboration and communication skills
  • Analytical thinking with attention to scalability and maintainability
  • Comfortable working in fast-paced product-oriented environments
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior MLOps / ML Platform Engineer
Senior MLOps / ML Platform Engineer

Sigma Software • Warszawa

On-site
PLN 180,000 - 320,000
Senior MLOps Engineer (Google Cloud)
Senior MLOps Engineer (Google Cloud)

Spyro Soft • Poland

Remote
USD 180,000 - 230,000
Senior MLOps Platform Engineer: Scale, Automate ML
Senior MLOps Platform Engineer: Scale, Automate ML

Sigmasoftware2 • Poland

On-site
PLN 150,000 - 230,000
MLOps Team Lead
MLOps Team Lead

Fetcherr • Warszawa

On-site
PLN 180,000 - 320,000
Senior Director – Data Platform, Cloud Services
Senior Director – Data Platform, Cloud Services

Jobtailor • Warszawa

On-site
PLN 350,000 - 520,000
Applied Machine Learning Engineer
Applied Machine Learning Engineer

Hitachi Energy • Kraków

On-site
PLN 45,000 - 65,000
Competitive benefits for financial wellbeing
Support for physical and mental wellbeing
Senior Data Engineer
Senior Data Engineer

Sigma Software • Warszawa

On-site
PLN 240,000 - 360,000
DevOps
DevOps

Complexio • Warszawa

On-site
PLN 250,000 - 420,000
Lead ML Architect (Google Cloud)
Lead ML Architect (Google Cloud)

Spyro Soft • Poland

Remote
USD 180,000 - 240,000
Senior MLOps & Observability Engineer | Kubernetes Cloud
Senior MLOps & Observability Engineer | Kubernetes Cloud

CloudFerro Sp. z o • Warszawa

On-site
PLN 180,000 - 270,000