Observability Platform Engineer for Large-Scale Infra

Mirantis

Poland

On-site

PLN 325,000 - 411,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive pay
Professional development
Conferences
Open-source culture
Team events

Job summary

Mirantis is building out our Neocloud service—managing large-scale infrastructure to a high SLA for demanding compute workloads. We are seeking an Observability Platform Engineer to design and build the monitoring, logging, tracing, and alerting platform that our operations teams depend on to detect incidents fast at scale.

This is a hands-on role: you’ll turn unclear telemetry into concrete signals, define SLO/SLI frameworks, and own the reliability of the observability stack while

Qualifications

  • Proven experience designing and building observability platforms for large-scale, production environments.
  • Hands-on with metrics, logging, and distributed tracing tooling (Prometheus, Grafana, OpenTelemetry, Loki, Jaeger, Tempo etc).
  • Experience with high-volume telemetry pipelines and tradeoffs (cardinality, retention, cost, query latency).
  • Strong software engineering skills in Go, Python, or Rust.
  • Experience with Kubernetes and cloud-native infrastructure.
  • Solid understanding of SLO/SLI and alerting design to minimize noise.
  • Excellent communication with operations and delivery teams.

Responsibilities

  • Design, build, and operate observability components for large-scale infrastructure.
  • Build telemetry pipelines capable of handling high cardinality, high volume data with cost and retention considerations.
  • Define and implement SLO/SLI frameworks and alerting strategies to reduce noise.
  • Collaborate with service delivery and operations to understand incident visibility needs.
  • Integrate observability tooling with incident management workflows including RCA data.
  • Improve detection speed and reduce MTTR/MTTD across the platform.
  • Contribute to AI-assisted operations tooling roadmap as it matures.
  • Own reliability, scalability, and security of the observability stack.
  • Document architecture and runbooks for maintainability.

Skills

Observability platform design
Incident response collaboration
Go
Python
Rust
Strong communication skills
SLO/SLI design

Tools

Prometheus
Grafana
OpenTelemetry
Loki
Jaeger
Thanos/Cortex/Mimir
Elasticsearch/OpenSearch
Kubernetes
eBPF tooling

Job description

Mirantis is building out our Neocloud service—managing large-scale infrastructure to a high SLA for demanding compute workloads. We are seeking an Observability Platform Engineer to design and build the monitoring, logging, tracing, and alerting platform that our operations teams depend on to detect incidents fast at scale.

This is a hands-on role: you’ll turn unclear telemetry into concrete signals, define SLO/SLI frameworks, and own the reliability of the observability stack while

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Observability Platform Engineer — Scale & Incident Readiness
Observability Platform Engineer — Scale & Incident Readiness

Mirantis • Warszawa

On-site
PLN 190,000 - 280,000
Observability Platform Engineer — Neocloud
Observability Platform Engineer — Neocloud

Mirantis • Warszawa

On-site
PLN 190,000 - 280,000
Observability Platform Engineer - Neocloud
Observability Platform Engineer - Neocloud

Mirantis • Poland

On-site
PLN 325,000 - 411,000
Competitive pay
Professional development
Conferences
+2
Remote Cloud Infrastructure Service Delivery Manager
Remote Cloud Infrastructure Service Delivery Manager

Mirantis • Poznań

Remote
PLN 180,000 - 240,000
Professional development
Conference attendance
Company events
+1
Remote Observability Platform Engineer for Scalable Systems
Remote Observability Platform Engineer for Scalable Systems

Whatnot Inc. • Kraków

Hybrid
PLN 240,000 - 360,000
Senior Software Engineer - Bare-Metal Cloud & Networking
Senior Software Engineer - Bare-Metal Cloud & Networking

Mirantis • Poland

On-site
PLN 240,000 - 420,000
Remote Observability Engineer - Scale & Reliability
Remote Observability Engineer - Scale & Reliability

Whatnot • Kraków

Hybrid
PLN 520,000 - 580,000
Senior AI Infra & Platform Reliability Engineer
Senior AI Infra & Platform Reliability Engineer

Mirantis • Poznań

On-site
PLN 180,000 - 300,000
Senior Infra Software Engineer - GPU Provisioning & Kubernetes
Senior Infra Software Engineer - GPU Provisioning & Kubernetes

Mirantis • Poznań

On-site
PLN 180,000 - 300,000
Remote AI Infrastructure & Platform Operations Engineer
Remote AI Infrastructure & Platform Operations Engineer

Mirantis • Poznań

Remote
PLN 180,000 - 240,000