Senior Platform SRE: Reliability & Observability Architect

IG KnowHow

Kraków

Hybrid

PLN 240,000 - 420,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Growth opportunities
Mentoring programs
Networking clubs
Volunteer time off
Perks details

Job summary

IG Group is seeking a Senior Platform SRE to own and improve the reliability platform within a hybrid AWS/on-premises environment. The role focuses on observability, SLOs, and automated resilience across hundreds of services and engineers.

You will mentor others, drive incident prevention, and contribute to IG’s reliability standards. The position requires hands-on engineering, chaos testing, and strong collaboration with Platform Engineering and product teams in a fast-paced fintech setting.

Qualifications

  • Hands-on observability and instrumentation with OpenTelemetry and production use of monitoring tools.
  • Design and maintain SLOs, error budgets, and multi-window burn-rate alerts.
  • Build and operate CI/CD pipelines with blue/green/canary releases and automated rollback.
  • Experience with Kubernetes and Nomad in a hybrid cloud environment.
  • Production-quality coding in Java and/or Python.
  • Strong knowledge of distributed systems and failure modes.
  • Experience with incident management and blameless post-incident reviews.
  • Familiarity with chaos engineering practices and tooling.
  • Ability to write RFCs and contribute to engineering standards.

Responsibilities

  • Build and own the reliability platform and tooling.
  • Implement comprehensive monitoring and observability across services.
  • Maintain production readiness including automated deployments and zero-downtime patching.
  • Design self-healing capabilities and automated traffic rerouting.
  • Run chaos experiments to improve resilience and incident response.
  • Develop CI/CD pipelines embedding reliability practices.
  • Mentor engineers on reliability patterns and production engineering discipline.
  • Collaborate with platform teams to define SLOs on customer journeys.

Skills

Observability engineering
SLOs and error budgets
CI/CD and release engineering
Container orchestration
Software engineering (Java/Python)
Distributed systems
Incident management
Chaos engineering
Standards and governance

Tools

OpenTelemetry
Datadog
Dynatrace
Grafana
Kubernetes (EKS/AKS/GKE)
HashiCorp Nomad
Terraform
PagerDuty/ServiceNow

Job description

IG Group is seeking a Senior Platform SRE to own and improve the reliability platform within a hybrid AWS/on-premises environment. The role focuses on observability, SLOs, and automated resilience across hundreds of services and engineers.

You will mentor others, drive incident prevention, and contribute to IG’s reliability standards. The position requires hands-on engineering, chaos testing, and strong collaboration with Platform Engineering and product teams in a fast-paced fintech setting.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

IG Group • Kraków

Hybrid
PLN 250,000 - 420,000
Career development
Mentoring programs
Networking clubs
+3
Senior Platform SRE
Senior Platform SRE

IG Group • Kraków

On-site
PLN 250,000 - 420,000
Career development
Mentoring programs
Networking clubs
+3
Senior Platform SRE
Senior Platform SRE

IG KnowHow • Kraków

Hybrid
PLN 240,000 - 420,000
Growth opportunities
Mentoring programs
Networking clubs
+2
Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Staff SRE: Scalable Infra, Observability & Automation
Staff SRE: Scalable Infra, Observability & Automation

Assured • Poland

Hybrid
PLN 80,000 - 120,000
Competitive salary and equity packages
Platinum medical, dental, and vision healthcare plan
Unlimited PTO
+4
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Województwo pomorskie

On-site
PLN 80,000 - 120,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Kraków

On-site
PLN 254,000 - 340,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior SRE: Cloud-Native Reliability & Automation
Senior SRE: Cloud-Native Reliability & Automation

OneRail USA • Kraków

Hybrid
Remote Senior SRE: Cloud Platform & Automation Lead
Remote Senior SRE: Cloud Platform & Automation Lead

FYUL • Warszawa

Hybrid
PLN 240,000 - 420,000
Private health insurance
Flexible hours
Remote work
SRE Engineer: Observability, AI-Driven Reliability
SRE Engineer: Observability, AI-Driven Reliability

Citibank (Switzerland) AG • Warszawa

Hybrid
Confidential
Pension plan
Private medical care
Life insurance
+1