Service Manager & Site Reliability Engineer

GFT Technologies Poland

Łódź

Hybrid

PLN 180,000 - 280,000

Full time

3 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model
Medical & life insurance
Lunch subsidy
Online training

Job summary

GFT Technologies Poland seeks an experienced professional to manage real-time incident response, impact assessment, external communications, and cross‑production coordination for global operations.

Working at the intersection of Site Reliability Engineering, Incident Response, and Partner Operations, you will ensure timely, SLA‑compliant updates and support RCA activities while driving reliability improvements with Python or Kotlin.

Qualifications

  • 5+ years of experience in Incident Operations or SRE.
  • Experience in on-call environments with SLA-driven responsibilities.
  • Strong understanding of distributed systems and production environments.
  • Experience with monitoring, alerting, and incident management tools.
  • Familiarity with APIs, system integrations, and observability platforms.
  • Hands-on experience with Python or Kotlin.
  • Understanding of SDLC and production reliability principles.
  • Strong communication and stakeholder management skills.
  • Ability to manage multiple priorities under pressure.
  • Ownership mindset and cross-functional collaboration.

Responsibilities

  • Monitor and respond to production incidents.
  • Coordinate incident response activities across teams.
  • Assess impact and determine incident severity.
  • Manage external communications and status page updates.
  • Support incident reporting, RCA activities, and SLA tracking.
  • Collaborate with Engineering teams to improve reliability and observability.
  • Drive process improvements and automation initiatives.
  • Contribute to internal reliability tooling using Python or Kotlin.

Skills

Python
Kotlin
Incident management
On-call experience
Distributed systems
Monitoring & alerting
Communication
Stakeholder management
Automation
Platform engineering

Tools

Datadog
Chronosphere
PagerDuty
Rootly
Slack workflows

Job description

You will manage real-time incident response, impact assessment, external communications, and coordination across production systems. Working at the intersection of Site Reliability Engineering, Incident Response, and Partner Operations, you will ensure timely, accurate, and SLA-compliant communication while supporting the scalability and reliability of global operations.

  • Monitor and respond to production incidents
  • Coordinate incident response activities across teams
  • Assess impact and determine incident severity
  • Manage external communications and status page updates
  • Support incident reporting, RCA activities, and SLA tracking
  • Collaborate with Engineering teams to improve reliability and observability
  • Drive process improvements and automation initiatives
  • Contribute to internal reliability tooling using Python or Kotlin
  • 5+ years of experience in Incident Operations, Site Reliability Engineering, Technical Operations, or a similar role
  • Experience working in on-call environments with SLA-driven responsibilities
  • Strong understanding of distributed systems and production environments
  • Experience with monitoring, alerting, and incident management tools
  • Familiarity with APIs, system integrations, and observability platforms
  • Hands‑on experience with Python or Kotlin
  • Understanding of SDLC and production reliability principles
  • Strong communication, stakeholder management, and decision‑making skills
  • Ability to work effectively in high-pressure environments and manage multiple priorities
  • Strong ownership mindset and cross‑functional collaboration skills
Nice to have
  • Experience with Datadog or Chronosphere
  • Experience with PagerDuty, Rootly, or Slack workflows
  • Experience managing external status pages
  • Experience with incident management automation and process improvements
  • Experience contributing to reliability tooling and platform engineering
We offer
  • Hybrid work in one of our locations: Lodz, Poznan, Krakow, Warsaw, Wroclaw (2 office days per week)
  • Working in a highly experienced and dedicated team
  • Benefit package tailored to your needs (medical, sport, lunch subsidy, life insurance, etc.)
  • Online training and certifications
  • Access to e-learning platform
  • Work From Anywhere (up to 140 days/year abroad)
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Service Manager & Site Reliability Consultant
Service Manager & Site Reliability Consultant

GFT Technologies Poland • Wrocław

On-site
PLN 130,000 - 210,000
Service Manager & Site Reliability Consultant
Service Manager & Site Reliability Consultant

GFT TECHNOLOGIES SE • Łódź

Hybrid
PLN 282,000 - 379,000
Site Reliability Engineer
Site Reliability Engineer

SIX Group Services Ltd. • Warszawa

Hybrid
PLN 180,000 - 280,000
Flexible Work Models
Personal Development
Agile Working Methods
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Wrocław

On-site
PLN 260,000 - 360,000
Great Place to Work
Profit sharing
Medical care
+3
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Piła

On-site
PLN 230,000 - 350,000
Great Place to Work
Profit sharing
Internal trainings
+2
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Lublin

On-site
PLN 240,000 - 380,000
Medical care
Profit sharing
Internal trainings
+1
Senior Site Reliability Engineer - OneRail’s Technology Hub - Poland
Senior Site Reliability Engineer - OneRail’s Technology Hub - Poland

OneRail USA • Kraków

Hybrid
Production Support & Operations Engineer Splunk, Sysdig, Prometheus, Grafana, Kubernetes, Docker Warszawa, Gdańsk
Production Support & Operations Engineer Splunk, Sysdig, Prometheus, Grafana, Kubernetes, Docker Warszawa, Gdańsk

Diverse CG Sp. z o.o. Sp.k. • Województwo pomorskie

Hybrid
PLN 180,000 - 240,000
Private medical care
Co-financing for the sports card
Constant support of dedicated consult
Production Support & Operations Engineer Splunk, Sysdig, Prometheus, Grafana, Kubernetes, Docker Warszawa, Gdańsk
Production Support & Operations Engineer Splunk, Sysdig, Prometheus, Grafana, Kubernetes, Docker Warszawa, Gdańsk

DCG Poland • Województwo pomorskie

Hybrid
PLN 180,000 - 300,000
Private medical care
Co-financing for the sports card
Dedicated consultant
Support Consultant with Kubernetes
Support Consultant with Kubernetes

GFT Technologies Poland • Kraków

Hybrid
PLN 150,000 - 230,000