Real-Time Site Reliability Engineer — Production Uptime & Scale

Alcor

Kraków

On-site

PLN 250,000 - 380,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Alcor is hiring a Site Reliability Engineer to own production reliability for a real-time platform where uptime and latency are the product. You will manage SLOs, incident response, on-call rotations, and production scaling with a focus on blameless postmortems and fast recovery.

Expect a startup tempo with weekly deploys, 1-week sprints, and a culture that values fault-tolerance and AI-assisted tooling to boost velocity while keeping reliability at the center of every decision.

Qualifications

  • 8+ years operating production systems at scale.
  • Strong Go or Python – you automate reliability.
  • Deep on event-driven and real-time systems reliability.
  • Strong monitoring and uptime mindset with proactive alerting.
  • Good networking understanding (TCP/UDP, TLS, WebSocket, DNS, load balancing).
  • GCP at scale; multi-cloud literacy a plus and multi-tenancy experience.

Responsibilities

  • Own SLOs and error budgets per tenant / service.
  • Incident response and blameless postmortems.
  • Production scaling and capacity planning.
  • Observability depth (p50/p95/p99 per event hop).
  • On-call rotation with DevOps; communicate outages clearly.
  • Deploy-safety collaboration with DevOps and automated rollbacks.

Skills

SRE & incident response
Go
Python
Cloud native / GCP
Observability & monitoring
Automation / runbooks
Systems at scale
Networking fundamentals
Chaos engineering
On-call ownership

Tools

NATS
WebSocket
Streaming pipelines
CI/CD pipelines

Job description

Alcor is hiring a Site Reliability Engineer to own production reliability for a real-time platform where uptime and latency are the product. You will manage SLOs, incident response, on-call rotations, and production scaling with a focus on blameless postmortems and fast recovery.

Expect a startup tempo with weekly deploys, 1-week sprints, and a culture that values fault-tolerance and AI-assisted tooling to boost velocity while keeping reliability at the center of every decision.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliaibility engineer
Site Reliaibility engineer

Alcor • Kraków

On-site
PLN 250,000 - 380,000
Senior Site Reliability Engineer — Remote Incident Leader
Senior Site Reliability Engineer — Remote Incident Leader

Affirm • Poland

Remote
PLN 308,000 - 428,000
Parental benefits
Health care coverage
Flexible Spending Wallets
+2
Platform-Driven Senior DevOps Engineer (Multi-Cloud)
Platform-Driven Senior DevOps Engineer (Multi-Cloud)

Alcor • Kraków

On-site
PLN 180,000 - 320,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Województwo pomorskie

On-site
PLN 80,000 - 120,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Kraków

On-site
PLN 254,000 - 340,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Remote SRE for AI-Ready Analytics Platform
Remote SRE for AI-Ready Analytics Platform

Tier4 • Poland

Remote
PLN 238,000 - 285,000
Fully remote
Cross-border collaboration
Competitive contractor compensation
On-Prem Reliability Engineer
On-Prem Reliability Engineer

OpsMill • Poland

Remote
PLN 180,000 - 300,000
Senior Site Reliability Engineer - Cloud & On-Prem Scale
Senior Site Reliability Engineer - Cloud & On-Prem Scale

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Akamai Technologies • Województwo małopolskie

On-site