SRE: Reliability, Observability & Cloud Infra

HeadHR

Wrocław

Hybrid

PLN 180,000 - 240,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

HeadHR in Wrocław, Poland, is seeking a Service Reliability Engineer to define SLIs/SLOs and improve reliability across Java/Spring Boot, Python and Node.js microservices. You will optimize latency and memory, perform JVM diagnostics, and implement circuit breakers and retries.

You will build Datadog observability and operate services on AWS with EKS/ECS, API Gateway, RDS and more. Strong Linux knowledge and CI/CD experience are essential.

Qualifications

  • At least 3 years of experience as a Service Reliability Engineer.
  • Strong hands-on experience with Java and Spring Boot.
  • Good proficiency in Python or another backend programming language.
  • Knowledge of AWS, including EKS, ECS, IAM, VPC, RDS, S3 and API Gateway.
  • Hands-on Datadog experience with APM, Logs, dashboards, SLOs, RUM and Synthetic Monitoring.
  • Docker, Kubernetes, SQL Server or another SQL dialect, and MongoDB indexes, replica sets and performance tuning.
  • CI/CD with GitHub Actions, Jenkins and GitLab CI; queues and messaging with SQS, SNS or Kafka.
  • Basic Linux and networking fundamentals, including HTTP, TLS, DNS and load balancing.

Responsibilities

  • Define SLIs and SLOs and improve reliability across Java/Spring Boot, Python and Node.js microservices.
  • Optimize latency, throughput, memory and GC; perform JVM diagnostics; implement circuit breakers, retries and graceful degradation.
  • Build Datadog observability covering metrics, logs, traces, synthetics, RUM, APM and dashboards.
  • Triage incidents, restore services, communicate proactively, lead post-incident reviews, perform root cause analysis and track remediation.
  • Operate services on EKS and ECS and manage AWS API Gateway, Lambda, RDS, IAM, S3, VPC, CloudWatch, ALB and NLB.
  • Maintain CI/CD pipelines using GitHub Actions, Jenkins and GitLab CI; build Python and shell automation.
  • Troubleshoot MSSQL/SQL and MongoDB, including replica sets, indexes and PITR, and ensure reliable backup, restore, failover and performance tuning.

Skills

Java
Spring Boot
Python

Tools

Datadog
Docker
Kubernetes
GitHub Actions
Jenkins
GitLab CI
SQS
SNS
Kafka
SQL Server
MongoDB
AWS

Job description

HeadHR in Wrocław, Poland, is seeking a Service Reliability Engineer to define SLIs/SLOs and improve reliability across Java/Spring Boot, Python and Node.js microservices. You will optimize latency and memory, perform JVM diagnostics, and implement circuit breakers and retries.

You will build Datadog observability and operate services on AWS with EKS/ECS, API Gateway, RDS and more. Strong Linux knowledge and CI/CD experience are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE: Kubernetes Reliability & Observability
SRE: Kubernetes Reliability & Observability

Software Mind • Województwo małopolskie

On-site
PLN 120,000 - 180,000
Flexible employment and remote work
International projects with leading全球?
International business trips
+6
SRE Lead: Drive Reliability, CI/CD & Observability
SRE Lead: Drive Reliability, CI/CD & Observability

Us3 Consulting • Warszawa

On-site
PLN 300,000 - 460,000
SRE Manager: Cloud Reliability & AI-Driven Ops
SRE Manager: Cloud Reliability & AI-Driven Ops

3003 Sabre Polska Sp. z o.o. • Kraków

Hybrid
PLN 260,000 - 380,000
Paid time off
Year-End-Break: fully paid days off
Paid parental leave: up to 12 weeks
+3
Senior AWS SRE - Hybrid Role in Łódź, Poland
Senior AWS SRE - Hybrid Role in Łódź, Poland

IDEMIA • Łódź

Hybrid
PLN 210,000 - 260,000
Hybrid work model
Office in Łódź
Service Reliability Engineer (K/M)
Service Reliability Engineer (K/M)

HeadHR • Wrocław

Hybrid
PLN 180,000 - 240,000
SRE: Infra Automation & Reliability Engineer
SRE: Infra Automation & Reliability Engineer

CoreWeave • Warszawa

On-site
PLN 223,000 - 298,000
Living Wage Accredited Employer
SRE Engineer: Cloud, CI/CD & Observability
SRE Engineer: Cloud, CI/CD & Observability

Genuine Parts Company • Kraków

Hybrid
PLN 180,000 - 280,000
SRE Engineer - Observability, AI & Automation
SRE Engineer - Observability, AI & Automation

Citi • Warszawa

Hybrid
PLN 165,000 - 281,000
Pension plan
Private medical care
Life insurance
+3
Cloud SRE: 24/7 Reliability, Automation & Incident Response
Cloud SRE: 24/7 Reliability, Automation & Incident Response

IBM Computing • Kraków

On-site
PLN 180,000 - 240,000
SRE Engineer - Observability & AI-Driven Reliability
SRE Engineer - Observability & AI-Driven Reliability

Citigroup Inc. • Warszawa

Hybrid
PLN 165,000 - 281,000
Pension Plan
Private Medical Care
Life Insurance
+5