Lead SRE Engineer: Kubernetes & Cloud Reliability

Talanto

Kraków

Hybrid

PLN 35,000 - 40,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Talanto is seeking a Principal Site Reliability Engineer to design, implement, and operate scalable systems on Kubernetes (AWS EKS). You will drive reliability, security, and technical direction while collaborating with development teams to embed best practices across the lifecycle.

The role is hands-on and requires real impact on production reliability, incident response, and resilience testing in a fast-paced environment.

Qualifications

  • 7+ years of commercial experience in SRE, systems engineering, infrastructure, or related roles.
  • University degree in Computer Science or a related field.
  • Strong hands-on experience with Kubernetes (AWS EKS), including networking, scaling, and security.
  • Proven experience with Terraform, ArgoCD, and GitHub Actions.

Responsibilities

  • Design, operate, and troubleshoot Kubernetes clusters (AWS EKS) with a focus on networking, scalability, security, and reliability.
  • Architect and maintain highly available, fault-tolerant infrastructure on AWS using Infrastructure as Code (Terraform).
  • Automate provisioning, deployment, and configuration processes following GitOps practices with ArgoCD and GitHub Actions.
  • Define and enforce guardrails for infrastructure, applications, and databases to ensure secure and consistent operations.
  • Implement and maintain monitoring and observability solutions using Prometheus, Grafana, and related tools.
  • Build and evolve CI/CD pipelines and progressive delivery strategies.
  • Collaborate closely with development teams to embed reliability and security best practices throughout the application lifecycle.
  • Participate in incident response, post‑incident reviews, and continuous improvement initiatives, including resilience testing and chaos engineering.
  • Design and manage secure networking solutions, including AWS VPCs, Kubernetes networking, and firewalls.

Skills

SRE
Distributed systems
Automation
Incident management

Education

Bachelor's degree in CS or related field

Tools

Terraform
ArgoCD
GitHub Actions
Prometheus
Grafana

Job description

Talanto is seeking a Principal Site Reliability Engineer to design, implement, and operate scalable systems on Kubernetes (AWS EKS). You will drive reliability, security, and technical direction while collaborating with development teams to embed best practices across the lifecycle.

The role is hands-on and requires real impact on production reliability, incident response, and resilience testing in a fast-paced environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer: Cloud-Native Reliability Architect
Senior Site Reliability Engineer: Cloud-Native Reliability Architect

SolarWinds • Kraków

Hybrid
PLN 200,000 - 360,000
Medical care with Luxmed
Pension plan
English/Polish classes
+1
Senior SRE - Managed Kubernetes Reliability & Scale
Senior SRE - Managed Kubernetes Reliability & Scale

Akamai Technologies GmbH • Kraków

Remote
PLN 100,000 - 130,000
Health benefits
Financial benefits
Family support
+2
Senior SRE: Kubernetes Reliability for Global Platforms
Senior SRE: Kubernetes Reliability for Global Platforms

Software Mind • Kraków

Remote
PLN 180,000 - 240,000
Private healthcare and insurance
Multisport card
Language classes
+3
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 120,000
Health benefits
Financial benefits
Family support
+2
SRE Lead: Reliability, Automation & Observability
SRE Lead: Reliability, Automation & Observability

Thrive IT Systems • Warszawa

Hybrid
PLN 180,000 - 260,000
Senior Site Reliability Engineer — Remote Incident Leader
Senior Site Reliability Engineer — Remote Incident Leader

Affirm • Poland

Remote
PLN 308,000 - 428,000
Parental benefits
Health care coverage
Flexible Spending Wallets
+2
Remote SRE Engineering Lead: Scale & Reliability
Remote SRE Engineering Lead: Scale & Reliability

XTB online investing • Poland

Hybrid
PLN 320,000 - 520,000
Remote work
Private medical care
Group insurance
+5
Senior SRE: Scale AWS, Kubernetes & CI/CD (Remote)
Senior SRE: Scale AWS, Kubernetes & CI/CD (Remote)

Tatari • Poland

On-site
PLN 220,000 - 320,000
Equity package
Quarterly tax assistance
Wellness days
+4
Senior Cloud Reliability Engineer (AWS & Kubernetes)
Senior Cloud Reliability Engineer (AWS & Kubernetes)

StrongSD GmbH • Poland

Remote
PLN 180,000 - 240,000
Medical insurance
Free corporate English classes
Mentorship program
+1
Senior SRE: Kubernetes Production Reliability, Remote
Senior SRE: Kubernetes Production Reliability, Remote

Software Mind • Kraków

On-site
PLN 180,000 - 300,000
Flexible work
Remote work
Global projects
+6