Cloud SRE: Resiliency, Observability & 24/7 On-Call

XM

Poland

On-site

PLN 180,000 - 240,000

Full time

39 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive remuneration package
International training opportunities
Confidential recruitment

Job summary

XM is seeking a senior SRE/Resiliency Engineer to drive reliability initiatives in a cloud-native environment. You will lead chaos experiments, design observability, and oversee migrations with strong Kubernetes and IaC skills.

Responsibilities include mentoring teammates, on-call rotations, and ensuring compliance with resiliency standards across cloud deployments. The role requires extensive AWS experience and a deep understanding of SRE practices.

Qualifications

  • BSc/MSc degree in Computer Science or related field.
  • 5+ years of cloud services experience, with at least 3 years on AWS cloud.
  • 3+ years of experience in SRE or a similar role.
  • Experience with monitoring, APM, logging, and notification tools.
  • Familiarity with incident, problem and change management procedures and practices.
  • Advanced knowledge of SRE practices and methods.
  • Understanding and practice of Service Levels.
  • Strong troubleshooting skills and the ability to mentor others.
  • Extensive experience with Kubernetes and related technologies, services, and ecosystem.
  • Advanced knowledge of CI/CD, Infrastructure as Code (IaC) concepts and tools, especially HCL Terraform and AWS CloudFormation.
  • Experience with versioning tools like Git.
  • Strong organizational and documentation skills.
  • Exceptional time management and research abilities.
  • Advanced Linux, networking, and scripting skills.

Responsibilities

  • Honor and practice the Resiliency pillar of the Well Architected Framework in all tasks and responsibilities.
  • Conduct Chaos Engineering experiments and relevant exercises to improve resiliency and fault-tolerance.
  • Research workloads for migrating to the cloud with minimal disruption and impact.
  • Monitor cloud migration projects to ensure seamless transitions.
  • Design, consult, re-platform, and re-factor the observability of current cloud infrastructure.
  • Coordinate with other IT departments and teams regarding observability for both individual and organizational needs.
  • Regularly assess cloud deployments for compliance with the company’s standards and best practices.
  • Investigate and correct areas where observability is lagging.
  • Stay up to date and provide training on new and current technologies, services, tools, methodologies, and practices.
  • Occasionally participate in service capacity planning, software performance analysis, and system tuning.
  • Mentor colleagues in technical skills and knowledge.
  • Analyze, oversee, and remediate the company’s resiliency.
  • Participate in on-call support 24/7 based on a rotation schedule

Skills

Cloud experience
AWS
SRE
Kubernetes
Terraform
CloudFormation
CI/CD
IaC
APM/Monitoring
Linux
Networking
Troubleshooting
Mentoring

Education

BSc/MSc in Computer Science or related field

Tools

Kafka (MSK)
Postgres
MySQL
Python
Go
Git

Job description

XM is seeking a senior SRE/Resiliency Engineer to drive reliability initiatives in a cloud-native environment. You will lead chaos experiments, design observability, and oversee migrations with strong Kubernetes and IaC skills.

Responsibilities include mentoring teammates, on-call rotations, and ensuring compliance with resiliency standards across cloud deployments. The role requires extensive AWS experience and a deep understanding of SRE practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: Multi-Cloud Reliability & Automation Lead
Senior SRE: Multi-Cloud Reliability & Automation Lead

Square One Resources Limited • Poland

Remote
PLN 180,000 - 300,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

XM • Poland

On-site
PLN 180,000 - 240,000
Competitive remuneration package
International training opportunities
Confidential recruitment
Senior Cloud Reliability Engineer (AWS & Kubernetes)
Senior Cloud Reliability Engineer (AWS & Kubernetes)

StrongSD GmbH • Poland

Remote
PLN 180,000 - 240,000
Medical insurance
Free corporate English classes
Mentorship program
+1
Senior Platform SRE: Reliability & Observability Architect
Senior Platform SRE: Reliability & Observability Architect

IG KnowHow • Kraków

Hybrid
PLN 240,000 - 420,000
Growth opportunities
Mentoring programs
Networking clubs
+2
Remote Site Reliability Engineer - Observability
Remote Site Reliability Engineer - Observability

XTB online investing • Warszawa

Hybrid
PLN 199,000 - 253,000
Training budget for courses
Birthday day off
Parental leave day off
+5
Senior Cloud SRE: Reliability & Observability
Senior Cloud SRE: Reliability & Observability

OpenTalent • Kraków

On-site
PLN 180,000 - 280,000
Relocation assistance
Senior Cloud Networking SRE — Observability & Incident Lead
Senior Cloud Networking SRE — Observability & Incident Lead

Akamai Career Site • Poland

Remote
PLN 180,000 - 280,000
FlexBase adapts to your job's needs
Senior SRE: Multi-Cloud Automation & Reliability Lead
Senior SRE: Multi-Cloud Automation & Reliability Lead

Sii Poland • Wrocław

On-site
PLN 260,000 - 360,000
Great Place to Work
Profit sharing
Medical care
+3
Cloud SRE: 24/7 Reliability, Automation & Incident Response
Cloud SRE: 24/7 Reliability, Automation & Incident Response

IBM Computing • Kraków

On-site
PLN 180,000 - 240,000
Senior Multi-Cloud SRE: Automate & Scale Reliability
Senior Multi-Cloud SRE: Automate & Scale Reliability

Sii Poland • Piła

On-site
PLN 230,000 - 350,000
Great Place to Work
Profit sharing
Internal trainings
+2