Site Reliability Engineer (SRE)

XM

Poland

On-site

PLN 180,000 - 240,000

Full time

44 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive remuneration package
International training opportunities
Confidential recruitment

Job summary

XM is seeking a senior SRE/Resiliency Engineer to drive reliability initiatives in a cloud-native environment. You will lead chaos experiments, design observability, and oversee migrations with strong Kubernetes and IaC skills.

Responsibilities include mentoring teammates, on-call rotations, and ensuring compliance with resiliency standards across cloud deployments. The role requires extensive AWS experience and a deep understanding of SRE practices.

Qualifications

  • BSc/MSc degree in Computer Science or related field.
  • 5+ years of cloud services experience, with at least 3 years on AWS cloud.
  • 3+ years of experience in SRE or a similar role.
  • Experience with monitoring, APM, logging, and notification tools.
  • Familiarity with incident, problem and change management procedures and practices.
  • Advanced knowledge of SRE practices and methods.
  • Understanding and practice of Service Levels.
  • Strong troubleshooting skills and the ability to mentor others.
  • Extensive experience with Kubernetes and related technologies, services, and ecosystem.
  • Advanced knowledge of CI/CD, Infrastructure as Code (IaC) concepts and tools, especially HCL Terraform and AWS CloudFormation.
  • Experience with versioning tools like Git.
  • Strong organizational and documentation skills.
  • Exceptional time management and research abilities.
  • Advanced Linux, networking, and scripting skills.

Responsibilities

  • Honor and practice the Resiliency pillar of the Well Architected Framework in all tasks and responsibilities.
  • Conduct Chaos Engineering experiments and relevant exercises to improve resiliency and fault-tolerance.
  • Research workloads for migrating to the cloud with minimal disruption and impact.
  • Monitor cloud migration projects to ensure seamless transitions.
  • Design, consult, re-platform, and re-factor the observability of current cloud infrastructure.
  • Coordinate with other IT departments and teams regarding observability for both individual and organizational needs.
  • Regularly assess cloud deployments for compliance with the company’s standards and best practices.
  • Investigate and correct areas where observability is lagging.
  • Stay up to date and provide training on new and current technologies, services, tools, methodologies, and practices.
  • Occasionally participate in service capacity planning, software performance analysis, and system tuning.
  • Mentor colleagues in technical skills and knowledge.
  • Analyze, oversee, and remediate the company’s resiliency.
  • Participate in on-call support 24/7 based on a rotation schedule

Skills

Cloud experience
AWS
SRE
Kubernetes
Terraform
CloudFormation
CI/CD
IaC
APM/Monitoring
Linux
Networking
Troubleshooting
Mentoring

Education

BSc/MSc in Computer Science or related field

Tools

Kafka (MSK)
Postgres
MySQL
Python
Go
Git

Job description

You will join a team working with Observability, Escalations, Post-mortems, Correction of Errors, and other practices that will contribute to the company's goal of cloud resiliency. You will be responsible for driving processes around reliability, best practices, cultural change, and enforcement of these practices.

The main responsibilities of the position include:

  • Honor and practice the Resiliency pillar of the Well Architected Framework in all tasks and responsibilities
  • Conduct Chaos Engineering experiments and relevant exercises to improve resiliency and fault-tolerance
  • Research workloads for migrating to the cloud with minimal disruption and impact
  • Monitor cloud migration projects to ensure seamless transitions
  • Design, consult, re-platform, and re-factor the observability of current cloud infrastructure
  • Coordinate with other IT departments and teams regarding observability for both individual and organizational needs
  • Regularly assess cloud deployments for compliance with the company’s standards and best practices
  • Investigate and correct areas where observability is lagging
  • Stay up to date and provide training on new and current technologies, services, tools, methodologies, and practices
  • Occasionally participate in service capacity planning, software performance analysis, and system tuning
  • Mentor colleagues in technical skills and knowledge
  • Analyze, oversee, and remediate the company’s resiliency
  • Participate in on-call support 24/7 based on a rotation schedule

Main requirements:

  • BSc/MSc degree in Computer Science or related field
  • 5+ years of cloud services experience, with at least 3 years on AWS cloud
  • 3+ years of experience in SRE or a similar role
  • Experience with monitoring, APM, logging, and notification tools
  • Familiarity with incident, problem and change management procedures and practices
  • Advanced knowledge of SRE practices and methods
  • Understanding and practice of Service Levels
  • Strong troubleshooting skills and the ability to mentor others
  • Extensive experience with Kubernetes and related technologies, services, and ecosystem
  • Advanced knowledge of CI/CD, Infrastructure as Code (IaC) concepts and tools, especially HCL Terraform and AWS CloudFormation
  • Experience with versioning tools like Git
  • Strong organizational and documentation skills
  • Exceptional time management and research abilities
  • Advanced Linux, networking, and scripting skills

The following will be considered an advantage:

  • Experience with platforms like Kafka (MSK)
  • Experience with RDBMSs, particularly Postgres and MySQL
  • Knowledge of scripting languages such as Python or Go

Benefit from:

  • Attractive remuneration package and perks
  • Intellectually stimulating work environment
  • Continuous personal development and international training opportunities

The Hiring Experience: What Awaits You

  • Show Your Skills – Online Technical Challenge
  • Let’s Connect – Intro Chat with Talent Acquisition
  • Deep Dive – First Interview with Your Future Team
  • Final Connection – Final Interview

All applications will be treated with strict confidentiality!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Principal SRE Engineer
Principal SRE Engineer

Talanto • Kraków

Hybrid
PLN 35,000 - 40,000
Software Engineering & SRE Lead
Software Engineering & SRE Lead

BELVEDERE • Warszawa

Hybrid
PLN 335,000 - 536,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Akamai Technologies • Kraków

On-site
PLN 180,000 - 270,000
Senior Site Reliability Engineer (Cloud and Networking) - Remote
Senior Site Reliability Engineer (Cloud and Networking) - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 130,000
Health benefits
Financial support
Family support
+1
Senior Site Reliability Engineer (SRE) – Kubernetes
Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind • Kraków

Remote
PLN 180,000 - 240,000
Private healthcare and insurance
Multisport card
Language classes
+3
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 120,000
Health benefits
Financial benefits
Family support
+2
Senior Software Quality & SRE Lead
Senior Software Quality & SRE Lead

Randstad Polska Sp. z o.o. • Warszawa

On-site
PLN 260,000 - 380,000
DevOps Engineer - SRE Observability - name
DevOps Engineer - SRE Observability - name

OpenTalent • Poland

Hybrid
PLN 180,000 - 300,000
Office as an option
Remote work option
Workation
+4
Senior Engineer - SRE & Infrastructure Services
Senior Engineer - SRE & Infrastructure Services

EPAM Systems • Poland

Hybrid
PLN 240,000 - 340,000
Hybrid work model
Remote within Poland
Relocation opportunities
+1