Site Reliability Engineer

XTB online investing

Warszawa

Hybrid

PLN 199,000 - 253,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Training budget for courses
Birthday day off
Parental leave day off
Equipment provided
Private medical care
English language platform access
Wellbeing platform and private therapy
Remote work options (Warsaw office /co

Job summary

XTB online investing is seeking a Site Reliability Engineer to elevate reliability for millions of clients. You will drive observability, implement standardized telemetry, and collaborate with product teams to balance feature velocity with system resilience.

The role requires hands-on engineering across Kubernetes environments, Ansible, and cloud architectures (Azure and on‑prem). You will be part of an experienced SRE team in a dynamic financial tech setting, with a strong emphasis on

Qualifications

  • SRE/DevOps background in large-scale environments.
  • Strong Python automation for tooling and internal solutions.
  • Experience with incident management and post-incident reviews.

Responsibilities

  • Develop and maintain Observability platform and telemetry model.
  • Partner with product/engineering to ensure reliability and SLO alignment.
  • Improve detection with AI/ML-led anomaly detection and proactive resilience.
  • Build internal automation and tooling to streamline SRE workflows.
  • Participate in on-call rotation and incident post-mortems.

Skills

SRE/DevOps
Python automation
Incident management
Post-incident analysis
Soft skills
Azure/cloud experience
Kubernetes
Observability
AI/ML for SRE

Tools

Prometheus
Grafana
ELK Stack
Tempo
Thanos
Jaeger
Ansible
Azure Kubernetes Service
Kubernetes

Job description

XTB is a global company from the financial industry, focusing on online trading of financial instruments. We are the largest FinTech in Poland and a leader in Central and Eastern Europe, and the range of our operations covers several countries, including Asia and South America. At XTB, we focus on the development of our employees, giving them opportunities to gain knowledge and skills in various fields, as well as offering a number of training and development programs. If you are looking for challenges and want to gain valuable experience in an international business environment, XTB is the right place for you.

We are a certified Great Place to Work company.

We are looking for a Site Reliability Engineer to define and drive the reliability of XTB systems at the scale of millions of clients. In this role, you will strengthen SRE practices and shape the resilience of our entire technology stack through high-impact observability, ensuring our systems remain robust and scalable.

Responsibilities
  • Observability Platform Engineering: Develop a standardized observability ecosystem. Implement a conscious telemetry model focusing on structured events, distributed tracing, and intelligent sampling strategies - that provides deep, actionable insights into system behavior.
  • Reliability Enablement: Act as a strategic partner to product engineering teams, providing the platform, standards, and data they need to own service reliability. Use error budgets and alerting as the primary language for balancing feature velocity with stability.
  • Proactive Resilience & Protection: Enhance detection capabilities to identify issues before they impact the customer. Leverage early-warning systems and AI/ML for automated anomaly detection and intelligent data analysis to continuously verify and strengthen system resilience.
  • Operations & Tooling: Build internal automation and tooling that streamlines SRE workflows, automates routine operational tasks, and enhances efficiency across the technology stack.
  • Incident Management & On-Call Rotation: Participate in an on-call rotation to provide incident management, ensuring rapid incident resolution, effective communication, and post-incident analysis to drive continuous improvement.
Requirements
  • Professional Background: Professional experience in SRE, Infrastructure, or DevOps roles managing high-scale, distributed environments.
  • Technical Engineering: Advanced programming skills in Python, with a strong focus on building scalable automation, internal tooling, and robust scripts.
  • Cloud & Orchestration: Hands-on expertise in managing production-grade Kubernetes environments, configuration management tools like Ansible, and designing resilient infrastructure architectures within Azure Kubernetes Service and on-prem environments.
  • Observability Engineering: Proficiency in building standardized telemetry ecosystems. You have mastered self-hosted opensource tools for observability data collection, storage and visualization, like Prometheus, Grafana, ELK Stack, Tempo, Thanos, Jaeger and similar.
  • Operational & Soft Skills: Ability to drive incident management, conduct thorough post-incident analysis, and foster a culture of reliability and shared ownership.
Nice to have
  • Experience with commercial APM platforms (e.g., Datadog, Splunk, New Relic) and chaos engineering tooling.
  • Experience with cloud cost management and FinOps principles.
  • Experience defining and tracking SRE metrics (SLI/SLOs) and managing error budgets to drive reliability.
  • Experience with AI/ML techniques for SRE tasks, such as AIOps, automated anomaly detection, log analysis, and optimizing reliability workflows.
  • Experience in building and managing strategies to proactively manage technical debt and align team output with organizational goals.
What We Offer
  • Real influence on the development of the company and the product.
  • Work in an experienced team that is happy to share its knowledge.
  • A clear vision of development thanks to regular feedback and clear career paths.
  • Regular team-building meetings.
Benefits
  • A training budget for courses and conferences that interest you.
  • An extra day off on your birthday.
  • An extra day off for parents.
  • Equipment tailored to your needs.
  • Private medical care and group insurance.
  • Access to an e-learning platform for learning English and a benefits platform.
  • Access to a wellbeing platform and the opportunity to take advantage of workshops and private therapy sessions.
  • Remote work, from the office in Warsaw or from a coworking space in your city.

17,850 zł - 22,700 zł a month

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Team Leader (Site Reliability Engineering)
Engineering Team Leader (Site Reliability Engineering)

XTB online investing • Warszawa

Hybrid
PLN 304,000 - 386,000
Training budget
Birthday off
Parental leave
+2
Site Reliability Engineer
Site Reliability Engineer

XTB online investing • Poland

Hybrid
PLN 180,000 - 300,000
Remote work
Office in Warsaw
Birthday leave
+4
Senior Java Software Engineer
Senior Java Software Engineer

XTB online investing • Warszawa

Hybrid
PLN 228,000 - 301,000
Training budget for courses
Birthday off
Parental leave
+4
Engineering Team Leader (Site Reliability Engineering)
Engineering Team Leader (Site Reliability Engineering)

XTB online investing • Poland

Hybrid
PLN 320,000 - 520,000
Remote work
Private medical care
Group insurance
+5
Platform Engineering Team Leader
Platform Engineering Team Leader

XTB online investing • Warszawa

Hybrid
PLN 304,000 - 386,000
Private medical care
Group insurance
Remote work flexibility
+1
Data Engineer
Data Engineer

XTB online investing • Warszawa

On-site
PLN 156,000 - 199,000
Training budget
Birthday off
Parental leave
+4
Senior Project Manager
Senior Project Manager

XTB online investing • Warszawa

Hybrid
PLN 199,000 - 253,000
Training budget
Birthday leave
Parental leave
+5
Data Engineer
Data Engineer

XTB online investing • Poland

On-site
PLN 180,000 - 240,000
Private medical care
Group insurance
Birthday leave
+4
Product Manager
Product Manager

XTB online investing • Warszawa

Hybrid
PLN 161,000 - 205,000
Real influence on product development
Knowledge-sharing team
Regular feedback and clear career path
+9
Specialist in Back-Office Team Regulated Markets
Specialist in Back-Office Team Regulated Markets

XTB • Warszawa

Hybrid
PLN 90,000 - 150,000
Hybrid work model
Private medical care
Meal co-financing
+3