Site Reliability Engineer

Balyasny Asset Management L.P.

Warszawa

On-site

PLN 180,000 - 300,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Balyasny Asset Management L.P. is seeking a Site Reliability Engineer to shape and drive our SRE practices across cloud and on‑prem environments.

You will establish standards, promote observability, and partner with DevOps to enhance reliability and performance of trading systems. The role requires hands-on experience with Prometheus, Grafana, Loki, Tempo, Kubernetes, Docker, and scripting in Python, Bash, or Go, plus AWS and on‑prem hosting exposure.

Qualifications

  • 5+ years of experience in SRE or similar roles within distributed systems environments.
  • Bachelor's degree in engineering, computer science, information systems, or equivalent.
  • SME with Prometheus, Grafana, Loki, Tempo (OTEL).
  • Extensive knowledge of Kubernetes and Docker.
  • Hands-on with cloud (AWS preferred) and on-prem hosting.
  • Scripting in Python, Bash, or Go for automation.
  • Strong CI/CD, agile, and DevOps culture.

Responsibilities

  • Develop and promote the SRE philosophy, practices, and scalability of infrastructure.
  • Scale observability and monitoring using Prometheus, Grafana, Loki, Tempo for visibility.
  • Participate in on-call rotation and ensure reliability of deployments.
  • Define standards for reliability in Kubernetes, optimize configurations.
  • Develop automation to improve deployment pipelines, health checks, and recovery.
  • Collaborate with dev teams to improve service stability and SLOs.

Skills

Prometheus
Grafana
Loki
Tempo
Kubernetes
Docker
AWS
Python
Bash
Go
CI/CD
DevOps
SRE

Education

Bachelor's degree in engineering or computer science

Job description

We are looking for a Site Reliability Engineer who can cultivate our SRE philosophy, processes, and technologies from the ground up. This role entails driving standards and fostering adoption across our technology teams, whilst closely partnering with our DevOps and Cloud teams.

With a hands-on approach, you'll work across both cloud and on-premises hosting platforms, ensuring the reliability and scalability of our trading systems and production environments. This is a chance to play a pivotal role in transforming our operational capabilities and enhancing performance across a wide array of environments and platforms.

Key Responsibilities:
  • Develop and promote our SRE philosophy, establishing best practices and processes that will be instrumental in scaling our infrastructure.
  • Implement and scale end-to-end observability and monitoring solutions using Prometheus, Grafana, Loki, and Tempo, ensuring high visibility into application performance and infrastructure health.
  • Participate in on-call rotation with approximately 1 week per month of on-call time shared equally across members of the team.
  • Review and define standards for application reliability requirements within our Kubernetes environment, ensuring application configuration is optimized for performance, cost and reliability.
  • Develop automation and tooling to improve efficiency and reliability of deployment pipelines, system health checks, and recovery procedures.
  • Collaborate with development teams to enhance service stability, scalability, and fault tolerance through SRE best practices like blameless post-mortems and service level objectives (SLOs).
To be considered a good fit, you must have:
  • 5+ years of experience in SRE or similar roles within complex, distributed systems environments.
  • A Bachelor's degree in engineering, computer science, information systems, or equivalent experience.
  • SME with key SRE technologies such as Prometheus, Grafana, Loki, Tempo (OTEL).
  • Extensive knowledge of container orchestration using Kubernetes and containerization with Docker.
  • Hands-on experience with both cloud (AWS preferred) and on-premises hosting platforms.
  • Proven ability to script in languages like Python, Bash, or Go, to automate routine tasks and deployment pipelines.
  • Strong understanding of CI/CD principles, agile methodologies, and DevOps culture.
  • High level of initiative, passion for reliability engineering, detail orientation, and follow-through capabilities.
  • Exceptional interpersonal and communication skills, with the ability to explain complex technical concepts to a diverse audience.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Województwo pomorskie

On-site
PLN 80,000 - 120,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Kraków

On-site
PLN 254,000 - 340,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Site Reliability Engineer
Site Reliability Engineer

Caspian One • Warszawa

On-site
PLN 180,000 - 280,000
Senior Systems Site Reliability Engineer, B2B
Senior Systems Site Reliability Engineer, B2B

Jobtailor • Poland

On-site
PLN 180,000 - 320,000
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 120,000
Health benefits
Financial benefits
Family support
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Wrocław

On-site
Competitive salary
Flexible schedule
Medical insurance
+4
Senior Site Reliability Engineer - Cloud & On-Prem Scale
Senior Site Reliability Engineer - Cloud & On-Prem Scale

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

Capital.com • Warszawa

On-site
PLN 180,000 - 320,000
Competitive salary
Generous time off
Comprehensive health benefits
+4
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

Capital Com Australia limited • Warszawa

Hybrid
PLN 180,000 - 280,000
Competitive Salary
Work-Life Harmony
Generous Time Off
+4
Software Reliability Engineer/Devops
Software Reliability Engineer/Devops

RE Partners • Warszawa

Hybrid
PLN 180,000 - 240,000
Hybrid work model in Warsaw, Poland