Senior Site Reliability Engineer

Grid Dynamics

Kraków

On-site

PLN 254,452 - 339,270

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Sports benefits
Professional development opportunities
Flexible schedule
Corporate social events

Job summary

Grid Dynamics in Kraków is seeking a Senior Site Reliability Engineer to enhance the reliability, resilience, and performance of its core enterprise products. You will work at the intersection of infrastructure and software engineering, proactively analyzing system architecture and optimizing application reliability.

The ideal candidate has over 5 years of experience and strong skills in Java, Docker, and Kubernetes. You will be responsible for incident response, automation, and mentoring while ensuring system health and security compliance.

Qualifications

  • 5+ years of experience in Site Reliability Engineering or Platform Engineering roles.
  • Strong proficiency in Java, Spring Boot, Hibernate, and Jenkins.
  • Expertise with Docker and Kubernetes.
  • Deep knowledge of Linux systems, networking, and distributed architectures.

Responsibilities

  • Understand product topology and identify bottlenecks.
  • Respond to outages and perform blameless Root Cause Analysis.
  • Fix defects directly in production and recommend improvements.
  • Manage application vulnerabilities and monitor compliance.

Skills

Java
Spring Boot
Docker
Kubernetes
Linux systems
Problem-solving
Communication skills

Education

Bachelor’s degree in Computer Science or Systems Engineering

Tools

Jenkins
Prometheus
Grafana
Splunk

Job description

We are looking for an experienced Senior Site Reliability Engineer to join our team and oversee the reliability, resilience, and performance of our core enterprise products.

In this role, you will bridge the gap between infrastructure operations and software engineering. You won't just react to alerts—you will proactively analyze system architecture, build automation, and dive deep into the application code (Java/Spring Boot) to fix bugs and eliminate issues at their root.

Responsibilities
  • Architecture & Reliability: Understand the end-to-end product topology from both infrastructure and application perspectives. Identify bottlenecks, scale limitations, and unstable components, driving long-term resolutions before they impact production.
  • Incident Response & RCA: Respond to outages, provide L3 on-call technical support (on rotation), and perform blameless Root Cause Analysis (RCA) to implement permanent fixes.
  • Hands-on Engineering: Address defects, perform code bug fixes directly in production, and recommend architectural improvements during incident analysis.
  • Security & Vulnerability Management: Oversee vulnerability management for applications and containers, manage patching processes, ensure compliance, and monitor certificate expirations and renewals according to global best practices.
  • SRE Advocacy & SDLC: Represent the SRE organization in design reviews, capacity planning, and operational readiness exercises. Partner closely with development teams to embed reliability best practices early in the SDLC.
  • Automation & Mentoring: Build automation tools to reduce manual toil and improve efficiency. Spread SRE culture, create standard documentation, and provide technical mentorship to junior team members.
  • System Health: Oversee the production environment by tracking availability, applying learnings from observability tools, and becoming a Subject Matter Expert (SME) on core issuing products.
Requirements
  • Experience: 5+ years of experience in Site Reliability Engineering (SRE) or Platform Engineering roles.
  • Software Engineering: Strong proficiency in Java, Spring Boot, Hibernate, and Jenkins. Ability to read, analyze, and fix application code.
  • Containerization: Hands-on expertise with Docker and container orchestration using Kubernetes.
  • Infrastructure: Deep knowledge of Linux systems, networking, and distributed architectures.
  • Observability: Strong understanding of monitoring, logging, and observability tools (e.g., Prometheus, Grafana, Splunk).
  • Education: Bachelor’s degree in Computer Science, Systems Engineering, or equivalent practical experience.
  • Soft Skills: Excellent problem-solving abilities and strong communication skills.
Nice to have
  • Infrastructure as Code & Cloud: Hands-on experience with tools like Terraform or Ansible, alongside familiarity with major public cloud providers (AWS, GCP, or Azure).
  • Advanced Networking & Service Mesh: Knowledge of service mesh technologies (e.g., Istio, Linkerd) for traffic management, security, and observability in microservices architectures.
  • Industry Experience: Previous background in the FinTech, payments, or banking sectors, with an understanding of high-security compliance standards (e.g., PCI-DSS).
We offer
  • Opportunity to work on bleeding-edge projects
  • Work with a highly motivated and dedicated team
  • Competitive salary
  • Flexible schedule
  • Benefits package - medical insurance, sports
  • Corporate social events
    We offer
    • Professional development opportunities
    • Well-equipped office
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Województwo pomorskie

On-site
PLN 80,000 - 120,000
Medical insurance
Sports benefits
Professional development opportunities
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Grid Dynamics • Wrocław

On-site
Competitive salary
Flexible schedule
Medical insurance
+4
Site Reliability Engineer
Site Reliability Engineer

Balyasny Asset Management L.P. • Warszawa

On-site
PLN 180,000 - 300,000
Senior Systems Site Reliability Engineer, B2B
Senior Systems Site Reliability Engineer, B2B

Jobtailor • Poland

On-site
PLN 180,000 - 320,000
Site Reliability Engineer
Site Reliability Engineer

Fáilte Ireland • Poland

On-site
PLN 180,000 - 280,000
Equity program
Cloud SRE — Reliability, Observability & Equity in InsurTech
Cloud SRE — Reliability, Observability & Equity in InsurTech

Fáilte Ireland • Poland

On-site
PLN 180,000 - 280,000
Site Reliability Engineer
Site Reliability Engineer

Caspian One • Warszawa

On-site
PLN 180,000 - 280,000
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

Akamai Technologies • Kraków

On-site
PLN 90,000 - 120,000
Health benefits
Financial benefits
Family support
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Experis ManpowerGroup Sp. z o.o. • Poland

On-site
PLN 90,000 - 120,000
MultiSport Plus
Medicover
Generali life insurance
+2
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

Capital.com • Warszawa

On-site
PLN 180,000 - 320,000
Competitive salary
Generous time off
Comprehensive health benefits
+4