Senior Site Reliability Engineer

SolarWinds

Kraków

Hybrid

PLN 200,000 - 360,000

Full time

11 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Medical care with Luxmed
Pension plan
English/Polish classes
Free lunches on Wednesdays

Job summary

SolarWinds is seeking a Senior Site Reliability Engineer to build, operate, and improve production systems across Kubernetes, AWS, and Azure. You will work hands-on in a fast-paced, collaborative environment with emphasis on reliability and automation.

The role requires 5+ years in SRE/DevOps, strong Linux, Terraform, and scripting skills, and the ability to troubleshoot complex distributed systems while collaborating across teams in Kraków and beyond.

Qualifications

  • 5+ years in Site Reliability Engineering, DevOps, or related field.
  • Hands-on production Kubernetes experience, including troubleshooting.
  • Linux systems administration and troubleshooting skills.
  • Strong cloud experience with AWS and Azure.

Responsibilities

  • Operate, maintain, upgrade and improve production Kubernetes clusters and workloads across AWS and Azure environments.
  • Manage Kubernetes platform components and technologies such as Helm, Kustomize, operators, istio, autoscaling, and cluster/node lifecycle management.
  • Support and maintain production database platforms such as ClickHouse, Aurora and other distributed data systems, including performance troubleshooting and operational health.
  • Build and maintain infrastructure using Terraform.
  • Develop automation and tooling to reduce operational toil and improve reliability using Python, Go, Bash, or similar technologies.
  • Participate in an on-call rotation, respond to production incidents, and lead or contribute to incident resolution and root-cause analysis.
  • Improve observability, monitoring, logging, alerting, and incident response across infrastructure and services.
  • Partner closely with software engineering and platform teams to design and deploy reliable, scalable services.
  • Contribute to capacity planning, performance optimization, patching, upgrades, and infrastructure lifecycle management.
  • Develop and maintain operational documentation, runbooks, and troubleshooting guides.
  • Participate in reliability initiatives, including automation, resilience engineering, disaster recovery, and continuous improvement.

Skills

Kubernetes
AWS
Azure
Linux
Terraform
Python
Go
Bash
Incident response
Automation

Tools

Helm
Kustomize
Istio
ClickHouse
Aurora

Job description

At SolarWinds, we're a people-first company. Our purpose is to enrich the lives of the people we serve—including our employees, customers, shareholders, partners, and communities. Join us in our mission to help customers accelerate business transformation with simple, powerful, and secure solutions.

The ideal candidate thrives in an innovative, fast-paced environment and is collaborative, accountable, ready, and empathetic. We're looking for individuals who believe they can accomplish more as a team and create lasting growth for themselves and others. We hire based on attitude, competency, and commitment. Solarians are ready to advance our world-class solutions in a fast-paced environment and accept the challenge to lead with purpose. If you're looking to build your career with an exceptional team, you've come to the right place. Join SolarWinds and grow with us!

We work inhybrid mode 3+2, at least 3 days at the office (with mandatory Wednesdays and Thursdays) and 2 days at the home office.

The location of ouroffice is Puszkarska 7J/Building E, 30-644 Kraków, Polska.

We employ only viaan employmentcontract - FTE.

What you will do

SolarWinds is looking for a Senior Site Reliability Engineer with 5+ years of experience to help build, operate, and improve the systems that support our products and internal platforms. This role is ideal for someone who is hands-on, dependable, and comfortable working across Kubernetes, AWS, Azure, Linux, and Database environments in a production setting.

The right candidate has a strong foundation in site reliability engineering, enjoys solving problems in live environments, and understands the importance of reliability, automation, and operational discipline.

  • Operate, maintain, upgrade and improve production Kubernetes clusters and workloads across AWS and Azure environments.
  • Manage Kubernetes platform components and technologies such as Helm, Kustomize, operators, istio, autoscaling, and cluster/node lifecycle management.
  • Support and maintain production database platforms such as ClickHouse, Aurora and other distributed data systems, including performance troubleshooting and operational health.
  • Build and maintain infrastructure using Terraform.
  • Develop automation and tooling to reduce operational toil and improve reliability using Python, Go, Bash, or similar technologies.
  • Participate in an on-call rotation, respond to production incidents, and lead or contribute to incident resolution and root-cause analysis.
  • Improve observability, monitoring, logging, alerting, and incident response across infrastructure and services.
  • Partner closely with software engineering and platform teams to design and deploy reliable, scalable services.
  • Contribute to capacity planning, performance optimization, patching, upgrades, and infrastructure lifecycle management.
  • Develop and maintain operational documentation, runbooks, and troubleshooting guides.
  • Participate in reliability initiatives, including automation, resilience engineering, disaster recovery, and continuous improvement.
Required Qualifications
  • 5+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Platform Engineering, or a related field.
  • Strong hands-on production Kubernetes experience, including operating, upgrading and troubleshooting Kubernetes clusters and workloads.
  • Strong understanding of Kubernetes fundamentals, including:
    • Pods, Deployments, StatefulSets, DaemonSets, Jobs, CronJobs
    • Services, Ingress, DNS, and Kubernetes networking
    • ConfigMaps, Secrets, and persistent storage
    • Resource requests/limits and scheduling
    • Autoscaling across workloads, clusters, and nodes
    • Pod Disruption Budgets and high availability patterns
    • Kubernetes upgrades and cluster lifecycle management
  • Experience troubleshooting Kubernetes at both the application workload and cluster infrastructure levels.
  • Strong hands-on experience with AWS and Azure cloud infrastructure.
  • Strong Linux systems administration and troubleshooting skills.
  • Experience operating customer facing, highly available production systems.
  • Experience participating in a production on-call rotation and responding to high-severity incidents.
  • Strong experience with Terraform and Infrastructure as Code.
  • Experience with scripting and automation using Python, Go, Bash, or similar languages.
  • Strong understanding of infrastructure concepts including compute, networking, storage, DNS, load balancing, and security.
  • Strong troubleshooting skills with the ability to methodically diagnose complex distributed-system failures.
  • Strong communication skills and the ability to collaborate effectively with engineering and cross-functional teams.
What we are looking for
  • Ownership mindset and strong operational discipline
  • Ability to stay calm and effective during incidents
  • Willingness to learn, improve systems, and drive reliability-focused changes
  • Practical approach to solving infrastructure and operational problems
  • Team player who values clear communication and documentation
  • Continuous learner stays current with Kubernetes, cloud infrastructure, and modern SRE practices.
On‑Call Expectations

This position requires participation in a scheduled on‑call rotation to support production systems and help ensure service reliability and availability.

  • 10 study days per year
  • 2 volunteering days per year
  • 30‑day holidays after 5-year tenure, Sabbatical Leave
  • 4 weeks of paternity leave
  • Up to 8700 PLN personal education budget per year
  • 300 PLN corrective glasses reimbursement every two years
  • Medical care with Luxmed - individual, partner, or family package fully paid by the company
  • The company fully pays for group life insurance
  • Pension scheme (employee capital plans) with 1.5% employer contribution
  • Unlimited access to LinkedIn Learning
  • English/Polish classes
  • MyBenefit platform with a monthly subsidy of 103 PLN (with various vouchers and Multisport cards available)
  • 500 PLN per year of race fee reimbursement
  • Employee Assistance Program
  • Free lunches at the office on Wednesdays

SolarWinds is an Equal Employment Opportunity Employer. SolarWinds will consider all qualified applicants for employment without regard to race, color, religion, sex, age, national origin, sexual orientation, gender identity, marital status, disability, veteran status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE Lead – Kubernetes (10/961)
Senior SRE Lead – Kubernetes (10/961)

INFOLET SP. Z O.O. • Kraków

On-site
PLN 350,000 - 500,000
Relocation package
Extended medical care
Multisport Benefit card
+1
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Kraków

On-site
PLN 180,000 - 300,000
Great Place to Work
Centre of internal trainings
Profit sharing
+2
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Poznań

On-site
PLN 260,000 - 340,000
Profit sharing
Medical care
Centre of internal trainings
+1
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Warszawa

On-site
PLN 210,000 - 270,000
Great Place to Work
Centre of internal trainings
Profit sharing
+2
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Łódź

On-site
PLN 240,000 - 360,000
Great Place to Work
Centre of internal trainings
Profit sharing
+2
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Toruń

On-site
PLN 240,000 - 360,000
Great Place to Work
Solid financial situation
Contracts with the biggest brands
+9
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Lublin

On-site
PLN 240,000 - 380,000
Medical care
Profit sharing
Internal trainings
+1
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Wrocław

On-site
PLN 260,000 - 360,000
Great Place to Work
Profit sharing
Medical care
+3
Senior Site Reliability Engineer (f/m/x)
Senior Site Reliability Engineer (f/m/x)

Sii Poland • Piła

On-site
PLN 230,000 - 350,000
Great Place to Work
Profit sharing
Internal trainings
+2
Senior Software Engineer (Senior UNIX Automation Engineer)
Senior Software Engineer (Senior UNIX Automation Engineer)

Sopra Steria • Katowice

Hybrid
PLN 134,000 - 179,000
Luxmed
Medicover Sport
Worksmile
+1