Staff Site Reliability Engineer

Cytiva

Kraków

Hybrid

PLN 260,000 - 380,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Remote work arrangement

Job summary

Cytiva, Kraków, seeks a Site Reliability Engineer to drive reliability for cloud-native services, set SLOs, and lead incident responses. You will design observability, improve performance, and mentor teams while collaborating across Cytiva’s operations.

The role supports a remote work arrangement and requires 5+ years in SRE/DevOps with strong automation skills in Python or Go.

Qualifications

  • 5+ years in SRE/DevOps on production systems.
  • Deep knowledge of SRE concepts: SLOs, error budgets, toil reduction.
  • Experience with observability platforms and dashboards.
  • Cloud platforms: AWS, Azure, or GCP; containers and IaC.
  • Proficient in Python or Go for automation.

Responsibilities

  • Drive reliability direction across teams.
  • Define and monitor SLOs and error budgets.
  • Design and maintain observability for cloud-native apps.
  • Lead incident response and postmortems.
  • Mentor engineers and set engineering standards.

Skills

SRE principles
Observability design
Cloud platforms
Kubernetes
IaC
Python/Go
Troubleshooting

Education

CS/Engineering degree

Tools

Prometheus
Grafana
Datadog
ELK stack
OpenTelemetry
Splunk
Terraform
OpenTofu
Pulumi

Job description

Bring more to life. At Danaher, our work saves lives. And each of us plays a part. Fueled by our culture of continuous improvement, we turn ideas into impact - innovating at the speed of life.

Bring more to life. At Danaher, our work saves lives. And each of us plays a part. Fueled by our culture of continuous improvement, we turn ideas into impact - innovating at the speed of life. Our 60,000+ associates work across the globe at more than 15 unique businesses within life sciences, diagnostics, and biotechnology. Are you ready to accelerate your potential and make a real difference? At Danaher, you can build an incredible career at a leading science and technology company, where we're committed to hiring and developing from within. You'll thrive in a culture of belonging where you and your unique viewpoint matter. Learn about the Danaher Business System which makes everything possible. The Site Reliability Architect is responsible for the availability, performance, and operational maturity of our cloud-native platform and the critical data and AI/ML services that run on it. This is a senior individual contributor role: you will set reliability direction across teams, raise the engineering bar through standards and mentorship, and do the hands-on work of making complex distributed systems predictable. This position reports to the Senior Director, Data and AI Platform and is part of the Chief Information Officer (CIO) Office and will be located onsite in Krakow, Poland. This is a Danaher Corporate role, hosted by our Cytiva operating company in Kraków.

In This Role, You Will Have The Opportunity To
  • Champion SRE practice at scale. Establish and monitor Service Level Objectives and error budgets for critical services, and use them to drive real decisions about reliability, availability, performance, and cost-efficiency - including when to slow down and when to ship.
  • Own observability end to end. Design, implement, and maintain monitoring, logging, and distributed tracing for cloud-native applications and infrastructure (Kubernetes, microservices), and build the dashboards, alerts, and runbooks that give teams deep insight into system health rather than alert noise.
  • Eliminate toil. Identify repetitive operational work and manual process across the production estate and automate it away, developing the tools, scripts, and pipeline improvements that make operations, deployment, and incident response faster and safer.
  • Lead incident response and learning. Participate across the full incident lifecycle - detection, triage, mitigation, resolution - and lead thorough blameless postmortems that get to root cause and produce preventative measures that stick.
  • Shape systems before they're built. Partner with development teams to influence the design of new services so operability, reliability, and cost-efficiency are engineered in from the start, and proactively surface performance bottlenecks and architectural weaknesses before they reach production.
  • Set technical direction and grow the team. Drive cross-team architecture and reliability decisions, establish standards and documentation, mentor engineers across our operating companies, and help foster a culture of technical rigor, blameless learning, and collaboration.
The Essential Requirements Of The Job Include
  • 5+ years of hands-on experience in a Site Reliability Engineering, DevOps, or equivalent role focused on production system reliability and operations; CS/Engineering degree or equivalent practical experience.
  • Strong understanding and practical application of SRE principles - SLOs, error budgets, toil reduction, and blameless culture - with proven experience participating in and improving incident management processes for business-critical systems.
  • Expertise designing, implementing, and managing observability platforms for cloud-native environments (e.g., Prometheus, Grafana, Datadog, ELK stack, OpenTelemetry, Splunk), including the dashboards, alerting, and runbooks that make them actionable.
  • Extensive hands-on experience with at least one major cloud platform (AWS, Azure, or GCP) across compute, networking, and database services; containerization and orchestration (Docker, Kubernetes); Infrastructure as Code (e.g., Terraform, OpenTofu, Pulumi); and proficiency in at least one language (Python, Go) for automation and tool development.
  • Proven ability to set technical direction at platform scale - driving cross-team architecture and design decisions, establishing standards, and mentoring engineers - grounded in strong troubleshooting skills across complex distributed systems, including microservices, CI/CD pipelines, and large-scale data infrastructure.
Travel, Motor Vehicle Record & Physical/Environment Requirements
  • Ability to travel - up to 10%
It would be a plus if you also possess previous experience in:
  • Life sciences, diagnostics, or biotechnology (e.g., partnering with R&D, quality, clinical, manufacturing, or commercial teams)
  • Reliability and efficiency practices at scale - chaos engineering and resilience testing, capacity planning, or cloud cost optimization (FinOps).
  • Working in a matrixed environment

Danaher offers a broad array of comprehensive, competitive benefit programs that add value to our lives. Whether it’s a health care program or paid time off, our programs contribute to life beyond the job. Check out our benefits at Danaher Benefits Info. At Danaher, we believe in designing a better, more sustainable workforce. We recognize the benefits of flexible, remote working arrangements for eligible roles and are committed to providing enriching careers, no matter the work arrangement. This position is eligible for a remote work arrangement in which you can work remotely from your home. Additional information about this remote work arrangement will be provided by your interview team. Explore the flexibility and challenge that working for Danaher can provide. Join our winning team today. Together, we'll accelerate the real-life impact of tomorrow's science and technology. We partner with customers across the globe to help them solve their most complex challenges, architecting solutions that bring the power of science to life. For more information, visit www.danaher.com.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Architect
Site Reliability Architect

Danaher Corporation • Kraków

Remote
PLN 260,000 - 420,000
Remote work eligibility
Competitive benefits
Senior DevOps Engineer
Senior DevOps Engineer

Danaher • Kraków

On-site
PLN 180,000 - 260,000
Director, Platform Engineering
Director, Platform Engineering

Danaher • Kraków

On-site
PLN 400,000 - 560,000
Lead AI SRE and QA Engineer
Lead AI SRE and QA Engineer

Danaher Corporation • Kraków

On-site
PLN 200,000 - 350,000
Remote SRE Architect: Cloud-Native Reliability Leader
Remote SRE Architect: Cloud-Native Reliability Leader

Danaher Corporation • Kraków

Remote
PLN 260,000 - 420,000
Remote work eligibility
Competitive benefits
Lead Data Scientist
Lead Data Scientist

Danaher Corporation • Kraków

On-site
PLN 320,000 - 460,000
Director AI Architecture, Standardization & Engineering Excellence
Director AI Architecture, Standardization & Engineering Excellence

Danaher Corporation • Kraków

On-site
PLN 260,000 - 380,000
Lead Data Governance Architect - Kraków (onsite) OR Poland (remote)
Lead Data Governance Architect - Kraków (onsite) OR Poland (remote)

Danaher Corporation • Warszawa

Hybrid
PLN 240,000 - 360,000
Principal AI Engineering Excellence
Principal AI Engineering Excellence

Danaher Corporation • Kraków

On-site
PLN 300,000 - 420,000
Director, Enterprise Vulnerability & Exposure Management (Kraków on-site OR Poland - remote)
Director, Enterprise Vulnerability & Exposure Management (Kraków on-site OR Poland - remote)

Danaher • Kraków

Hybrid
PLN 360,000 - 540,000