Kubernetes Infrastructure Reliability Engineer

Roche

Madrid

Presencial

EUR 90.000 - 130.000

Jornada completa

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca para este puesto: genera un currículum y una carta de presentación adaptados en cuestión de un minuto.

Supera los filtros ATS

Descripción de la vacante

Roche seeks a Kubernetes Infrastructure Reliability Engineer to build and maintain a highly available, cloud-native platform across hybrid cloud deployments. You will apply software engineering to operations, emphasizing automation, security, and observability, while enabling Roche applications and processes to scale globally.

The role requires leading incident prevention, PoCs, and stakeholder engagement, with a focus on reliability, capacity planning, and cross-region coordination in a

Formación

  • 4–7 years relevant experience without a degree.
  • Bachelor’s or Master’s degree increases eligibility.
  • Experience in multinational environment is a plus.

Responsabilidades

  • Focus on capacity planning and launch reviews for services before they go live.
  • Resolve complex problems in a global Kubernetes-based infrastructure and lead end-to-end design.
  • Bridge between engineering and operations; communicate complex solutions to cross-functional teams.

Conocimientos

Kubernetes
CI/CD
Python
Bash
Go
IaC (Ansible/Terraform)
AWS
Networking
Observability

Educación

Bachelor's degree or equivalent

Herramientas

Rancher
Portworx
GitLab
Terraform
Ansible
Jenkins

Descripción del empleo

At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.

The Position

The Kubernetes Infrastructure Reliability Engineer is a highly skilled expert responsible for solving complex business problems using advanced cloud native technologies. The engineer will build and maintain a Kubernetes-based infrastructure, enabling the modernization of business applications and processes.

This role combines software and systems engineering to optimize systems, increase efficiency, and eliminate operational work through automation.

You will be part of the global CaaS infrastructure team at a leading healthcare company, working with members across different regions. The team's mandate is to deliver, maintain, and continuously improve a highly available Kubernetes platform across hybrid cloud deployments, including on-premise data centers and public clouds like AWS. In this role, you will apply software engineering principles to operations to build and run massively distributed, fault-tolerant systems, focusing heavily on automation, security, and observability.

Job Responsibilities
  • Service Reliability and Optimization: Focus on capacity planning and launch reviews for services before they go live. Perform blameless postmortems and proactive identification of potential outages to foster iterative improvements
  • Accountability/Problem Solving: Resolves complex problems in a global Kubernetes-based infrastructure through in-depth evaluation of variable factors, including inter-organizational impact, balanced with effective consultative engagement of key stakeholders. Leads end-to-end design of infrastructure solutions and maintains component standards. Evaluates promising solutions via Proof of Concept (PoCs) and feasibility studies across multiple areas, and serves as an internal escalation point for major incidents
  • Stakeholder Management: Acts as a bridge between engineering and operations. Communicates and presents complex information and potential solutions to cross-functional teams and the business in non-technical terms. Represents the organization as a prime contact on initiatives and interacts with senior internal and external personnel. Uses deep knowledge to influence IT infrastructure vendor product evaluations and collaborates with multiple IT partners (e.g. Enterprise Architects, Solution Owners) to integrate feedback. Mentors and shares DevOps culture, guiding developers on how to create and deploy cloud-native applications
  • Impact/Strategy: Provides technical leadership and direction for small-to-medium sized initiatives (projects, lifecycle work, PoCs). Ensures solutions comply with Quality/Regulatory standards and that designs adhere to the organization’s Technical Architecture Framework (TAF) policies and directions. Assists in planning technology projects, estimating engineering resources, dependencies, risks and timelines for successful delivery
  • Business / Technical ability: Applies extensive cloud native technical expertise, acting as a recognized expert in Kubernetes and maintaining in-depth knowledge across related cloud native technologies (containers, AWS, etc.). Demonstrates a detailed understanding of how IT infrastructure impacts respective Roche business processes and outcomes
Qualifications
Education & Professional Experience
  • Without Degree: 4–7 years of relevant experience
  • Bachelor’s Degree: 2–5 years of relevant experience
  • Master’s Degree: 1–3 years of relevant experience
  • At least 1 year of experience working in a multinational environment; healthcare industry experience is a plus
Technical Skills
  • Kubernetes & Containers: Strong hands‑on experience navigating, managing, and hardening Kubernetes clusters and containers, including knowledge of distributed storage. A Certified Kubernetes Administrator (CKA) certification is a strong plus. Knowledge of tools like Rancher or Portworx is beneficial
  • Infrastructure as Code (IaC): Hands-on experience delivering and managing infrastructure automation using tools like Ansible and Terraform
  • Scripting & Software Engineering: Proficiency in scripting and programming languages, primarily Python, Bash, or Go, including experience with test automation (e.g., pytest) and APIs deployment and management
  • CI/CD Tools: Expert knowledge of implementing software delivery pipelines using tools (e.g., Jenkins, Rundeck, or GitLab)
  • Systems & Networking: Strong understanding of Linux operating systems and core networking principles, including DNS, load balancing, firewalls, routing, and service meshes
  • Observability: Experience configuring logging, metrics, and monitoring tools, specifically focusing on setting up alerts based on symptoms rather than waiting for system outages
  • Cloud Infrastructure: Experience with public cloud platforms, with a strong preference for AWS, specifically involving managed services for compute, networking, security, and identity (e.g., EKS, VPC, IAM)
General and Operational Knowledge
  • Proven experience applying best practices in an always‑up, always‑available service environment utilizing Scrum and Agile methodologies
  • Deep understanding of Technical Architecture Frameworks (TAF) and Quality/Regulatory compliance standards
Additional Qualifications
  • Excellent problem‑solving skills, decision‑making ability, and sound judgment
  • A strong team‑oriented mindset with the ability to function independently with low supervision and navigate ambiguity
  • Highly fluent oral and written English communication skills are required.
  • Ability to work across multiple time zones
  • Strong customer & delivery focus
  • 24/7 on‑call rotation is required for this role****

#RDT2026

Who we are

A healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life‑changing healthcare solutions that make a global impact.

Let’s build a healthier future, together.

Roche is an Equal Opportunity Employer.
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Cloud DevOps Infrastructure Engineer - Roche Cloud Platform Azure
Cloud DevOps Infrastructure Engineer - Roche Cloud Platform Azure

Roche • Madrid

Presencial
EUR 60.000 - 90.000
Cloud Devops Infrastructure Engineer - Roche Cloud Platform Azure
Cloud Devops Infrastructure Engineer - Roche Cloud Platform Azure

Roche & Company • Madrid

Presencial
EUR 90.000 - 120.000
IT Infrastructure Engineer
IT Infrastructure Engineer

Roche • Madrid

Híbrido
EUR 55.000 - 75.000
Equal Opportunity Employer
IT Software Engineer - RDT Engineering Excellence & Experience
IT Software Engineer - RDT Engineering Excellence & Experience

Roche • Madrid

Presencial
EUR 70.000 - 100.000
Cloud Devops & Validation Engineer - Roche Cloud Platform
Cloud Devops & Validation Engineer - Roche Cloud Platform

Roche & Company • Madrid

Presencial
EUR 65.000 - 90.000
OSS Lead - RDT Quality, Risk & Compliance
OSS Lead - RDT Quality, Risk & Compliance

Roche • Madrid

Presencial
EUR 85.000 - 120.000
DevOps Infrastructure Engineer
DevOps Infrastructure Engineer

Roche • Madrid

Presencial
EUR 65.000 - 95.000
Infrastructure Management and Provisioning Engineer
Infrastructure Management and Provisioning Engineer

Roche • Madrid

Presencial
EUR 90.000 - 125.000
CaaS Release Train Engineer
CaaS Release Train Engineer

F. Hoffmann-La Roche AG • Madrid

Presencial
EUR 70.000 - 100.000
Technical Lead - Observability - RDT Digital Operations and Reliability
Technical Lead - Observability - RDT Digital Operations and Reliability

Roche • Madrid

Presencial
EUR 100.000 - 130.000