Senior Site Reliability Engineer

F. Hoffmann-La Roche AG

España

On-site

PHP 4,822,182 - 6,027,728

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

F. Hoffmann-La Roche AG is seeking a Senior Site Reliability Engineer to join our global SRE team. In this role, you'll design and maintain cutting-edge tools, focusing on automating tasks, enhancing system reliability, and seamless collaboration with development teams.

We value diversity and are committed to creating an inclusive environment. Join us to influence healthcare innovation on a global scale.

Come help us shape the backbone of technology that drives healthcare solutions!

Qualifications

  • Minimum bachelor’s degree in computer science, Engineering, or related field.
  • Experience in site reliability engineering or software engineering with production on-call experience.
  • Proficiency in scripting languages for automation purposes.

Responsibilities

  • Design and maintain tools and frameworks that automate tasks and streamline deployments.
  • Lead incident management and response, conducting root cause analyses.
  • Collaborate with engineering teams to improve system resilience and reduce operational toil.

Skills

AWS
Azure
Python
Kubernetes
Incident management
Observability tools

Education

Bachelor’s degree in Computer Science or Engineering

Tools

Terraform
CI/CD tools

Job description

Senior Site Reliability EngineerSkip to main content#Senior Site Reliability Engineer page is loaded## Senior Site Reliability EngineerApplylocations: Sant Cugat del Vallèstime type: Full timeposted on: Posted Todayjob requisition id: 202606-114677At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.### ### The Position**The Position** We are building a global Site Reliability Engineering (SRE) team to support critical commercial and internal platforms and applications. As an SRE, you will help design, build, and scale reliable distributed systems that power healthcare innovation worldwide. This role is focused on reliability, scalability, automation and operational excellence. You will influence system design, define reliability standards and reduce operational toil through engineering solutions. This role includes participation in a structured on-call rotation.**Who We Are** At Roche, we are passionate about transforming patients’ lives, and we are bold in both decision and action - we believe that good business means a better world. That is why we come to work every single day. We commit ourselves to scientific rigor, unassailable ethics and access to medical innovations for all. We do this today to build a better tomorrow. Roche is strongly committed to a diverse and inclusive workplace. We strive to build teams that represent a range of backgrounds, perspectives and skills. Embracing diversity enables us to create a great place to work and to innovate for patients.**Step into the Future of IT with Roche!** As a seasoned Site Reliability Engineer (SRE) at Roche, you will leverage your deep software engineering expertise to propel our products to new heights of robustness, scalability and reliability. This isn't just a role—it's an invitation to shape the backbone of technological innovations forward.**Your Mission** Design and maintain cutting-edge tools, scripts and frameworks that automate repetitive tasks, streamline software deployment and manage expansive systems with unparalleled efficiency. Partner closely with forward-thinking development teams to architect and implement high-performance solutions that elevate system efficiency, optimize resource utilization and enhance deployment processes for superior uptime and user satisfaction.**Your Impact** Lead the charge in incident management and response. Detect system anomalies, troubleshoot swiftly and conduct thorough root cause analyses to prevent recurring issues. Champion continuous improvement by refining monitoring and alerting mechanisms, conducting insightful post-incident reviews and embedding best practices in software lifecycle management. Your strategic foresight and meticulous planning will ensure our systems are not only reliable but also superlatively performant. By joining our elite team, you will play a pivotal role in delivering seamless experiences to our end-users, exceeding business and customer demands, and solidifying Roche's reputation as a leader in IT innovation.**Your Core Responsibilities** Reliability Engineering & Architecture* Define and implement SLIs, SLOs, and error budgets with product and engineering teams* Conduct reliability reviews for new and existing services* Design scalable, fault-tolerant architectures in AWS and Azure environments* Lead capacity planning, performance and cost optimization initiatives* Improve system resilience through automation and self-healing patterns* Drive organizational observability maturity (metrics, logs, traces, alert quality)**Incident Management & Continuous Improvement*** Perform complex root cause analysis and drive rapid mitigation* Participate in blameless postmortems and follow-through* Improve MTTR, reduce incident frequency, and elevate production standards* Collaborate seamlessly with engineering teams to enable timely and effective resolutions* Handle requests and incidents, create and maintain runbooks* Participation in a structured 24\\*7 on-call rotation**Automation & Platform Engineering*** Reduce operational toil through tooling and automation (Python or similar)* Improve CI/CD reliability and deployment safety mechanisms* Build and maintain infrastructure-as-code (Terraform or equivalent)* Enhance Kubernetes platform reliability (EKS, AKS, or similar)**Cross-Functional Leadership*** Partner with business, engineering, security, and cloud teams to embed reliability early in the software development life cycle* Mentor mid-level engineers and help shape SRE best practices* Championing a culture of ownership, accountability, and continuous improvement**Who You Are:*** Minimum bachelor’s degree in computer science, Engineering, or a related field, or equivalent professional experience.* Experience in either site reliability engineering, software engineering or related fields with production on-call experience.* Solid experience with AWS and/or Azure, including setting up, monitoring, and maintaining cloud resources (incl. Kubernetes, EKS, AKS, GKE, etc knowledge).* Proficiency with observability tools* Hands-on experience with incident management tools* Proficiency in scripting languages for automation purposes* Demonstrated proficiency in troubleshooting, especially in cloud and distributed system environments* Excellent communication, teamwork and documentation skills, with a proactive and self-motivated approach to improving system reliability and operational efficiencies.* We value and encourage candidates from diverse backgrounds and experiences, believing that diverse perspectives drive innovation and success.* Excelling in both spoken and written English communication.### Who we areA healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life-changing healthcare solutions that make a global impact.Let’s build a healthier future, together.**Roche is an Equal Opportunity Employer.**
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer
Senior DevOps Engineer

F. Hoffmann-La Roche AG • España

On-site
PHP 5,424,000 - 7,837,000
Senior Software Engineer
Senior Software Engineer

F. Hoffmann-La Roche AG • España

On-site
PHP 5,424,000 - 7,837,000
Career growth opportunities
Inclusive work culture
Senior DevOps Automation Engineer
Senior DevOps Automation Engineer

F. Hoffmann-La Roche AG • España

Hybrid
PHP 3,953,000 - 5,392,000
Hardware Apple novo
Aulas de idiomas
Benefícios de transporte
+1
Expert Software Developer / Architect C#
Expert Software Developer / Architect C#

F. Hoffmann-La Roche AG • España

On-site
PHP 3,498,000 - 4,899,000
Lead UX Designer & Frontend Engineer
Lead UX Designer & Frontend Engineer

F. Hoffmann-La Roche AG • España

On-site
PHP 3,491,000 - 4,889,000
Senior Site Reliability Engineer — Scale, Automate & Uptime
Senior Site Reliability Engineer — Scale, Automate & Uptime

F. Hoffmann-La Roche AG • España

On-site
PHP 4,822,000 - 6,028,000
Sales Data and Systems Specialist
Sales Data and Systems Specialist

F. Hoffmann-La Roche AG • Taguig

On-site
PHP 600,000 - 1,000,000
Finance Enterprise Partner
Finance Enterprise Partner

F. Hoffmann-La Roche AG • Philippines

On-site
PHP 1,800,000 - 2,900,000
Global Category Manager - Product Engineering Services
Global Category Manager - Product Engineering Services

F. Hoffmann-La Roche AG • España

On-site
PHP 4,888,000 - 6,984,000
Safety, Health, Environment (SHE) & Facility Officer
Safety, Health, Environment (SHE) & Facility Officer

Roche • Philippines

On-site
PHP 1,200,000 - 1,800,000