Senior Site Reliability Engineer

deephealth

Amsterdam

Hybrid

EUR 110,000 - 140,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

DeepHealth in the Netherlands is seeking a Senior Site Reliability Engineer to own the reliability, availability, and performance of cloud components deployed at client sites. You will help ensure the DeepHealth solution is scalable, resilient, and secure while supporting the operational team and providing technical leadership within the platform engineering practice.

The role requires hands-on SRE/DevOps experience, strong cloud and containerization skills (GCP/AWS, Docker, Kubernetes, Istio),

Qualifications

  • Fluency in English, both written and spoken.
  • 7+ years hands-on SRE/DevOps experience.
  • Strong software development practices across design, testing, and deployment.
  • Networking and security best practices knowledge.
  • Experience with GitOps and fleet management (ArgoCD).
  • Proficiency with Docker, Kubernetes and Linux systems.
  • Experience with automation tools (Ansible, Terraform) and SQL DB admin.
  • CI/CD pipelines expertise and cloud environments (GCP or AWS).
  • Observability tooling experience (Prometheus, Grafana).
  • Excellent communication and leadership skills.

Responsibilities

  • Own reliability, availability, and performance of cloud components at client sites.
  • Define and monitor SLOs, error budgets, and reliability metrics.
  • Develop automation to reduce toil in deployment, monitoring, and incidents.
  • Design and maintain observability tooling; resolve issues proactively.
  • Write technical specs and ensure regulatory compliance and best practices.
  • Lead incident response and blameless post-mortems with corrective actions.
  • Perform capacity planning and performance tuning for growth.
  • Support deployment of software solutions at customer sites.
  • Provide ongoing maintenance and support to the operations team.
  • Mentor engineers and promote SRE/platform engineering best practices.
  • Maintain secure infrastructure configurations and CI/CD security.

Skills

English fluency
7+ years SRE/DevOps
Cloud experience (GCP or AWS)
Kubernetes & Docker
GitOps (ArgoCD)
CI/CD pipelines
Observability (Prometheus Grafana)
Networking & security best practices
Terraform/Ansible
Linux systems

Tools

Docker
Kubernetes
Istio
KEDA
ArgoCD
Terraform
Ansible

Job description

Job Summary

The Senior Site Reliability Engineer plays a vital role in ensuring the reliability, availability, and performance of DeepHealth software applications, which integrate AI algorithms to deliver clinically relevant information for enhanced decision support. This role takes ownership of the reliability of cloud components deployed at client sites, ensuring that the DeepHealth solution is scalable, resilient, and secure, providing support to the operational team, and providing technical leadership within the platform engineering practice.



Essential Duties and Responsibilities


  • Own the reliability, availability, and performance of cloud components deployed at client sites and of the DeepHealth solution.

  • Define and monitor service level objectives (SLOs), error budgets, and key reliability metrics.

  • Develop and implement automation tools and processes to eliminate toil and streamline deployment, monitoring, and incident response operations.

  • Design and maintain observability tooling (monitoring, logging, alerting, and tracing), and resolve issues before they impact clients.

  • Contribute to the writing of technical specifications and documentation, ensuring compliance with regulatory requirements and industry best practices.

  • Lead incident response, conduct blameless post-mortems, and drive the implementation of corrective and preventive actions.

  • Perform capacity planning and performance tuning to anticipate growth and ensure optimal resource utilization.

  • Support the deployment of software solutions at customer sites, ensuring smooth implementation and optimal performance.

  • Provide ongoing support and maintenance for deployed solutions and to the operational team, addressing any issues or challenges promptly to maintain high levels of customer satisfaction.

  • Mentor and onboard engineers, providing technical leadership in SRE and platform engineering best practices.

  • Implement and maintain secure infrastructure configurations per approved baselines.

  • Ensure CI/CD pipeline security, including integrity verification and access controls.

  • Perform day-to-day technical security controls including system hardening and log monitoring.

  • Document all infrastructure changes and maintain audit trails.



PLEASE NOTE: This is not an exhaustive list of all duties, responsibilities and requirements of the position described above. Other functions may be assigned and management retains the right to add or change duties at any time.



Minimum Qualifications, Education and Experience


  • Fluency in English, both written and spoken.

  • Significant hands-on experience (7+ years) in Site Reliability Engineering or DevOps.

  • In-depth knowledge of software development practices, including design, implementation, testing, and deployment.

  • Strong knowledge of networking and security best practices.

  • Experience with fleet management and GitOps (e.g., ArgoCD).

  • Proficiency in containerization technologies, such as Docker and Kubernetes and its ecosystem (e.g., Istio, KEDA), and with Linux systems.

  • Experience with virtualization technologies, automation tools (such as Ansible or Terraform) and with SQL database administration.

  • Proficiency with continuous integration and continuous deployment (CI/CD) pipelines.

  • Strong proficiency with cloud-based environments (GCP or AWS).

  • Strong experience with observability and monitoring tools (e.g., Prometheus, Grafana).

  • Excellent communication skills, both written and verbal, and strong problem-solving and analytical skills.

  • Strong technical leadership and mentoring abilities



Preferred:


  • Familiarity with the medical device industry and the specific requirements for software applications within this domain.

  • Understanding of AI and machine learning concepts, with the ability to integrate algorithms into software applications.

  • Experience with information security standards (e.g., ISO 27001, SOC 2).



Travel

Occasional travel may be required (typically less than 10%), primarily within Europe, for audits, customer meetings, partner / vendor visits, or company offsites.



Working Environment


  • France or The Netherlands – Remote-friendly.

  • The role can be based in France or The Netherlands with flexible remote working arrangements.

  • There are offices in Paris, Amsterdam and Rotterdam.

  • Periodic on-site presence may be required for team meetings, audits, or workshops.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer, Clinical AI Devices
Software Engineer, Clinical AI Devices

DeepHealth • Amsterdam

Hybrid
EUR 90,000 - 120,000
Technical Services Engineer - Spanish Speaking
Technical Services Engineer - Spanish Speaking

Quantib BV • Amsterdam

Hybrid
EUR 55,000 - 70,000
Technical Services Engineer - German Speaking
Technical Services Engineer - German Speaking

Quantib BV • Netherlands

On-site
EUR 55,000 - 75,000
Technical Support Engineer
Technical Support Engineer

deephealth • Amsterdam

On-site
EUR 35,000 - 50,000
Senior Site Reliability Engineer: Cloud Reliability & Observability Lead
Senior Site Reliability Engineer: Cloud Reliability & Observability Lead

deephealth • Amsterdam

Hybrid
EUR 110,000 - 140,000
Senior DevOps Engineer
Senior DevOps Engineer

Jobgether SRL • Netherlands

Remote
EUR 90,000 - 130,000
Remote-first organization
European-wide teams
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Harnham • Rotterdam

Hybrid
EUR 70,000 - 110,000
Competitive salary
Hybrid working
Exposure to cloud & AI platforms
+1
Project Manager, Clinical AI
Project Manager, Clinical AI

Quantib BV • Amsterdam

On-site
EUR 60,000 - 90,000
Senior Reliability Engineer
Senior Reliability Engineer

Harnham • Rotterdam

Hybrid
EUR 90,000 - 130,000
Competitive salary
Incentives & Benefits
Hybrid work model
Director of Data Governance
Director of Data Governance

DeepHealth • Netherlands

Hybrid
EUR 115,000 - 130,000