Site Reliability Engineer

Harnham

Rotterdam

On-site

EUR 57,000 - 95,000

Full time

48 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Competitive salary
Benefits package
Ownership of platform reliability
AI operations exposure

Job summary

Harnham in Rotterdam is seeking a Senior Site Reliability Engineer to own AI platform operations and overall service reliability in a fast-moving technology environment.

You will shape SLOs/SLIs, lead observability, and drive CI/CD improvements while optimizing cloud costs and scalability. This role offers hands-on ownership, a collaborative international team, and opportunities to work with cutting-edge cloud and AI technologies.

Qualifications

  • Extensive experience operating high-traffic distributed production systems.
  • Strong SRE principles, observability, and platform operations.
  • Hands-on with GCP and IaC (Terraform).
  • Proactive automation and incident management mindset.

Responsibilities

  • Define and manage Service Level Objectives (SLOs), SLIs, and error budget policies across critical services.
  • Lead initiatives across observability, distributed tracing, monitoring, and incident response.
  • Improve deployment safety through CI/CD best practices, automated rollbacks, and progressive delivery techniques.
  • Own AI platform operations, including runtime performance, scalability, reliability, and cost optimisation.
  • Drive FinOps initiatives across cloud infrastructure, identifying opportunities to improve efficiency and manage costs.
  • Lead incident management activities and develop automated safeguards to prevent recurring issues.
  • Enhance resilience through capacity planning, load testing, disaster recovery planning, and service reliability improvements.
  • Manage infrastructure through Infrastructure as Code and cloud automation practices.
  • Build tooling, runbooks, and self-service capabilities that improve the developer experience and reduce operational overhead.

Skills

SRE fundamentals
Observability
Monitoring
Linux
CI/CD
Distributed systems
Automation
FinOps awareness
Cloud-native platforms

Tools

Google Cloud Platform
Terraform
Kubernetes

Job description

Senior Site Reliability Engineer (SRE & AI Platform Operations)
Rotterdam, 5 days in office
Up to €95k annually + benefits

This is an opportunity to take ownership of reliability, observability, cloud infrastructure, and AI platform operations within a fast-moving technology environment. You will play a key role in building the guardrails, automation, and operational excellence that enable engineering teams to innovate at pace while maintaining stability and performance.

The Company

They are a technology-driven organisation undergoing significant platform transformation, modernising core systems and investing heavily in cloud-native architecture and AI-enabled capabilities. Their engineering culture is focused on innovation, automation, and continuous improvement. You will join a collaborative environment where technical expertise is valued and where your work will have a direct impact on business performance and customer experience.

The Role
  • Define and manage Service Level Objectives (SLOs), SLIs, and error budget policies across critical services.
  • Lead initiatives across observability, distributed tracing, monitoring, and incident response.
  • Improve deployment safety through CI/CD best practices, automated rollbacks, and progressive delivery techniques.
  • Own AI platform operations, including runtime performance, scalability, reliability, and cost optimisation.
  • Drive FinOps initiatives across cloud infrastructure, identifying opportunities to improve efficiency and manage costs.
  • Lead incident management activities and develop automated safeguards to prevent recurring issues.
  • Enhance resilience through capacity planning, load testing, disaster recovery planning, and service reliability improvements.
  • Manage infrastructure through Infrastructure as Code and cloud automation practices.
  • Build tooling, runbooks, and self-service capabilities that improve the developer experience and reduce operational overhead.
Your Skills & Experience
  • Strong commercial experience operating high-traffic, distributed production systems.
  • Deep knowledge of Site Reliability Engineering principles, monitoring, and platform operations.
  • Hands-on experience with Google Cloud Platform and Infrastructure as Code using Terraform.
  • Strong troubleshooting capabilities across Linux environments, cloud infrastructure, containers, databases, and distributed systems.
  • Experience with observability tooling, monitoring platforms, and distributed tracing.
  • Familiarity with modern software engineering environments and cloud-native architectures.
  • Understanding of AI workloads, LLM integrations, API-driven services, or automated processing pipelines.
  • A proactive, solutions-focused approach with a passion for automation and operational excellence.
What They Offer
  • Competitive salary and benefits package.
  • The opportunity to shape platform reliability and AI operations at scale.
  • Real ownership and influence across engineering and infrastructure strategy.
  • An international and collaborative working environment.
  • Clear opportunities for professional growth and career progression.
  • The chance to work on cutting-edge cloud and AI technologies.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cluster - Data professionals • Den Haag

Hybrid
EUR 61,000 - 102,000
Hybrid work
Home office allowance
Pension scheme
+8
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Elevation Group • Den Haag

Hybrid
EUR 61,000 - 102,000
Hybrid working
Pension scheme
Personal development budget
+8
Senior SRE: AI Platform & Cloud Reliability Lead
Senior SRE: AI Platform & Cloud Reliability Lead

Harnham • Rotterdam

On-site
EUR 57,000 - 95,000
Competitive salary
Benefits package
Ownership of platform reliability
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

TOPdesk • Delft

Hybrid
EUR 100,000 - 130,000
Hybrid work environment
10 to Grow programme
Excellent employment conditions
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Topdesk-7 • Delft

Hybrid
EUR 59,000 - 86,000
Site Reliability Engineer
Site Reliability Engineer

Amoria Bond • Rotterdam

On-site
EUR 75,000 - 110,000
Site Reliability Engineer (EU Remote) - AI infrastructure
Site Reliability Engineer (EU Remote) - AI infrastructure

Hamilton Barnes Associates Limited • Amstelveen

Remote
EUR 180,000 - 220,000
IPO Equity
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Elevation Group • Den Haag

Hybrid
EUR 113,000 - 138,000
Hybrid working
Home office allowance
Pension scheme with employer contrib.
+3
Head of Platform Engineering
Head of Platform Engineering

Harnham • Randstad

On-site
EUR 120,000 - 180,000
Senior AI Engineer
Senior AI Engineer

Harnham • Rotterdam

On-site
EUR 90,000 - 120,000
Competitive salary
International team in Rotterdam
Senior leadership exposure
+1