Senior Site Reliability Engineer — Remote (AWS, Kubernetes)

Jobgether

Netherlands

On-site

EUR 120,000 - 160,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote within Europe
Technical ownership
Cross-team collaboration

Job summary

Jobgether is seeking a Staff Site Reliability Engineer to lead reliability across AI-driven production environments from a fully remote European base. You will shape resilient infrastructure, drive SRE best practices, and collaborate across platform, product, data, and ML teams to improve availability, performance and security.

Responsibilities include building scalable Kubernetes workloads, implementing IaC with Terraform, and defining SLOs/SLIs with robust observability.

Qualifications

  • Extensive hands-on SRE/Production Engineering experience in large-scale environments.
  • Proven ability to establish or scale SRE practices in high-growth settings.
  • Deep AWS or Azure expertise and modern cloud-native architectures.
  • Strong Kubernetes production experience and security hardening.
  • Infra-as-code mastery with Terraform or similar tooling.
  • Experience designing end-to-end CI/CD pipelines and observability practices.
  • Practical MLOps experience including model deployment and monitoring.

Responsibilities

  • Architect, deploy, operate, and continuously improve scalable, secure production environments.
  • Lead reliability initiatives and standardize SRE practices across teams.
  • Design and optimize Kubernetes infrastructure with production hardening.
  • Define and enforce SLOs/SLIs, error budgets, and observability across systems.
  • Improve release safety, deployment frequency, and incident response.
  • Automate toil and implement scalable processes for enterprise environments.
  • Collaborate with product/data teams to integrate telemetry and ML metrics.

Skills

AWS
Kubernetes
CI/CD pipelines
Observability
MLOps
Distributed systems
Automation
Security hardening

Tools

Terraform

Job description

Jobgether is seeking a Staff Site Reliability Engineer to lead reliability across AI-driven production environments from a fully remote European base. You will shape resilient infrastructure, drive SRE best practices, and collaborate across platform, product, data, and ML teams to improve availability, performance and security.

Responsibilities include building scalable Kubernetes workloads, implementing IaC with Terraform, and defining SLOs/SLIs with robust observability.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE - AWS, Kubernetes & Cloud Reliability (Remote)
Senior SRE - AWS, Kubernetes & Cloud Reliability (Remote)

Cerebras • Amsterdam

On-site
EUR 121,000 - 164,000
Staff Site Reliability Engineer - Remote, AI-Driven Platform
Staff Site Reliability Engineer - Remote, AI-Driven Platform

Tamarind Intelligence • Amsterdam

Hybrid
EUR 140,000 - 210,000
Remote Work
Health Plan wherever you are
Stock Options
+2
Site Reliability Engineer — On-Call & Automation Lead
Site Reliability Engineer — On-Call & Automation Lead

kaiko.ai • Amsterdam

On-site
EUR 60,000 - 80,000
Competitive salary
Good pension plan
25 vacation days
+2
Senior SRE: Kubernetes & AWS Reliability Leader (Hybrid)
Senior SRE: Kubernetes & AWS Reliability Leader (Hybrid)

Manychat • Amsterdam

Hybrid
EUR 70,000 - 90,000
Comprehensive health insurance
Professional development budget
Flexible benefits package
+3
Senior Platform Engineer: Cloud & ML (Remote Europe)
Senior Platform Engineer: Cloud & ML (Remote Europe)

Jobgether • Netherlands

On-site
EUR 32,000 - 46,000
Fully remote Europe
Flexible hours
Professional development budget
+3
SRE Team Lead - Reliability, Automation & Scale
SRE Team Lead - Reliability, Automation & Scale

Together AI • Amsterdam

On-site
EUR 80,000 - 120,000
Remote Site Reliability Engineer: Observability & Cloud Automation
Remote Site Reliability Engineer: Observability & Cloud Automation

Sanders & Creemers • Amsterdam

On-site
Senior SRE: Azure, Hybrid & Automation
Senior SRE: Azure, Hybrid & Automation

Cluster - Data professionals • Den Haag

Hybrid
EUR 61,000 - 102,000
Hybrid work
Home office allowance
Pension scheme
+8
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Triwill Group • Netherlands

On-site
EUR 90,000 - 130,000
Senior Platform Engineer — Kubernetes, AWS & Reliability
Senior Platform Engineer — Kubernetes, AWS & Reliability

Workato • Amsterdam

On-site
EUR 110,000 - 150,000