Senior Site Reliability Engineer

Salt

Amsterdam

Hybrid

EUR 83,000 - 124,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Salt is seeking an experienced Senior Site Reliability Engineer to join a high-performance platform team in Amsterdam on a 6-month hybrid contract. You will own end-to-end reliability, lead incident response, and drive automation across CI/CD and IaC to support a cloud migration and a high-volume event processing platform.

You will collaborate with software engineers to improve observability, capacity planning, and platform efficiency, while maintaining strong on-call coverage and ensuring

Qualifications

  • Experience as SRE/DevOps in high-scale production.
  • Strong incident management and root-cause analysis skills.
  • Experience with cloud migrations and distributed systems.

Responsibilities

  • Own end-to-end reliability of production services.
  • Lead incident response, RCA, post-mortems, remediation.
  • Improve observability with monitoring, logging, tracing, dashboards.
  • Build and enhance CI/CD pipelines and IaC solutions.
  • Drive automation to reduce toil and improve efficiency.
  • Support performance testing, capacity planning, and scalability.
  • Contribute to cloud migration of a large event platform.
  • Collaborate with software engineers to improve reliability.
  • Participate in on-call rotation and maintain service availability.

Skills

SRE
Platform Engineer
DevOps
Java
Kafka
AWS
CI/CD
Observability
On-call

Tools

Terraform
Helm
GitOps
Prometheus
Grafana
Datadog
ELK/EFK

Job description

Senior Site Reliability Engineer (SRE) – Travel – Amsterdam
Duration: 6-Month Contract
Start: ASAP
Hybrid: 2 days a week (on-call 1/7)

My client is looking for an experienced Site Reliability Engineer to join a high-performing platform team responsible for operating and evolving a mission-critical event processing platform that handles billions of events every day.

This is an exciting opportunity to work on large-scale distributed systems, drive operational excellence, and support a major cloud migration initiative while ensuring the reliability, scalability, and performance of a business-critical platform.

What You'll Be Doing
  • Own end-to-end reliability of production services.
  • Lead incident response, root cause analysis, post-mortems, and remediation activities.
  • Improve observability through monitoring, alerting, logging, tracing, and dashboards.
  • Build and enhance CI/CD pipelines and Infrastructure-as-Code solutions.
  • Drive automation initiatives to reduce operational toil and improve platform efficiency.
  • Support performance testing, capacity planning, and scalability initiatives.
  • Contribute to the migration of a large-scale event streaming platform to a cloud-native architecture.
  • Collaborate with software engineers and platform teams to improve reliability, resilience, and operational maturity.
  • Participate in a shared on-call rotation and help maintain high service availability.
What We're Looking For
  • Proven experience as a Site Reliability Engineer, SRE, Platform Engineer, or DevOps Engineer in high-scale production environments.
  • Strong hands-on experience with:
  • Java
  • Kafka
  • AWS
  • Terraform, Helm, GitOps, or similar Infrastructure-as-Code tools
  • CI/CD pipelines and deployment automation
  • Experience with modern observability tooling such as Prometheus, Grafana, OpenTelemetry, ELK/EFK, Datadog, or similar.
  • Strong background in incident management, production troubleshooting, and reliability engineering.
  • Experience operating and scaling distributed systems handling high transaction or event volumes.
  • Excellent communication skills and the ability to work effectively across engineering teams.
Nice to Have
  • Experience with Confluent Cloud.
  • Experience migrating Kafka workloads from on-premises environments to cloud platforms.
  • Knowledge of large-scale event-driven architectures and stream processing systems.
  • Experience optimising platform performance, scalability, and operational costs.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Amoria Bond • Rotterdam

On-site
EUR 75,000 - 110,000
Senior SRE – Travel to Amsterdam, Cloud-Scale (6mo)
Senior SRE – Travel to Amsterdam, Cloud-Scale (6mo)

Salt • Amsterdam

Hybrid
EUR 83,000 - 124,000
Site Reliability Engineer - Banking & Finance
Site Reliability Engineer - Banking & Finance

Hamilton Barnes Associates Limited • Amsterdam

On-site
EUR 180,000 - 220,000
Exceptional bonus scheme
Leading company benefits package
SRE
SRE

Gazelle Global • Amstelveen

On-site
EUR 70,000 - 110,000
Senior SRE - Hybrid Amsterdam - Equity
Senior SRE - Hybrid Amsterdam - Equity

DataSnipper • Amsterdam

Hybrid
EUR 90,000 - 140,000
Equity
Pension
Vacation days
+7
Site Reliability Engineer
Site Reliability Engineer

Salt Digital Recruitment • Amsterdam

On-site
EUR 82,656 - 165,312
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Elevation Group • Den Haag

Hybrid
EUR 61,000 - 102,000
Hybrid working
Pension scheme
Personal development budget
+8
Site Reliability Engineer
Site Reliability Engineer

Harnham • Rotterdam

On-site
EUR 57,000 - 95,000
Competitive salary
Benefits package
Ownership of platform reliability
+1
SRE Engineer – AWS, Observability & Automation (6-Month)
SRE Engineer – AWS, Observability & Automation (6-Month)

Amoria Bond • Rotterdam

On-site
EUR 75,000 - 110,000
Sr. Platform Engineer (Core Platform Engineering)
Sr. Platform Engineer (Core Platform Engineering)

Computer Futures • Amsterdam

Hybrid
EUR 96,000 - 165,000