Site Reliability Engineer

kaiko.ai

Amsterdam

On-site

EUR 60,000 - 80,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Good pension plan
25 vacation days
Learning and development budget
Annual commuting subsidy

Job summary

Kaiko.ai in Amsterdam is looking for a Site Reliability Engineer to ensure platform reliability through on-call support and incident resolution. In this role, you will manage alerts, drive incidents to resolution and enhance the SRE programme.

Ideal candidates will have significant experience in Kubernetes and Linux, solid programming skills, plus a collaborative approach to problem-solving. The position offers competitive salary, benefits, and the autonomy to work flexibly.

Qualifications

  • Deep Kubernetes experience with debugging under pressure.
  • Strong Linux fundamentals for OS-level troubleshooting.
  • Solid programming ability for developing automation.

Responsibilities

  • Carry on-call, triage alerts, and drive incidents to resolution.
  • Transform noisy alerts into actionable signals.
  • Contribute to blameless post-mortems and follow through on action items.

Skills

Kubernetes
Linux fundamentals
Programming ability
Observability and incident-response
SRE principles
Production SRE experience

Tools

Terraform
Helm

Job description

About kaiko

Kaiko is building a next‑generation agentic clinical AI assistant that helps clinicians reason across patient data, guidelines, and diagnostics.

About the role

This is a frontline reliability role focused on keeping the platform healthy through on‑call, incident response, and continuous improvement of the SRE programme. You’ll have room to shape how reliability works at kaiko rather than inherit a rigid process.

Responsibilities
  • Own the reactive frontline: carry on‑call, triage alerts quickly and methodically, and drive incidents to resolution with clear communication and clean handoffs.
  • Make alerts trustworthy: treat noisy or low‑value alerts as defects to fix, not background noise; move toward structured, queryable signals and precise, actionable alerting.
  • Turn incidents into durable fixes: run and contribute to blameless post‑mortems, then follow through across teams so action items land in our services.
  • Strengthen the SRE programme: write automation and tooling that removes toil, improves runbooks and dashboards, and leaves the on‑call rotation better documented.
Qualifications
  • Deep Kubernetes experience: debug real failure modes under pressure—scheduling, networking, resource pressure, control‑plane versus workload issues.
  • Strong Linux fundamentals: comfortable debugging at the OS level—processes, networking, filesystems, resource limits.
  • Solid programming ability: eliminate toil, develop maintainable automation and tooling; experience building products shows good problem‑solving.
  • Observability and incident‑response fluency: comfortable in metrics, logs, traces; write queries to isolate problems and stay composed during incidents.
  • Internalized SRE principles: SLO/SLI, error‑budget thinking, bias toward prevention, instinct for reducing toil.
  • Experience of roughly 3‑6 years in production SRE, platform, or infrastructure roles—skill demonstrated over experience.
Soft Skills
  • Solution‑oriented and pragmatic; willing to build end‑to‑end solutions when needed.
  • Collaborative communicator who coaches teams and writes clear, actionable guidance.
  • Bias to automate and remove toil.
Nice to Have
  • Experience with structured logging and modern alerting practices.
  • Infrastructure‑as‑code and CI/CD fluency (Terraform, Helm, GitOps, etc.).
  • Familiarity with incident‑management tooling (Rootly, PagerDuty, incident.io) and related post‑mortem discipline.
  • Exposure to regulated or high‑stakes domains (health, fintech, critical infrastructure).
Benefits
  • Competitive salary, good pension plan and 25 vacation days per year.
  • Great offsites and team events.
  • EUR 1000 learning and development budget.
  • Autonomy to work flexible style.
  • Annual commuting subsidy.
Equal Employment Opportunity

Kaiko is committed to building a diverse team and to an inclusive, equitable hiring process. We welcome applicants of all backgrounds.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer — On-Call & Automation Lead
Site Reliability Engineer — On-Call & Automation Lead

kaiko.ai • Amsterdam

On-site
EUR 60,000 - 80,000
Competitive salary
Good pension plan
25 vacation days
+2
Senior Clinical AI Data Lead
Senior Clinical AI Data Lead

kaiko.ai • Amsterdam

Hybrid
EUR 90,000 - 110,000
25 vacation days per year
Great offsites and team events
EUR 1000 learning and development预算
+1
Head of Engineering
Head of Engineering

kaiko.ai • Amsterdam

On-site
EUR 80,000 - 110,000
Attractive competitive salary
Good pension plan
25 vacation days per year
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Cluster - Data professionals • Den Haag

Hybrid
EUR 61,000 - 102,000
Hybrid work
Home office allowance
Pension scheme
+8
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Topdesk-7 • Delft

Hybrid
EUR 59,000 - 86,000
Site Reliability Engineer, Mistral Cloud
Site Reliability Engineer, Mistral Cloud

Mistral • Amsterdam

On-site
EUR 90,000 - 130,000
Content Lead
Content Lead

kaiko.ai • Amsterdam

Hybrid
EUR 70,000 - 100,000
Pension plan
25 vacation days
Learning & development budget
+2
Data ML Engineering Lead
Data ML Engineering Lead

kaiko.ai • Amsterdam

Hybrid
EUR 80,000 - 120,000
Competitive salary
Good pension plan
25 vacation days per year
+2
Partnerships Lead, API Partnerships
Partnerships Lead, API Partnerships

kaiko.ai • Amsterdam

On-site
EUR 120,000 - 180,000
Competitive salary
25 vacation days per year
Learning and development budget
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Manychat • Amsterdam

Hybrid
EUR 70,000 - 90,000
Comprehensive health insurance
Professional development budget
Flexible benefits package
+3