Senior Site Reliability Engineer

Forto

Berlin

Vor Ort

EUR 90.000 - 140.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Forto is seeking a senior Site Reliability Engineer to enhance reliability and developer experience across our global logistics platform. You will shape the runtime platform, empower 70+ engineers with self-service tooling, and drive platform enhancements in collaboration with product teams.

We value strong backend/infra experience, IaC, GitOps, and mature SLOs, enabling cost optimization and improved security.

Qualifikationen

  • 5+ years in backend or infrastructure engineering, with SRE/Platform focus.
  • Hands-on experience with GCP/AWS, Kubernetes, Terraform, and Helm in production.
  • Strong software development background building frameworks and internal tooling.
  • Experience with observability platforms (Datadog) at scale, including alerting and SLOs.
  • Proven track record designing and operating distributed systems.

Aufgaben

  • Build out the runtime platform as a self-service product for engineering teams.
  • Advance platform engineering practices including code quality and test-driven development.
  • Own the developer portal and CI/CD overhaul in collaboration with product teams.
  • Ensure site reliability with observability, deployment, and disaster recovery capabilities.
  • Own end-to-end reliability standards via SLOs and error budgets.
  • Drive cost optimization across Kubernetes, databases, and monitoring at scale.
  • Improve security posture through tooling and collaboration with security stakeholders.
  • Collaborate with the broader engineering org on platform architecture and tech refresh.
  • Improve developer productivity via platform services and tooling.
  • Serve as second line of defense for incidents and on-call rotation.

Kenntnisse

Backend engineering
SRE
Platform engineering
GCP/AWS
Kubernetes
Terraform
Helm
Datadog observability
IaC / GitOps
SLOs & error budgets
Distributed systems

Tools

Kubernetes
Terraform
Helm
GCP
Datadog

Jobbeschreibung

About Us

What if your work could drive change in a globally established industry, shaping processes that touch every corner of the world? At Forto, we are at the forefront of change, harnessing the power of AI to revolutionise logistics. We want to reinvent digital supply chains to be transparent, frictionless and sustainable. From day one, our mission has been to simplify global trade – creating a seamless and efficient logistics process.

Your role & Mission

The Site Reliability Engineering team at Forto is responsible for reliability and developer experience. We enable our development teams to write complex business logic by providing best‑in‑class tooling and infrastructure. We have a production environment based on GCP, Kubernetes, Terraform, and Helm. On top of that, we have self‑service tooling written in TypeScript. This is a high‑ownership role on a lean team that directly shapes how 70+ engineers build and ship software. If you care about platform quality and want your work felt immediately across an engineering org, this is a great match for you.

What You Will Do
  • Build out our runtime platform as a self‑service product that enables our engineering teams to write code, run workloads, and drive engineering culture forward.
  • Bring software development skills and practices into platform engineering, such as code quality, domain‑driven design, and test‑driven development.
  • Own the developer portal and internal platform roadmap, including leading this year’s overhaul of our CI/CD pipelines in collaboration with all product teams.
  • Ensure site reliability by building observability solutions, deployment, and disaster recovery capabilities.
  • Own reliability standards end‑to‑end through SLOs and error budgets – shaping how teams balance velocity and risk.
  • Drive infrastructure cost optimisation across Kubernetes, MongoDB, and Datadog at scale.
  • Improve our security posture through tooling, compliance work, and partnership with security stakeholders.
  • Work closely with the entire Engineering function as a steward of platform architecture – embracing new technologies and cleaning up old ones.
  • Improve developer productivity by working with engineers on platform services and developer tooling.
  • Serve as a second line of defense for incidents and be the secondary on‑call for our developers.
Required Skills And Experience
  • 5+ years in backend or infrastructure engineering, with at least 2 years in an SRE or platform engineering role.
  • Hands‑on experience with GCP/AWS, Kubernetes, Terraform, and Helm in a production environment.
  • Strong software development background – building frameworks, internal tooling, and infrastructure.
  • Experience with an observability platform (Datadog or equivalent) at scale – not just dashboards, but alerting strategy, cost management, and SLO instrumentation.
  • Experience defining and operating SLOs and error budgets as a reliability mechanism, not just as metrics.
  • Solid understanding of Infrastructure as Code (IaC) and a GitOps‑first mindset.
  • Proven track record designing, developing, and troubleshooting complex distributed systems.
Why work with us?

Our team is hard‑working, constantly seeking to maximise the impact of their work, but we put our people first, always winning with care. We value efficient systems and swift, direct communication. We want everyone to have their time to speak, so that we can embrace diverse perspectives to help drive towards solutions always.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Site Reliability Engineer (m/w/d)
Senior Site Reliability Engineer (m/w/d)

Impower • München

Hybrid
EUR 70.000 - 90.000
Flexible hours
Ownership in projects
Diverse team culture
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Stuttgart

Vor Ort
EUR 70.000 - 100.000
Senior Site Reliability Engineer (all genders)
Senior Site Reliability Engineer (all genders)

FACT-Finder • Pforzheim

Hybrid
EUR 90.000 - 140.000
Hybrid work model
Impact on product reliability
Competitive compensation
+1
(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz
(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz

United States Digital Space LLC • Deutschland

Hybrid
EUR 60.000 - 90.000
Hybrid working arrangements
Flexible hours
Subsidized sports or yoga courses
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobgether • Deutschland

Vor Ort
EUR 90.000 - 130.000
Fully remote
High ownership
Tech exposure
Senior Site Reliability Engineer – Kubernetes Platform
Senior Site Reliability Engineer – Kubernetes Platform

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Lead, Site Reliability Engineer
Lead, Site Reliability Engineer

CardWorks • Deutschland

Hybrid
EUR 126.000 - 140.000
Competitive Pay
Bonus Program
Benefits Package
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

B Capital • Deutschland

Remote
EUR 46.000 - 105.000
Work from anywhere
Flexible paid time off
Mental health support services
+3
Senior Infrastructure Engineer (Core Infra)
Senior Infrastructure Engineer (Core Infra)

United States Digital Space LLC • Berlin

Vor Ort
EUR 120.000 - 180.000
Site Reliability Engineer
Site Reliability Engineer

Contorion • Berlin

Hybrid
EUR 60.000 - 80.000
Flexible schedule
30 days of vacation
Subsidized job ticket or bike
+2