Cloud Platform Engineer – Data Reliability, Backing Services

Jobtailor

Deutschland

Vor Ort

EUR 90.000 - 120.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Jobtailor is seeking a Platform Engineer to own reliability, performance, and scalability of our data platforms including MySQL, PostgreSQL, Kafka, and search systems. You will define best practices for how product teams use transaction, document, search, messaging, and analytical platforms, and design highly available services with backup, replication, and disaster recovery capabilities.

You will collaborate with cross-functional teams to improve reliability and security, perform capacity

Qualifikationen

  • 4–7 years in Platform Engineering, SRE, Data Platform Engineering, DevOps, or related roles.
  • Hands-on production experience with MySQL, PostgreSQL, Kafka, and at least one of MongoDB or Elasticsearch/OpenSearch.
  • Experience with analytical or distributed data platforms such as StarRocks, ClickHouse, Doris, Druid, Pinot, or similar OLAP systems is highly desirable.
  • Hands-on experience operating stateful workloads in Kubernetes-based environments.
  • Good understanding of high availability, replication, backup/recovery, disaster recovery, capacity planning, and performance tuning.
  • Familiarity with distributed systems concepts including sharding, replication, partitioning, consistency, and query optimization.
  • Experience with major cloud platforms (AWS, GCP, or OCI).
  • Experience automating provisioning, deployment, configuration, monitoring, and lifecycle management using Terraform, Helm, Ansible, GitOps, or similar.
  • Strong scripting or programming skills in Python, Bash, Go, or similar.
  • Experience with observability platforms such as Prometheus/Grafana, ELK/OpenSearch, Datadog, or equivalent.
  • Strong troubleshooting and collaboration, documentation skills; ownership mindset and adaptability.

Aufgaben

  • Own reliability, performance, scalability, and operational health of database, messaging, search, and analytical platforms.
  • Define best practices for how product engineering teams use transactional, document, search, and analytical platforms.
  • Design highly available, scalable platform services including replication, backup, recovery, failover, and DR capabilities.
  • Perform capacity planning, performance tuning, workload reviews, upgrades, patching, and lifecycle management for platform services.
  • Identify and resolve risks such as slow queries, replication lag, and storage saturation.
  • Troubleshoot complex production issues across distributed data platforms.
  • Enable and support Kubernetes-based deployments using cloud‑native patterns and best practices.
  • Use automation, Infrastructure as Code, GitOps, and CI/CD to make backing services repeatable and reliable.
  • Collaborate with Product Engineering, SRE, Security, Data Engineering, and Cloud Platform teams to improve reliability, performance, availability, and security posture.

Kenntnisse

MySQL
PostgreSQL
Kafka
MongoDB
Elasticsearch/OpenSearch
Kubernetes
Terraform
Go
Python
GitOps

Jobbeschreibung

Responsibilities
  • Own reliability, performance, scalability, and operational health of MySQL, PostgreSQL, MongoDB, Elasticsearch/OpenSearch, Kafka, StarRocks, ClickHouse, and similar platforms
  • Define best practices for how product engineering teams use transactional, document, search, messaging, and analytical platforms
  • Design and maintain highly available, scalable, and resilient platform services, including replication, backup, recovery, failover, and disaster recovery capabilities
  • Perform capacity planning, performance tuning, workload reviews, upgrades, patching, and lifecycle management for platform services
  • Identify and resolve risks such as slow queries, hot partitions, consumer lag, replication lag, index growth, retention issues, and storage saturation
  • Troubleshoot and resolve complex production issues related to databases, messaging systems, search platforms, and distributed data platforms
  • Enable and support Kubernetes-based deployments of database, messaging, search, and analytical platforms using cloud‑native patterns and operational best practices
  • Use automation, Infrastructure as Code, GitOps, and CI/CD to make backing services repeatable, reliable, and easier to operate
  • Collaborate with Product Engineering, SRE, Security, Data Engineering, and Cloud Platform teams to improve reliability, performance, availability, and security posture
Requirements
  • 4-7 years of experience in Platform Engineering, SRE, Database Reliability Engineering, Data Platform Engineering, DevOps, or related roles
  • Strong hands‑on experience with MySQL, PostgreSQL, Kafka, and at least one of MongoDB or Elasticsearch/OpenSearch in production environments
  • Experience with analytical or distributed data platforms such as StarRocks, ClickHouse, Apache Doris, Druid, Pinot, or similar OLAP systems is highly desirable
  • Hands‑on experience operating stateful workloads in Kubernetes‑based environments
  • Good understanding of high availability, replication, backup and recovery, disaster recovery, capacity planning, and performance tuning concepts
  • Familiarity with distributed systems concepts including sharding, replication, partitioning, consistency, compaction, backpressure, consumer lag, and query optimization
  • Experience with at least one major cloud platform (AWS, GCP, or OCI)
  • Experience automating provisioning, deployment, configuration, monitoring, and lifecycle management using tools such as Terraform, Helm, Ansible, GitOps, or similar automation frameworks
  • Strong scripting or programming skills in Python, Bash, Go, or similar
  • Experience with observability platforms such as LTGM, Prometheus/Grafana, ELK/OpenSearch, Datadog, or equivalent
  • Strong troubleshooting, problem‑solving, and debugging skills across distributed systems
  • Excellent communication, collaboration, and documentation skills
  • Demonstrated curiosity, ownership mindset, adaptability, and ability to guide product engineering teams
Core Competencies

Demonstrates expertise in managing and optimizing database platforms such as MySQL, PostgreSQL, and Kafka, with a strong focus on high availability, performance tuning, and disaster recovery. Proficient in automation and cloud‑native deployments, ensuring operational excellence across distributed systems.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Principal Platform Engineer – DevOps/Developer Experience
Principal Platform Engineer – DevOps/Developer Experience

Jobtailor • Deutschland

Remote
EUR 120.000 - 180.000
Staff Software Engineer - Cloud Data Storage
Staff Software Engineer - Cloud Data Storage

Embedded Shishya • Deutschland

Remote
EUR 90.000 - 130.000
Senior Software Engineer, Data Platform
Senior Software Engineer, Data Platform

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Senior Site Reliability Engineer – Kubernetes Platform
Senior Site Reliability Engineer – Kubernetes Platform

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Database Platform Engineer
Database Platform Engineer

Jobtailor • Frankfurt

Vor Ort
EUR 90.000 - 130.000
Senior Backend Engineer – Platform
Senior Backend Engineer – Platform

Jobtailor • Frankfurt

Vor Ort
EUR 90.000 - 130.000
Staff Data Engineer - Sponsored Content & Products (all genders)
Staff Data Engineer - Sponsored Content & Products (all genders)

ABOUT YOU SE & Co. KG • Hamburg

Vor Ort
EUR 85.000 - 100.000
Senior Platform Engineer, Backend
Senior Platform Engineer, Backend

Jobtailor • Köln

Vor Ort
EUR 90.000 - 130.000
Senior Engineer, Azure, Terraform, Kubernetes
Senior Engineer, Azure, Terraform, Kubernetes

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Senior Infrastructure Engineer (Core Infra)
Senior Infrastructure Engineer (Core Infra)

United States Digital Space LLC • Berlin

Vor Ort
EUR 120.000 - 180.000