¡Activa las notificaciones laborales por email!

Senior Database Reliability Engineer (DBRE) & Architect (worldwide remote)

CloudLinux

Madrid

A distancia

EUR 60.000 - 90.000

Jornada completa

Hoy
Sé de los primeros/as/es en solicitar esta vacante

Genera un currículum adaptado en cuestión de minutos

Consigue la entrevista y gana más. Más información

Descripción de la vacante

A leading cloud technology company in Spain seeks a Database Engineer with deep expertise in PostgreSQL and ClickHouse. Your role will involve designing self-service platforms, managing analytics clusters, and implementing SRE practices. This position offers fully remote work, flexible hours, 24 days paid vacation, and support for professional development. Join us to impact services used by thousands globally.

Servicios

Professional development support
Paid vacation, national holidays, and sick leaves
Medical insurance compensation
Co-working and gym reimbursement
Education budget

Formación

  • 5+ years of PostgreSQL expertise including MVCC internals and lock mechanics.
  • Experience operating large ClickHouse clusters and understanding ZooKeeper.
  • Skill in writing complex Terraform modules and Ansible roles.

Responsabilidades

  • Design and implement a self-service platform based on Terraform and Ansible.
  • Manage exponentially growing analytics clusters with ClickHouse.
  • Implement SRE practices and automated self-healing mechanisms.

Conocimientos

Deep PostgreSQL Expertise
ClickHouse Mastery
Engineering Mindset (SRE / DevOps)
Hybrid Environment Experience
Systems Approach

Herramientas

Terraform
Ansible
Python
Go
Descripción del empleo
Responsibilities
  • DBaaS Architecture : Design and implement a self-service platform based on Terraform and Ansible , enabling the deployment of HA clusters (PostgreSQL and ClickHouse, MongoDB, Redis) in a heterogeneous environment (Bare Metal + OpenNebula + Kubernetes + Public Clouds). You will turn infrastructure into a product.
  • Scaling ClickHouse : Manage exponentially growing analytics clusters (12+ clusters, tens of terabytes of data). You will tackle sharding, table engine optimization (ReplicatedMergeTree), and building reliable S3 backup pipelines under high load.
  • Data Platform & Analytics Support : Maintain and scale the infrastructure for Apache Airflow and Redash. You will ensure the reliability of ETL pipelines and visualization tools, bridging the gap between raw infrastructure and the data analytics team.
  • Reliability as Code : Implement SRE practices in data management. Replace manual incident response with automated self-healing mechanisms. Define and implement SLO / SLI for all databases.
  • Stack Modernization : Lead the migration process from legacy solutions to modern cloud patterns. Participate in decision‑making regarding the implementation of Kubernetes operators for stateful workloads.
  • Expertise & Mentorship : Serve as the technical authority for product teams, helping them optimize data schemas and SQL queries for high‑load systems.
Our Tech Stack :
  • Databases : PostgreSQL 15+ (Patroni, PgBouncer), ClickHouse (Sharded / Replicated), MongoDB, Redis, Kafka
  • Data & Analytics : Apache Airflow, Redash (Infrastructure & Integration).
  • Infrastructure : Own 3+DC colocation (OpenNebula, Kubernetes, Bare Metal), AWS, Google Cloud, Azure, DO – Hybrid Cloud.
  • Automation & IaC : Terraform, Ansible, Python / Go, GitLab, Jenkins, Gerrit.
  • Observability : VictoriaMetrics, Grafana, Loki.
Why CloudLinux?
  • Culture : A Remote‑first company with an "Employees First" principle. We value results, not hours in the office.
  • Impact : Your architectural decisions will determine the stability of services used by thousands of companies around the world.
  • Growth : We support professional development and pay for training and conferences.
Requirements
What We Expect From You :
  • Deep PostgreSQL Expertise (5+ years) : You know MVCC internals, understand locking mechanics, can configure Patroni and PgBouncer "with your eyes closed," and have experience with seamless major version upgrades under load.
  • ClickHouse Mastery : Experience operating large clusters, understanding ZooKeeper / ClickHouse Keeper, sharding, replication internals, and the ability to diagnose performance issues at the data-part level.
  • Engineering Mindset (SRE / DevOps) : You hate doing the same task twice by hand. Experience writing complex Terraform modules and Ansible roles is mandatory. Programming skills in Python or Go for automation are a huge plus.
  • Hybrid Environment Experience : You understand the differences between running DBs on Bare Metal vs. Kubernetes vs. Cloud and know how to optimize TCO and disk subsystem performance (NVMe, Network Storage).
  • Systems Approach : You see the big picture - from the network packet to the application business logic. You understand the importance of security (FIPS, Audit logs) and Disaster Recovery.
Nice to Have :
  • Experience building an Internal Developer Platform (IDP).
  • Experience operating databases in Kubernetes (CloudNativePG, Altinity Operator).
  • Experience working in Cloud and Hosting providers on similar services.
Benefits
What's in it for you?
  • A focus on professional development.
  • Interesting and challenging projects.
  • Fully remote work with flexible working hours, which allows you to schedule your day and work from any location worldwide.
  • Paid 24 days of vacation per year, 10 days of national holidays, and unlimited sick leaves.
  • Compensation for private medical insurance.
  • Co‑working and gym / sports reimbursement.
  • Budget for education.
  • The opportunity to receive a reward for the most innovative idea that the company can patent.

By applying for this position, you agree with and give us your consent to maintain and process your personal data with this respect. Please read our Privacy Policy for more information.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra un archivo en formato PDF, DOC, DOCX, ODT o PAGES de hasta 5 MB.