SRE - DataPlatform

Veepee

Paris

Hybride

EUR 50 000 - 75 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Variable bonus
Dynamic and creative environment
E-learning courses
Participation in meetups and conferences
Flexible Office with up to 2 days at home

Résumé du poste

Veepee is looking for a Site Reliability Engineer (SRE) to enhance the reliability and efficiency of their modern data platform. The role involves contributing to the migration towards a new lakehouse architecture while ensuring high performance of data services.

The ideal candidate will have strong experience in Kubernetes and SRE principles. Benefits include a dynamic environment, flexible working options, and opportunities for professional development.

Qualifications

  • Strong experience with Kubernetes in production environments.
  • Experience with distributed data systems or strong willingness to learn.
  • Solid understanding of SRE principles including monitoring and alerting.
  • Familiarity with Infrastructure as Code concepts.
  • Experience with observability tools like Prometheus and Grafana.

Responsabilités

  • Ensure reliability and performance of our data platform services.
  • Define and implement SRE best practices: SLIs/SLOs and observability.
  • Operate services running on Kubernetes and automate infrastructure provisioning.
  • Contribute to cloud migration to VeepeeCloud lakehouse stack.
  • Collaborate with teams to optimize compute/storage usage.

Connaissances

Kubernetes in production environments
Distributed data systems
SRE principles
Infrastructure as Code (Terraform)
GitOps workflows
Observability tools (Prometheus, Grafana)
Working in cloud environments
Collaboration mindset
Fluent in English

Outils

Terraform
Prometheus
Grafana

Description du poste

Being an SRE at VeepeeTech means being part of a transversal SRE community while integrating a product-oriented Data Platform team.

You will contribute to the reliability, scalability, and operability of critical data services by applying SRE and DevOps practices, while sharing knowledge across teams.

The Data Platform is currently evolving toward a modern lakehouse architecture deployed on VeepeeCloud (our on-prem platform), based on technologies such as Trino, Iceberg, and object storage, with strong ambitions around performance, cost efficiency, and platform ownership.

You will work in a distributed environment (France & Spain), within a team of 40-50 data professionals across engineering, analytics, data science, and governance.

You will play a key role in ensuring the reliability and scalability of this next-generation data platform, while supporting the transition from public cloud to hybrid/on-prem architectures.

Tasks
Platform Reliability & Operations
  • Ensure reliability and performance of our data platform services (Trino, Iceberg, S3, Kafka, Flink)
  • Define and implement SRE best practices: SLIs/SLOs, error budgets, observability
  • Build and maintain monitoring, alerting, and incident response frameworks (Prometheus, Grafana, etc.)
Cloud Migration & Architecture
  • Contribute to the migration from public datawarehouse cloud to VeepeeCloud lakehouse stack
  • Support coexistence between cloud and on-prem systems and ensure consistency and reliability
  • Help design resilient architectures for ingestion, transformation, and serving layers
Kubernetes & Infrastructure
  • Operate and improve services running on Kubernetes (GKE/EKS & on-prem clusters)
  • Automate infrastructure provisioning using Terraform, Atlantis, and/or Crossplane
  • Improve GitOps workflows for platform deployment and configuration
FinOps & Performance Optimization
  • Collaborate with teams to optimize compute/storage usage (Trino queries, BigQuery slots, etc.)
  • Build tools and dashboards to track cost, usage, and efficiency
  • Support the transition toward cost-efficient on-prem workloads
Developer Enablement
  • Improve self-service capabilities for data teams (e.g., provisioning Trino/Iceberg resources)
  • Help teams adopt best practices in reliability, observability, and deployment
  • Write clear technical documentation and runbooks
Resilience & DRP
  • Contribute to Disaster Recovery Plan (DRP) definition and implementation
  • Ensure multi-DC resilience (FR1 / NL1) and data replication strategies
  • Participate in incident management and postmortems
MUST HAVE Skills
  • Strong experience with Kubernetes in production environments
  • Experience with distributed data systems (or strong willingness to learn)
  • Solid understanding of SRE principles (monitoring, alerting, SLAs/SLOs)
  • Experience with Infrastructure as Code (Terraform or similar)
  • Familiarity with GitOps workflows
  • Experience with observability tools (Prometheus, Grafana, logging systems)
  • Comfortable working in cloud environments
  • Strong collaboration mindset and ability to work across teams
  • Fluent in English
NICE TO HAVE Skills
  • Experience with Trino, Iceberg, or data lakehouse architectures
  • Experience with Ceph S3 or object storage systems
  • Knowledge of Kafka / Flink / Airflow
  • Experience with FinOps practices and cost optimization
  • Experience with Crossplane or platform self-service models
  • Programming skills (Python, Java, or Go)
  • Experience with multi-region / multi-DC architectures
Benefits
  • Variable bonus
  • The dynamic and creative environment within international teams
  • The variety of self-education courses on our e-learning platform
  • Participation in meetups and conferences locally and internationally
  • Flexible Office with up to 2 days at home

For the service of diversity and inclusion, Veepee is committed to reviewing all applications received on an equal basis.

For more information about our ecosystem: https://careers.veepee.com/en/home-page-en/

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

SRE (DataPlatform)
SRE (DataPlatform)

Veepee • Paris

Hybride
EUR 65 000 - 90 000
Health Insurance
Up to 2 days of remote work per week
E-learning platform access
SRE
SRE

Veepee • Paris

Sur place
EUR 50 000 - 70 000
Dynamic and creative environment
E-learning courses
Participation in meetups and conferences
+3
SRE / DevOps Engineer - CDI - H/F/X
SRE / DevOps Engineer - CDI - H/F/X

Veepee • Paris

Hybride
EUR 50 000 - 80 000
Health insurance
Self-education courses
Participation in meetups and conferences
+1
Software Engineer Go, Rust or Scala (Infrastructure - Foundation) - Freelance (H/F/X)
Software Engineer Go, Rust or Scala (Infrastructure - Foundation) - Freelance (H/F/X)

Veepee • Paris

Hybride
EUR 85 000 - 130 000
Variable bonus
International teams environment
E-learning courses
+2
Software Engineer Go, Rust or Scala (Infrastructure) - Foundation (H/F/X)
Software Engineer Go, Rust or Scala (Infrastructure) - Foundation (H/F/X)

Veepee • Paris

Sur place
EUR 90 000 - 130 000
Variable bonus
Flexible Office
International teams
+2
Hybrid SRE: Data Platform Lakehouse & Reliability
Hybrid SRE: Data Platform Lakehouse & Reliability

Veepee • Paris

Hybride
EUR 65 000 - 90 000
Health Insurance
Up to 2 days of remote work per week
E-learning platform access
Senior Data Scientist
Senior Data Scientist

Veepee • Paris

Hybride
EUR 65 000 - 85 000
Variable bonus
Dynamic international work environment
Access to self-learning courses
+2
SRE Data Platform Engineer – Hybrid Lakehouse
SRE Data Platform Engineer – Hybrid Lakehouse

Veepee • Paris

Hybride
EUR 50 000 - 75 000
Variable bonus
Dynamic and creative environment
E-learning courses
+2
Remote-Friendly SRE: Build Resilient, Scalable Infra
Remote-Friendly SRE: Build Resilient, Scalable Infra

Veepee • Paris

Hybride
EUR 50 000 - 70 000
Dynamic and creative environment
E-learning courses
Participation in meetups and conferences
+3
Security Engineer - Purple Team specialist H/F
Security Engineer - Purple Team specialist H/F

Veepee • Paris

Hybride
EUR 60 000 - 85 000
Variable bonus
International teams
Self-education platform
+3