Head of Engineering - GPU Cloud

Scaleway

Paris

Hybride

EUR 140 000 - 190 000

Plein temps

Il y a 4 jours
Soyez parmi les premiers à postuler

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Hybrid work
Office spaces
Meal service
Well-being benefits
International environment
Career mobility

Résumé du poste

Scaleway seeks a senior engineering leader to drive GPU Cloud infrastructure, leading Support Engineering, HPC, and SRE teams. You will own architecture decisions, validate proposals, and guide large-scale GPU cluster deployments with a focus on reliability and performance.

You will manage 14 engineers across two squads within the GPU Cloud organization, reporting to the SVP and collaborating with GTM, Operations, and Product to deliver scalable AI and HPC platforms.

Qualifications

  • 10+ years in infrastructure engineering with leadership of senior technical teams.
  • Experience designing, deploying, or operating large-scale compute clusters.
  • Expertise in GPU and HPC environments (NVIDIA/AMD).
  • Strong distributed systems, cluster architecture, and production ops knowledge.
  • Experience with Kubernetes, Proxmox, Warewulf, Prometheus, Grafana.
  • Knowledge of high-performance networking and storage (InfiniBand, Lustre, DDN, VAST).
  • Ability to translate complex constraints into engineering and business decisions.

Responsabilités

  • Lead Support Engineering, HPC, and SRE orgs; manage two Engineering Managers for 14 engineers.
  • Own technical strategy, architecture, and key technology choices for GPU Cloud infra.
  • Review and validate technical and service dimensions of strategic proposals.
  • Oversee design, deployment, and readiness of new GPU clusters.
  • Drive evolution of cluster management, capacity, automation, and ops capabilities.
  • Ensure reliability, scalability, performance, and maintainability of GPU infra.
  • Provide leadership on complex AI and HPC infra projects.
  • Align Engineering with GTM, Product, and Operations; develop teams long-term.

Connaissances

Leadership
Architecting complex infra
Team management
Strategic decision making
Communication with stakeholders
GPU/HPC infrastructure
Distributed systems
Kubernetes / orchestration
Monitoring & observability
Networking (InfiniBand, NICs)

Formation

Bachelor's or higher in CS/Engineering

Outils

Kubernetes
Proxmox
Warewulf
Prometheus
Grafana
NVIDIA GPU tech
AMD GPU tech
Lustre/Storage tech
InfiniBand networking

Description du poste

OUR STORY:

Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

Our offices are located in Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.

WHY WE NEED YOU ?

As our GPU Cloud business continues to scale, we are strengthening our engineering leadership to support the deployment and operation of increasingly large and complex GPU clusters.

Your mission will be to lead our Support Engineering, HPC, and SRE teams, own key technology and architecture decisions, and ensure that our most strategic GPU infrastructure projects are successfully designed, delivered, and operated.

You will also play a critical role in validating the technical feasibility of commercial proposals and ensuring that the commitments we make to customers can be delivered reliably on our sovereign cloud infrastructure.

YOUR FUTURE TEAM

We work in a collaborative and international environment where the diversity of Scalers, combined with a strong culture of knowledge sharing, helps us bring ambitious projects to life.

You will lead an organization of 14 engineers across two squads, each managed by an Engineering Manager reporting directly to you.

As part of the broader GPU Cloud organization, reporting to the SVP GPU Cloud, you will work closely with GTM, Operations, Service Management, Product, and other engineering teams to build and operate large-scale AI and HPC infrastructure.

YOUR DAILY ROUTINE

Tasks

  • Lead the Support Engineering, HPC, and SRE organizations, directly managing two Engineering Managers responsible for 14 engineers
  • Own the technical strategy, architecture, and key technology choices for GPU Cloud infrastructure
  • Review and validate the technical and service dimensions of strategic commercial proposals, ensuring commitments are realistic and deliverable
  • Oversee the design, deployment, and operational readiness of new GPU clusters
  • Drive the evolution of our cluster management, capacity management, automation, and operational capabilities
  • Ensure the reliability, scalability, performance, and maintainability of our GPU infrastructure
  • Provide technical leadership on complex AI and HPC infrastructure projects
  • Build strong alignment between Engineering, GTM, Product, and Operations
  • Develop the engineering organization through clear direction, effective delegation, coaching, and long-term team development
  • Establish and maintain high standards of engineering rigor, operational excellence, and technical decision-making
ABOUT YOU
HARDSKILLS:
  • 10+ years of experience in infrastructure engineering, including significant experience leading senior technical teams and managers
  • Proven experience designing, deploying, or operating large-scale infrastructure and compute clusters
  • Expertise in GPU and/or HPC environments, ideally involving NVIDIA and AMD technologies
  • Strong understanding of distributed infrastructure, cluster architecture, reliability, and production operations
  • Experience with orchestration, provisioning, and observability technologies such as Kubernetes, Proxmox, Warewulf, Prometheus, and Grafana
  • Strong knowledge of high-performance networking technologies such as InfiniBand, NVIDIA Spectrum-X, and Broadcom Tomahawk
  • Experience with high-performance and distributed storage technologies such as Lustre, DDN, and VAST Data
  • Ability to assess architectural trade-offs and translate complex technical constraints into clear engineering and business decisions
SOFT SKILLS:
  • Strong leadership skills with the ability to lead experienced engineers and Engineering Managers
  • Ability to navigate complex technical, organizational, and business situations
  • High level of rigor and a strong sense of ownership
  • Excellent organizational, prioritization, and planning skills
  • Strong communication and stakeholder-management abilities
  • Ability to synthesize complex engineering topics and communicate them clearly to both technical and non-technical stakeholders
  • Comfortable making decisions in fast-moving environments with high technical and operational stakes
WHAT YOU WILL FIND AT SCALEWAY
  • Hybrid work:We offer up to 3 days of remote work per week.
  • Offices: Our offices are spacious, dynamic workspaces with bold design, conveniently located near public transport. Most of our offices feature outdoor spaces (terraces) and bike parking facilities.

  • Dining: Our chef provides a healthy meal service at the headquarters, and breakfast is available across all our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches.

  • Well-being commitments: Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life.

  • International environment: With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.

  • Career & Mobility: Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers.

Why join the Scaleway adventure?

A rich and diverse product offering: Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

A cutting-edge technical environment: Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

Commitment to responsible cloud: Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

THE NEXT STEPS …
  • HR discovery call (30 min)

  • Interview with to understand your technical skills and approach to the role (45 min)

  • Technical interview with CTO to validate your expertise (1h)

  • Interview with Management to deepen discussions and assess your fit with the team (45 min)

  • HR interview to tour our offices and meet your future colleagues

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.

All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Engineering Manager AI GPU Cloud
Engineering Manager AI GPU Cloud

Scaleway • Paris

Hybride
EUR 90 000 - 130 000
Hybrid work up to 3 days per week
Modern offices near public transport
Healthy meals at HQ and Swile lunchカード
Engineering Manager AI GPU Cloud
Engineering Manager AI GPU Cloud

Scaleway • Paris

Hybride
EUR 110 000 - 140 000
Hybrid work up to 3 days remote per wk
Office near public transport
Swile meal card
Head of Operations GPU Cloud
Head of Operations GPU Cloud

Scaleway • Paris

Hybride
EUR 90 000 - 130 000
Hybrid work up to 3 days remote per wk
Modern offices near public transport
Healthy meal service at HQ
+3
Pre-Sales Solutions Architect - AI & GPU Infrastructure
Pre-Sales Solutions Architect - AI & GPU Infrastructure

Scaleway • Lille

Hybride
EUR 85 000 - 110 000
Hybrid work up to 3 days per week
International environment with diverse
SRE Engineering Manager – GPU Cloud
SRE Engineering Manager – GPU Cloud

Webhosting • Paris

Hybride
EUR 90 000 - 130 000
Hybrid work
Dining service
Swile card
Head of Operations GPU Cloud
Head of Operations GPU Cloud

Webhosting • Paris

Hybride
EUR 110 000 - 150 000
Hybrid work
Modern offices
Healthy meals
+3
Head of Operations GPU Cloud
Head of Operations GPU Cloud

Scaleway • Paris

Hybride
EUR 90 000 - 130 000
Hybrid work up to 3 days
Hardware Architect
Hardware Architect

Scaleway • Paris

Hybride
EUR 60 000 - 80 000
Healthy meal service
Remote work flexibility
Access to gym and daycare services
+1
Hardware Architect
Hardware Architect

Scaleway • Paris

Hybride
EUR 70 000 - 90 000
Hybrid work model
Healthy meal service
Access to a gym
+1
Technicien Support Elite
Technicien Support Elite

Scaleway • Rennes

Hybride
EUR 60 000 - 90 000
Hybrid work up to 3 days/week
Office perks and meals