Director of Deployment & Platform Reliability

Nscale

Amer

Presencial

EUR 209.000 - 307.000

Jornada completa

Hace 11 días
Generador de candidaturas

Transforma esta oferta en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Descripción de la vacante

Nscale is seeking a Director of Deployment Engineering to lead the systems engineering function responsible for deploying, validating, and operating the core compute platforms behind our AI infrastructure.

You will lead a high-performing team across systems, compute operations, and deployment validation, defining standards, tooling, and processes to ensure reliability and rapid delivery. Travel up to 50% will be required to support deployments and site readiness.

Formación

  • Bachelors degree in Computer Science, Engineering, or a related technical field.
  • 10+ years of systems engineering, compute operations, infrastructure, or engineering-management experience.
  • Experience in a large cloud provider, hyperscale data center, or similarly complex infrastructure environment.
  • Experience leading systems or compute operations teams in a high-availability production environment.
  • Strong knowledge of Linux/Unix systems administration, OS-level tuning, server architecture, and GPU hardware.
  • Experience with virtualization, containerization, and distributed systems, including Kubernetes and Docker.
  • Experience with infrastructure-as-code and configuration-management tools such as Ansible, Terraform, Puppet, or Chef.
  • Strong scripting, automation, and data center systems-design experience.
  • Familiarity with system architecture, data synchronization, fault tolerance, state management, and distributed-system reliability.
  • Experience with enterprise storage, networking, compute, monitoring, observability, or telemetry platforms.
  • Excellent judgment, organizational skills, and written and verbal communication.
  • The ability to influence technical direction, engineering priorities, and cross-functional decisions.

Responsabilidades

  • Lead and scale the systems engineering team responsible for core platform systems, deployment readiness, and compute operations.
  • Define and deliver multi-quarter systems engineering initiatives that improve deployment velocity, platform reliability, validation quality, and operational performance.
  • Establish systems validation standards for servers, GPUs, networking, storage, and supporting infrastructure before production acceptance.
  • Lead GPU burn-in and validation testing at scale, including thermal, power, stress, and performance qualification.
  • Partner with infrastructure, network engineering, deployment, product, and operations leaders to balance long-term platform architecture with immediate business needs.
  • Turn ambiguous, high-impact technical challenges into clear plans, priorities, milestones, and accountable execution.
  • Drive alignment across interdependent teams working on compute platforms, internal infrastructure, developer tooling, and operational systems.
  • Raise the bar for engineering quality, automation, observability, documentation, and operational excellence.
  • Build scalable approaches for systems monitoring, telemetry, incident learning, and continuous platform improvement.
  • Travel up to 50% to support deployments, site readiness, vendor collaboration, and operational execution.

Conocimientos

Linux/Unix admin
Kubernetes
Docker
Automation scripting
Leadership

Educación

Bachelor's degree in Computer Science/Engineering or related field

Herramientas

Terraform
Ansible
Puppet
Chef

Descripción del empleo

Nscale is seeking a Director of Deployment Engineering to lead the systems engineering function responsible for deploying, validating, and operating the core compute platforms behind our AI infrastructure.

You will lead a high-performing team across systems, compute operations, and deployment validation, defining standards, tooling, and processes to ensure reliability and rapid delivery. Travel up to 50% will be required to support deployments and site readiness.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Director, Deployment Engineering — Systems Engineering
Director, Deployment Engineering — Systems Engineering

Nscale • Amer

Presencial
EUR 209.000 - 307.000
Senior Site Reliability Engineer — CI/CD Platform & Cloud
Senior Site Reliability Engineer — CI/CD Platform & Cloud

N26 Inc. • Barcelona

Presencial
EUR 70.000 - 100.000
Relocation package with visa support
Premium N26 bank account subscriptions
Work from home budget
+1
Agentic Devops Architect: Ai-Driven Sdlc & Platforms
Agentic Devops Architect: Ai-Driven Sdlc & Platforms

Act Digital • Barcelona

Presencial
EUR 60.000 - 90.000
Health benefits
Wellness program
Financial rewards
AI-Driven Kubernetes Platform Architect
AI-Driven Kubernetes Platform Architect

Act Digital • Barcelona

Presencial
EUR 60.000 - 90.000
Health benefits
Wellness program
Financial rewards
Senior SRE: Cloud Platform Reliability & Automation
Senior SRE: Cloud Platform Reliability & Automation

United States Digital Space LLC • España

Presencial
EUR 76.000 - 102.000
Health coverage for you and family
Flexible locations & schedules
Generous vacation days
+2
Platform Engineer — AI/ML Infra & Kubernetes
Platform Engineer — AI/ML Infra & Kubernetes

Alten Spain • Lugo

Presencial
EUR 70.000 - 100.000
Senior AI Infra & Platform SRE — Remote EU
Senior AI Infra & Platform SRE — Remote EU

Front Door Defense • Barcelona

A distancia
EUR 90.000 - 130.000
Remote work flexibility
Competitive compensation
Career growth opportunities
Senior Platform Engineer
Senior Platform Engineer

Intellias • España

Presencial
EUR 70.000 - 100.000
Kubernetes Platform Engineer for AI Ops
Kubernetes Platform Engineer for AI Ops

Alten Delivery Centre Spain • Madrid

Presencial
EUR 60.000 - 90.000
Senior Platform Engineer — AI-First Infra & Scale
Senior Platform Engineer — AI-First Infra & Scale

Haddock • Barcelona

Híbrido
EUR 50.000 - 60.000
Hybrid work
Office in Barcelona Poblenou
AI tooling access
+2