GPU Cluster Architect: AI Infrastructure Design Lead
Nebius
Amsterdam
Hybrid
EUR 65,000 - 80,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Competitive salary and benefits
Opportunities for professional growth
Hybrid working arrangements
Dynamic and collaborative work environment
Job summary
A dynamic technology company is seeking a GPU Cluster Architect in Amsterdam to design next-generation AI infrastructure. This hands-on role requires extensive experience in designing GPU clusters and optimizing performance for modern AI workloads. The ideal candidate will collaborate with experts to ensure robust and scalable architecture across multiple data centers, focusing on high-performance applications.
Qualifications
5+ years of experience designing clusters.
Deep understanding of modern GPU architecture (NVIDIA, AMD, etc.).
Experience with HPC interconnects (InfiniBand & RoCE).
Solid background in systems architecture, networking, and hardware reliability.
Experience in scripting for automation and telemetry pipelines.
Responsibilities
Architect scalable GPU cluster topologies including compute nodes and interconnect.
Analyze AI/ML workloads to inform design tradeoffs across latency and bandwidth.
Validate low-latency, high-throughput interconnects at POD and DC scale.
Work with storage teams to optimize performance.
Partner with teams to operationalize and scale architecture.
Skills
Cluster Design
Performance Modeling
Network Architecture
Storage Integration
Reliability & Monitoring
Collaboration
Tools
Python
Go
InfiniBand
RoCE
Job description
A dynamic technology company is seeking a GPU Cluster Architect in Amsterdam to design next-generation AI infrastructure. This hands-on role requires extensive experience in designing GPU clusters and optimizing performance for modern AI workloads. The ideal candidate will collaborate with experts to ensure robust and scalable architecture across multiple data centers, focusing on high-performance applications.