Turn this role into an interview — a resume and cover letter built around what this employer wants.
NVIDIA is seeking a deeply technical Senior HPC Cluster Administrator to lead the design, deployment, and reliability of our large-scale GPU compute clusters, spanning DGX/HGX to Grace Blackwell systems. You will own the full lifecycle of GPU clusters, design storage and networking solutions, and drive automation with Ansible, Terraform, and CI/CD pipelines.
You will collaborate with ML engineers and software teams to optimize workloads and reliability.
NVIDIA is seeking a deeply technical Senior HPC Cluster Administrator to lead the design, deployment, and reliability of our large-scale GPU compute clusters, spanning DGX/HGX to Grace Blackwell systems. You will own the full lifecycle of GPU clusters, design storage and networking solutions, and drive automation with Ansible, Terraform, and CI/CD pipelines.
You will collaborate with ML engineers and software teams to optimize workloads and reliability.