Senior Slurm HPC & Scheduling Engineer

Bitdeer Technologies Group

Austin (TX)

On-site

USD 150,000 - 190,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Bitdeer is seeking a Staff Slurm Cluster & HPC Scheduling Engineer to own Slurm as a first-class scheduling layer across our GPU fleet. You will be the technical owner of cluster architecture, multi-tenant policy, and reliability on bare-metal and VM-based nodes, driving adoption of the Slinky stack for shared GPU pools.

The role is hands-on, customer-facing during onboarding and escalations, and sets engineering standards for the platform team.

Qualifications

  • 8+ years in HPC, cloud infra.
  • 4+ years operating production Slurm clusters at 100 GPU-node scale with real users and SLAs.
  • Strong GPU and fabric fundamentals (NVIDIA, InfiniBand, DCGM).
  • Experience with Kubernetes and Slurm-on-Kubernetes stacks.

Responsibilities

  • Design, deploy, and operate production Slurm clusters on bare metal and VMs.
  • Lead topology-aware scheduling for GPU fabrics and multi-tenant policy.
  • Advance Slinky on Kubernetes and evaluate slurm-bridge for co-scheduling.
  • Ensure cluster health with diagnostics and failure handling.
  • Automate with Terraform/Ansible and reusable images.
  • Provide technical leadership and customer engagement.

Skills

HPC engineering
Slurm expertise
Kubernetes
Python
Bash
Go
Multi-tenant security

Tools

Terraform
Ansible
Redfish/IPMI

Job description

Bitdeer is seeking a Staff Slurm Cluster & HPC Scheduling Engineer to own Slurm as a first-class scheduling layer across our GPU fleet. You will be the technical owner of cluster architecture, multi-tenant policy, and reliability on bare-metal and VM-based nodes, driving adoption of the Slinky stack for shared GPU pools.

The role is hands-on, customer-facing during onboarding and escalations, and sets engineering standards for the platform team.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Slurm & HPC Systems Engineer (GPU/Kubernetes)
Senior Slurm & HPC Systems Engineer (GPU/Kubernetes)

Bitdeer (NASDAQ: BTDR) • San Jose (CA)

On-site
USD 180,000 - 250,000
Senior Slurm HPC & Kubernetes Scheduler Engineer
Senior Slurm HPC & Kubernetes Scheduler Engineer

Bitdeer Technologies Group • San Jose (CA)

On-site
USD 190,000 - 270,000
Senior Slurm & HPC Cluster Engineer (GPU/AI Infra)
Senior Slurm & HPC Cluster Engineer (GPU/AI Infra)

Bitdeer (NASDAQ: BTDR) • Austin (TX)

On-site
USD 150,000 - 190,000
Staff Slurm Cluster & HPC Engineer
Staff Slurm Cluster & HPC Engineer

Bitdeer Technologies Group • Austin (TX)

On-site
USD 150,000 - 190,000
Staff Slurm Cluster & HPC Engineer
Staff Slurm Cluster & HPC Engineer

Bitdeer (NASDAQ: BTDR) • Austin (TX)

On-site
USD 150,000 - 190,000
Staff Slurm Cluster & HPC Engineer
Staff Slurm Cluster & HPC Engineer

Bitdeer (NASDAQ: BTDR) • San Jose (CA)

On-site
USD 180,000 - 250,000
Staff Slurm Cluster & HPC Engineer
Staff Slurm Cluster & HPC Engineer

Bitdeer Technologies Group • San Jose (CA)

On-site
USD 190,000 - 270,000
Senior HPC Systems Engineer — Slurm, GPU, Cloud-Native
Senior HPC Systems Engineer — Slurm, GPU, Cloud-Native

Nscale • New York (NY)

On-site
USD 180,000 - 260,000
Bonus
Equity
Medical Insurance
+5
Senior HPC Scheduler Engineer (LSF/Slurm) - Hybrid & Equity
Senior HPC Scheduler Engineer (LSF/Slurm) - Hybrid & Equity

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 152,000 - 288,000
Equity
Benefits package
Senior HPC Scheduler (LSF/Slurm) Engineer
Senior HPC Scheduler (LSF/Slurm) Engineer

NVIDIA AI • Durham (NC)

On-site
USD 152,000 - 288,000
Equity
Comprehensive benefits
Career growth opportunities