Staff HPC Systems Engineer — Slurm Platform Lead (GPU Cloud)

Greenhouse Software, Inc.

New York (NY)

On-site

USD 225,000 - 275,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Nscale is hiring a Staff HPC Systems Software Engineer to define the technical direction and evolution of a core HPC platform domain. You will shape multi-team efforts to build, automate, and operate Slurm-based capabilities within a cloud-native platform.

This high-impact role requires deep hands-on software engineering, strong systems judgment, and the ability to drive architecture across teams while ensuring robust, maintainable HPC services in GPU-backed infrastructure.

Qualifications

  • Extensive experience designing and building production software and automation for HPC systems, especially Slurm-based environments.
  • Strong track record of writing maintainable, testable, and resilient software in Go, Python, or similar languages.
  • Proven ability to define technical direction across a domain spanning multiple teams or services.
  • Strong understanding of Slurm internals, scheduler behaviour, cluster lifecycle concerns, and operational trade-offs.

Responsibilities

  • Own and evolve the technical direction for a defined HPC systems domain, such as Slurm platform architecture, scheduler integrations, cluster lifecycle, workload environments, or service automation.
  • Define how proven Slurm implementations should be packaged, automated, and exposed as a service.
  • Establish shared patterns and standards for automation, service lifecycle management, observability, reliability, and supportability across the HPC platform.
  • Lead technically critical initiatives spanning 2–4 teams or a defined HPC platform area.

Skills

HPC systems design
Slurm-based environments
Go/Python development
Platform orchestration
Kubernetes integration

Tools

Slurm
Kubernetes

Job description

Nscale is hiring a Staff HPC Systems Software Engineer to define the technical direction and evolution of a core HPC platform domain. You will shape multi-team efforts to build, automate, and operate Slurm-based capabilities within a cloud-native platform.

This high-impact role requires deep hands-on software engineering, strong systems judgment, and the ability to drive architecture across teams while ensuring robust, maintainable HPC services in GPU-backed infrastructure.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff HPC Platform Engineer — Slurm & Cloud
Staff HPC Platform Engineer — Slurm & Cloud

Nscale • United States

Remote
USD 225,000 - 275,000
Medical benefits
Dental benefits
Vision benefits
+3
Senior HPC Systems Engineer — Slurm, GPU, Cloud-Native
Senior HPC Systems Engineer — Slurm, GPU, Cloud-Native

Nscale • New York (NY)

On-site
USD 180,000 - 260,000
Bonus
Equity
Medical Insurance
+5
Staff HPC Systems Engineer - Slurm & Cloud Platform Lead
Staff HPC Systems Engineer - Slurm & Cloud Platform Lead

Nscale Ltd. • United States

On-site
USD 225,000 - 275,000
Highly competitive compensation package
Flexible workplace
Dynamic progression plan
Staff HPC Systems Software Engineer
Staff HPC Systems Software Engineer

Nscale • New York (NY)

On-site
USD 180,000 - 260,000
Bonus
Equity
Medical Insurance
+5
Staff HPC Systems Software Engineer
Staff HPC Systems Software Engineer

Greenhouse Software, Inc. • New York (NY)

On-site
USD 225,000 - 275,000
Senior GPU Infra Lead: Slurm, Kubernetes & Platform
Senior GPU Infra Lead: Slurm, Kubernetes & Platform

Jobgether SRL • United States

Remote
USD 170,000 - 250,000
Head of HPC Systems Engineering & Platform Strategy
Head of HPC Systems Engineering & Platform Strategy

Nscale • San Francisco (CA)

On-site
USD 200,000 - 300,000
Senior Slurm & HPC Systems Engineer (GPU/Kubernetes)
Senior Slurm & HPC Systems Engineer (GPU/Kubernetes)

Bitdeer (NASDAQ: BTDR) • San Jose (CA)

On-site
USD 180,000 - 250,000
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Staff HPC Systems Software Engineer
Staff HPC Systems Software Engineer

Nscale • United States

Remote
USD 225,000 - 275,000
Medical benefits
Dental benefits
Vision benefits
+3