HPC Systems Engineer - Scheduling & Performance

Referment

New York (NY)

On-site

USD 90,000 - 130,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Referment is seeking an experienced systems engineer to design and improve workload scheduling, fleet management, and clustered file-system operations for large-scale compute environments. You will also tune kernel and network performance while building tools to analyze infrastructure behavior for trading and research workloads.

The role requires strong programming skills in Python, Go, or Rust, and a Computer Science degree or up to five years of applicable experience.

Qualifications

  • Candidates must have strong programming ability in Python, Go, Rust, or a similar language.
  • A computer science degree or up to five years of comparable practical experience is required.

Responsibilities

  • Design workload scheduling, fleet management, and clustered file-system operations for large-scale compute environments.
  • Tune kernel and network performance while building tools to analyze infrastructure behavior for trading and research workloads.

Skills

Python
Go
Rust
Linux
Networking
HPC
Storage
GPU Infrastructure
Data-centre Hardware
Workload Scheduling
Fleet Management
Clustered File Systems
Kernel Tuning
Performance Tuning
Distributed Computing
Software Engineering

Education

Computer Science degree or equivalent experience

Job description

Referment is seeking an experienced systems engineer to design and improve workload scheduling, fleet management, and clustered file-system operations for large-scale compute environments. You will also tune kernel and network performance while building tools to analyze infrastructure behavior for trading and research workloads.

The role requires strong programming skills in Python, Go, or Rust, and a Computer Science degree or up to five years of applicable experience.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Developer: HPC Scheduling & Clustered File Systems
Systems Developer: HPC Scheduling & Clustered File Systems

Referment • New York (NY)

On-site
USD 120,000 - 160,000
Quant Systems: Systems Developer (New York) (4295F7B)
Quant Systems: Systems Developer (New York) (4295F7B)

Referment • New York (NY)

On-site
USD 90,000 - 130,000
Quant Systems: Systems Developer (New York) (2D66E00)
Quant Systems: Systems Developer (New York) (2D66E00)

Referment • New York (NY)

On-site
USD 120,000 - 160,000
Quant Systems: Systems Developer (New York) (B419E59)
Quant Systems: Systems Developer (New York) (B419E59)

Referment • New York (NY)

On-site
USD 120,000 - 180,000
Senior HPC Scheduler Engineer (LSF/Slurm) - Hybrid & Equity
Senior HPC Scheduler Engineer (LSF/Slurm) - Hybrid & Equity

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 152,000 - 288,000
Equity
Benefits package
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux

Strategic Business Systems, Inc (SBS) • Chantilly (VA)

Hybrid
USD 120,000 - 180,000
Flexible work arrangements
Senior HPC Scheduler & Platform Reliability Engineer
Senior HPC Scheduler & Platform Reliability Engineer

NVIDIA • Austin (TX)

On-site
USD 152,000 - 242,000
Senior HPC Scheduler & Platform Reliability Engineer
Senior HPC Scheduler & Platform Reliability Engineer

NVIDIA • Westford (MA)

On-site
USD 152,000 - 242,000
HPC Systems Engineer
HPC Systems Engineer

Radix Trading Experienced Job Board • New York (NY), Chicago (IL)

On-site
USD 120,000 - 150,000
Senior HPC Scheduler & Platform Reliability Engineer
Senior HPC Scheduler & Platform Reliability Engineer

NVIDIA • Durham (NC)

On-site
USD 152,000 - 242,000