Senior HPC Systems Engineer — Clusters & Automation

The University Of Chicago

Austin (TX)

Hybrid

USD 100,000 - 125,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

The University of Chicago is seeking a highly qualified Senior HPC System Administrator to join the RCC team responsible for HPC cluster infrastructure and facility operations. This hybrid role requires three days on site and involves procurement and management of HPC hardware and software.

The role includes designing, deploying, and maintaining Linux systems, job scheduling tools, and storage subsystems, with a focus on reliability, security, and performance for research computing.

Qualifications

  • Education: college/university degree in related field.
  • 5-7 years of work experience in a related job discipline.
  • Master's preferred for advanced roles.
  • Experience with HPC environments and large clusters is highly desirable.

Responsibilities

  • Install, configure, and maintain large clusters/servers and software.
  • Handle day-to-day system operations, monitoring, and storage performance.
  • Manage network switches, parallel file systems, and HPC software stacks.
  • Configure scheduling/queuing systems and diagnose operational issues quickly.
  • Coordinate with vendors to resolve hardware/software problems.
  • Assist users with access requests and help desk tickets.
  • Develop system automation scripts and tools for security maintenance and patching.
  • Build and deploy open-source software and vendor software as needed.
  • Provide reliable backups/restores for all managed systems.
  • Document procedures for routine and complex tasks.

Skills

Linux admin
HPC systems
MPI/OpenMP
Job schedulers
Python scripting
Storage networking

Education

Master's degree in CS
Bachelor's degree in CS

Tools

SLURM/Torque

Job description

The University of Chicago is seeking a highly qualified Senior HPC System Administrator to join the RCC team responsible for HPC cluster infrastructure and facility operations. This hybrid role requires three days on site and involves procurement and management of HPC hardware and software.

The role includes designing, deploying, and maintaining Linux systems, job scheduling tools, and storage subsystems, with a focus on reliability, security, and performance for research computing.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Systems Administrator - Hybrid
Senior HPC Systems Administrator - Hybrid

uchicago • Chicago (IL)

Hybrid
USD 120,000 - 180,000
HPC workstation access
Sr. HPC System Administrator
Sr. HPC System Administrator

uchicago • Chicago (IL)

Hybrid
USD 120,000 - 180,000
HPC workstation access
Sr. HPC System Administrator
Sr. HPC System Administrator

The University Of Chicago • Austin (TX)

Hybrid
USD 100,000 - 125,000
Systems Administrator III - HPC & Networking
Systems Administrator III - HPC & Networking

University of Chicago • Chicago (IL)

On-site
USD 100,000 - 110,000
Senior HPC Systems Architect – Research Computing
Senior HPC Systems Architect – Research Computing

The University of North Carolina • Charlotte (NC)

On-site
USD 100,000 - 140,000
Senior HPC Systems Engineer - Research Computing Lead
Senior HPC Systems Engineer - Research Computing Lead

University of North Carolina at Charlotte • Charlotte (NC)

On-site
USD 120,000 - 160,000
Senior HPC Systems Administrator - Hybrid Research Compute
Senior HPC Systems Administrator - Hybrid Research Compute

University of Pennsylvania • Philadelphia

Hybrid
USD 84,000 - 105,000
Health, Life, and Flexible Spending Accounts
Tuition assistance
Generous retirement plans
+9
HPC Cluster Engineer - Linux, Automation & Lab Ops
HPC Cluster Engineer - Linux, Automation & Lab Ops

Procom Consultants Group • Champaign (IL)

On-site
USD 65,000 - 85,000
Senior HPC Systems Engineer: Linux Clusters, AI & GPU
Senior HPC Systems Engineer: Linux Clusters, AI & GPU

United States Digital Space LLC • Starbase (TX)

On-site
USD 120,000 - 190,000
Systems Administration - HPC Cluster
Systems Administration - HPC Cluster

Metasys Technologies • Newton (MA)

On-site