Senior HPC Systems Engineer

Autonomai Recruitment

Chicago (IL)

On-site

USD 110,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Autonomai Recruitment in Chicago seeks a Senior HPC Operations Engineer to join one of the world's most advanced HPC environments supporting cutting-edge quantitative research. You will operate and scale a production HPC platform used by researchers running demanding workloads.

You will operate large-scale Linux infrastructure, high-speed interconnects, parallel storage, and automation, troubleshooting issues and improving reliability with researchers and engineers.

Qualifications

  • 5+ years of Linux systems administration experience.
  • Experience supporting HPC environments.
  • Exposure to technologies such as Slurm, Lustre, GPFS, BeeGFS or similar.
  • Strong troubleshooting and root cause analysis skills.

Responsibilities

  • Operate and support large-scale Linux HPC environments.
  • Support compute clusters, storage platforms and high-speed interconnects.
  • Work with Slurm, Lustre/GPFS, RDMA/InfiniBand and FUSE.
  • Troubleshoot complex production issues and perform root cause analysis.
  • Monitor infrastructure performance and improve reliability.
  • Partner with researchers and engineering teams to support critical workloads.

Skills

Linux systems administration
HPC environments
Slurm/Lustre/GPFS
Troubleshooting

Job description

Senior HPC Operations Engineer | Chicago

Join one of the world's most advanced high-performance computing environments supporting cutting-edge quantitative research.

We're looking for an experienced HPC Systems Engineer to help operate and scale a large production HPC platform used by researchers running some of the industry's most demanding workloads.

You’ll work on large-scale Linux infrastructure, high-speed networking, parallel storage, automation and performance-critical systems while solving complex operational challenges alongside a highly technical engineering team.

What you’ll be doing:
  • Operating and supporting large-scale Linux HPC environments
  • Supporting compute clusters, storage platforms and high-speed interconnects
  • Working with technologies such as Slurm, Lustre/GPFS, RDMA/InfiniBand and FUSE
  • Troubleshooting complex production issues and performing root cause analysis
  • Monitoring infrastructure performance and improving reliability
  • Partnering with researchers and engineering teams to support critical workloads
What we’re looking for
  • 5+ years of Linux systems administration or Linux infrastructure experience
  • Experience supporting High Performance Computing (HPC) environments
  • Exposure to technologies such as Slurm, Lustre, GPFS, BeeGFS or similar
  • Strong troubleshooting and root cause analysis skills
Why apply?
  • Work on one of the most advanced HPC environments in the world
  • Solve technically challenging infrastructure problems every day
  • Work with cutting-edge compute, networking and storage technologies
  • High ownership with excellent engineering culture
  • Industry-leading compensation and benefits
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Linux Systems Administrator / HPC
Senior Linux Systems Administrator / HPC

Kelly Services • Chicago (IL)

On-site
USD 100,000 - 120,000
Medical, Dental, and Vision Insurance
Vacation Time
Stock Options
+1
Senior HPC Systems Engineer — Scale & Performance
Senior HPC Systems Engineer — Scale & Performance

Autonomai Recruitment • Chicago (IL)

On-site
USD 110,000 - 170,000
Senior HPC & Infrastructure Engineer
Senior HPC & Infrastructure Engineer

Soni • Cherry Hill Township (NJ)

On-site
USD 135,000 - 155,000
HPC System Administrator
HPC System Administrator

Cybotic System • Savannah (GA)

On-site
USD 90,000 - 150,000
HPC Operations Engineer
HPC Operations Engineer

Career Techniques • New York (NY)

Hybrid
USD 175,000 - 225,000
Senior HPC Specialist (3202-1) Denver, CO
Senior HPC Specialist (3202-1) Denver, CO

ESR Healthcare • Denver (CO)

On-site
USD 90,000 - 130,000
HPC Data Center Infrastructure Planning Lead
HPC Data Center Infrastructure Planning Lead

Autonomai Recruitment • Chicago (IL)

On-site
USD 140,000 - 210,000
HPC Storage Engineer | Experienced Hire
HPC Storage Engineer | Experienced Hire

SIG Susquehanna • Bala Cynwyd (PA)

On-site
USD 100,000 - 130,000
HPC Systems Engineer
HPC Systems Engineer

Radix Trading Experienced Job Board • New York (NY), Chicago (IL)

On-site
USD 120,000 - 150,000
HPC Linux Software Engineer
HPC Linux Software Engineer

ClearanceJobs • Colorado Springs (CO)

On-site
USD 150,000 - 210,000