Junior HPC Systems Engineer: Learn to Run Large Clusters

Parallel Works

Chicago (IL)

Hybrid

USD 70,000 - 100,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical coverage
Vision coverage
Dental coverage
401(k) with company match
Short term disability
Generous paid vacation
Generous sick time

Job summary

Parallel Works is hiring a Junior HPC Systems Engineer to learn supercomputing operations on production systems. The role starts with monitoring, node health, and account and allocation work, and moves into cluster builds and escalations.

The systems are hybrid: on-premises clusters the customer owns, accredited Government cloud regions, and commercial GPU providers, often in the same day. We are looking for strong Linux fundamentals and interest in how large systems behave.

Qualifications

  • 2+ years of hands-on Linux administration (RHEL/ Rocky/ Alma, Debian/Ubuntu).
  • Understanding HPC fundamentals: batch scheduling, shared filesystems, MPI job launch.
  • Bash and Python scripting for operational work.
  • Practical networking: DNS, routing, firewalls, SSH keys, bastion access.
  • Git and ticket-driven workflow; clear written communication.
  • Interest in on-premises estate: bare metal, out-of-band consoles, site networking.
  • US citizenship and eligibility for Secret clearance; active clearance helpful.

Responsibilities

  • Monitor cluster health, node state, queue behavior, and alerting; take initial action on node failures and file system alerts.
  • Manage accounts and Slurm allocations; keep users and groups consistent across venues.
  • Oversee node lifecycle: health checks, draining, returning nodes, escalation paths.
  • Extend Ansible playbooks and operational scripts; automate repetitive steps.
  • Patch and harden systems; document baselines; accompany security review.
  • Maintain runbooks and participate in on-call rotation after training.

Job description

Parallel Works is hiring a Junior HPC Systems Engineer to learn supercomputing operations on production systems. The role starts with monitoring, node health, and account and allocation work, and moves into cluster builds and escalations.

The systems are hybrid: on-premises clusters the customer owns, accredited Government cloud regions, and commercial GPU providers, often in the same day. We are looking for strong Linux fundamentals and interest in how large systems behave.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Junior HPC Applications Engineer — Launch Research Clusters
Junior HPC Applications Engineer — Launch Research Clusters

Parallel Works, Inc. • Chicago (IL)

On-site
USD 65,000 - 90,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Junior HPC Systems Engineer
Junior HPC Systems Engineer

Parallel Works • Chicago (IL)

Hybrid
USD 70,000 - 100,000
Medical coverage
Vision coverage
Dental coverage
+4
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters
Senior HPC Systems Engineer — Secure Hybrid GPU Clusters

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Junior HPC Applications Engineer
Junior HPC Applications Engineer

Parallel Works, Inc. • Chicago (IL)

On-site
USD 65,000 - 90,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Senior HPC Systems Engineer
Senior HPC Systems Engineer

Parallel Works • Chicago (IL)

Hybrid
USD 140,000 - 190,000
Medical, vision, dental coverage
401(k) with company match
Short term disability
+1
Senior HPC Systems Engineer: Linux Clusters, AI & GPU
Senior HPC Systems Engineer: Linux Clusters, AI & GPU

United States Digital Space LLC • Starbase (TX)

On-site
USD 120,000 - 190,000
Senior HPC Systems Engineer: Scale AI Clusters
Senior HPC Systems Engineer: Scale AI Clusters

SpaceX • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Stock options
Health, vision and dental coverage
401(k) retirement plan
+2
HPC Systems Architect for AI & GPU Clusters
HPC Systems Architect for AI & GPU Clusters

United States Digital Space LLC • Hawthorne (CA)

On-site
USD 165,000 - 230,000
Senior HPC Systems Engineer: Linux Clusters & AI Readiness
Senior HPC Systems Engineer: Linux Clusters & AI Readiness

Future Ventures • Town of Texas (WI)

On-site
USD 140,000 - 210,000
Junior HPC & GPU System Engineer
Junior HPC & GPU System Engineer

Amax 1 • Colorado

On-site
USD 80,000 - 95,000
Medical, Dental, Vision Insurance
Flexible spending account
Health savings account
+4