HPC Systems Engineer: Scale & Optimize High-Performance Compute

Socket.dev

Houston (TX)

On-site

USD 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401(k) plan with company match
Vacation days
Sick days
Parental leave

Job summary

n² Group's HPC Services team in Houston, TX, is seeking a High Performance Computing Systems Engineer to design, deploy, operate, and optimize large-scale HPC platforms.

You'll work with computational scientists and partners to ensure performance and reliability, manage Linux-based environments, and implement security controls while evaluating new technologies.

This full-time on-site role offers growth in a mature organization with a strong focus on collaboration and innovation.

Qualifications

  • Bachelor's degree or equivalent in CS/CE/IS or related field.
  • 5+ years Linux production experience (RHEL/CentOS).
  • 5+ years deploying/ supporting production HPC.
  • Experience with Lustre/GPFS, InfiniBand/Omni-Path, Slurm/PBS Pro.
  • Strong communication with scientists/ stakeholders.

Responsibilities

  • Configure and manage HPC clusters, storage, and networking.
  • Maintain day-to-day operation and improvements of HPC environments.
  • Diagnose hardware, OS, networking, storage, and app issues across the HPC stack.
  • Implement security controls and maintain platform integrity.
  • Collaborate with data scientists, researchers, and domain specialists to support workflows.
  • Monitor system health, investigate bottlenecks, and optimize performance.
  • Plan installations, upgrades, patches, and platform improvements.
  • Evaluate new hardware/software to improve capability and performance.
  • Work with vendors to resolve complex infrastructure issues.

Skills

Linux administration
Bash scripting
Python
C/C++
HPC optimization
Communication
Collaboration

Education

Bachelor's degree in CS/CE/IS or related

Tools

Lustre
GPFS
InfiniBand
Omni-Path
Slurm
PBS Pro
Conda
Spack
RPM

Job description

n² Group's HPC Services team in Houston, TX, is seeking a High Performance Computing Systems Engineer to design, deploy, operate, and optimize large-scale HPC platforms.

You'll work with computational scientists and partners to ensure performance and reliability, manage Linux-based environments, and implement security controls while evaluating new technologies.

This full-time on-site role offers growth in a mature organization with a strong focus on collaboration and innovation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Systems Engineer — Scale & Performance
Senior HPC Systems Engineer — Scale & Performance

Autonomai Recruitment • Chicago (IL)

On-site
USD 110,000 - 170,000
Elite HPC Systems Architect & Platform Engineer
Elite HPC Systems Architect & Platform Engineer

CGG Services SAS • Houston (TX)

On-site
USD 90,000 - 130,000
Senior HPC Systems Engineer — Large-Scale Linux & Storage
Senior HPC Systems Engineer — Large-Scale Linux & Storage

ExxonMobil • Spring (TX)

On-site
USD 100,000 - 130,000
Pension Plan
Savings Plan
Comprehensive medical, dental, and vision plans
+2
Senior HPC Specialist (3202-1) Denver, CO
Senior HPC Specialist (3202-1) Denver, CO

ESR Healthcare • Denver (CO)

On-site
USD 90,000 - 130,000
HPC Systems Engineer — Cloud & On-Prem Compute Lead
HPC Systems Engineer — Cloud & On-Prem Compute Lead

Staffing Technologies • United States

On-site
USD 150,000 - 220,000
HPC System Administrator
HPC System Administrator

Cybotic System • Savannah (GA)

On-site
USD 90,000 - 150,000
HPC Performance & Systems Engineer
HPC Performance & Systems Engineer

Jobtailor • Irving (TX)

On-site
USD 85,000 - 115,000
Remote HPC Engineer: Compute-at-Scale & Cloud Ops
Remote HPC Engineer: Compute-at-Scale & Cloud Ops

RCH Solutions • Wayne (PA)

Remote
USD 80,000 - 120,000
Competitive salary and bonus package
Comprehensive health and wellness benefits
Company-sponsored 401(k) plan
+1
Senior HPC & GPU Cluster Architect — Scale & Automate
Senior HPC & GPU Cluster Architect — Scale & Automate

San Francisco Compute Company • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Generous equity grant
Competitive salary
Visa sponsorship
+6
HPC Systems Engineer: Large-Scale Compute & Performance
HPC Systems Engineer: Large-Scale Compute & Performance

ASML • San Jose (CA)

On-site
USD 148,000 - 222,000