Senior HPC Cluster Engineer: AI Workloads & OpenShift

Abile Group, LLC

Springfield (VA)

On-site

USD 140,000 - 190,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Abile Group is seeking an HPC Infrastructure & Cluster Engineer on a 10-year contract to support a Intelligence Community customer across multiple networks and locations. You will manage Linux clusters, optimize hardware and networks, and orchestrate AI workloads with RunAI and SLURM in OpenShift/Kubernetes environments.

The role requires a Bachelor's degree and 5+ years' Linux experience, with DoD 8570 IAT Level II certification or equivalent, TS/SCI with CI poly eligibility.

Qualifications

  • Bachelor's degree or equivalent experience in a related discipline.
  • At least 5+ years in Linux systems administration and infra management for HPC environments.
  • DoD 8570 IAT Level II certification or equivalent (Security+ CE, CND, SSCP, GSEC, GICSP, CySA+, CCNA).

Responsibilities

  • Manage day-to-day operations of customer compute clusters, including Linux admin and system upgrades.
  • Configure and optimize workload management and AI orchestration platforms (RunAI, SLURM).
  • Tune performance across hardware, OS, and networking to maximize throughput.
  • Administer storage and high-speed networks, including InfiniBand GPU-to-GPU topology.
  • Provision environments and containers using OpenShift/Kubernetes; automate maintenance with scripts.
  • Ensure security, compliance, and accreditation of all infrastructure components.

Skills

Linux systems administration
RunAI/SLURM workload management
OpenShift / Kubernetes
Automation scripting (Bash, Python)
Troubleshooting hardware/network/OS

Education

Bachelor's Degree in related discipline

Tools

InfiniBand networking

Job description

Abile Group is seeking an HPC Infrastructure & Cluster Engineer on a 10-year contract to support a Intelligence Community customer across multiple networks and locations. You will manage Linux clusters, optimize hardware and networks, and orchestrate AI workloads with RunAI and SLURM in OpenShift/Kubernetes environments.

The role requires a Bachelor's degree and 5+ years' Linux experience, with DoD 8570 IAT Level II certification or equivalent, TS/SCI with CI poly eligibility.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead HPC Cluster Engineer for AI/ML & OpenShift
Lead HPC Cluster Engineer for AI/ML & OpenShift

Abile Group, Inc • Springfield (VA)

On-site
USD 130,000 - 180,000
HPC Cluster Engineer: AI Workloads & Secure Infra Ops
HPC Cluster Engineer: AI Workloads & Secure Infra Ops

General Dynamics Information Technology • Springfield (VA)

On-site
USD 150,000 - 190,000
401K match
Health benefits
Internal mobility
+3
HPC Cluster Engineer for AI Workloads | TS/SCI
HPC Cluster Engineer for AI Workloads | TS/SCI

Socket.dev • Springfield (VA)

On-site
USD 148,000 - 179,000
Health/Dental/Vision
401(k)
Paid Time Off
+2
HPC Infrastructure & AI Compute Cluster Engineer
HPC Infrastructure & AI Compute Cluster Engineer

INflow • Springfield (VA)

On-site
USD 140,000 - 185,000
HPC Cluster Engineer (TS/SCI) – Secure Compute
HPC Cluster Engineer (TS/SCI) – Secure Compute

D2 Consulting • Springfield (VA)

On-site
USD 170,000 - 180,000
Health/Dental/Vision
401(k) match
Accrued PTO
+3
HPC Infrastructure & Cluster Engineer
HPC Infrastructure & Cluster Engineer

Abile Group, Inc • Springfield (VA)

On-site
USD 130,000 - 180,000
HPC Cluster Engineer: Linux, InfiniBand & AI Workloads
HPC Cluster Engineer: Linux, InfiniBand & AI Workloads

Inflowfed • Springfield (VA)

On-site
USD 120,000 - 150,000
Senior HPC-AI Cluster Architect (Equity)
Senior HPC-AI Cluster Architect (Equity)

NVIDIA • Santa Clara (CA)

On-site
USD 176,000 - 334,000
Equity
Benefits
HPC Cluster Engineer: Linux, InfiniBand & OpenShift
HPC Cluster Engineer: Linux, InfiniBand & OpenShift

INflow Federal • Springfield (VA)

On-site
USD 140,000 - 185,000
Security-Cleared HPC Cluster Engineer
Security-Cleared HPC Cluster Engineer

D2 Consulting • Springfield (VA)

On-site
USD 170,000 - 180,000
Health/Dental/Vision
401(k) match
Accrued PTO
+3