GPU Systems Engineer III — AI Clusters & Linux

RPMGlobal

Bethesda (MD)

On-site

USD 120,000 - 260,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

RPMGlobal is seeking a senior, security-cleared engineer to design, deploy, and optimize GPU clusters for enterprise AI mission systems. You will work with Linux-based environments, collaborate with AI/ML teams, and drive performance improvements across hardware and software layers.

Responsibilities include integration with NVIDIA data center platforms, driver optimization, tooling automation, documentation, and ensuring compliance with federal security standards.

Qualifications

  • Active TS/SCI clearance with ability to obtain CI Polygraph.
  • Bachelor's degree with six years experience or 9 years without a degree.
  • Experience managing NVIDIA GPU data center platforms (DGX, HGX, H200, H100, L4s).
  • Strong Linux expertise (RHEL, Ubuntu, Oracle, Rocky).
  • IAT Level II/III certification requirements.
  • U.S. citizenship required for government contracts.

Responsibilities

  • Design, configure, and maintain GPU clusters.
  • Collaborate to optimize architectures for performance and efficiency.
  • Work with AI/ML engineers to integrate GPUs with Linux systems.
  • Optimize GPU drivers for reliability and performance.
  • Analyze GPU performance and remove bottlenecks across hardware/software.
  • Build and maintain debugging tools and performance analysis software for Linux.
  • Use Bash, Python, Ansible, Puppet, Salt for tooling/automation.
  • Document architectures and Linux best practices.
  • Support ATO activities and ensure DoD security compliance.

Skills

Linux administration
Python scripting
Bash scripting
Ansible
Puppet
Salt
Kubernetes
Prometheus
Grafana
Slurm
LSF

Education

Bachelor's Degree
High School Diploma / GED
Associates Degree
Master's Degree
PhD

Tools

NVIDIA DGX
HGX
H200
H100
L4s
DoD 8570.11 IAT certifications

Job description

RPMGlobal is seeking a senior, security-cleared engineer to design, deploy, and optimize GPU clusters for enterprise AI mission systems. You will work with Linux-based environments, collaborate with AI/ML teams, and drive performance improvements across hardware and software layers.

Responsibilities include integration with NVIDIA data center platforms, driver optimization, tooling automation, documentation, and ensuring compliance with federal security standards.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior GPU Systems Architect for AI Clusters
Senior GPU Systems Architect for AI Clusters

RPMGlobal • Bethesda (MD)

On-site
USD 140,000 - 190,000
GPU Systems Engineer 3
GPU Systems Engineer 3

RPMGlobal • Bethesda (MD)

On-site
USD 120,000 - 260,000
GPU Systems Engineer 4
GPU Systems Engineer 4

RPMGlobal • Bethesda (MD)

On-site
USD 140,000 - 190,000
Senior HPC Systems Engineer: GPU Clusters & AI Infra
Senior HPC Systems Engineer: GPU Clusters & AI Infra

Nebius • United States

Remote
USD 180,000 - 240,000
Competitive pay
Career growth
Flexibility and ownership
+3
Senior HPC-AI Systems Architect (Equity)
Senior HPC-AI Systems Architect (Equity)

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 176,000 - 334,000
Senior AI Infrastructure Architect – GPU Clusters
Senior AI Infrastructure Architect – GPU Clusters

NVIDIA • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
AI Systems Engineer: HPC & GPU Clusters
AI Systems Engineer: HPC & GPU Clusters

Advanced Micro Devices, Inc. • San Jose (CA)

On-site
USD 180,000 - 260,000
AI Kernel / Cluster Engineer
AI Kernel / Cluster Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD 150,000 - 210,000
HPC AI Systems Architect (On-Prem GPU Cluster)
HPC AI Systems Architect (On-Prem GPU Cluster)

MRE Consulting • Houston (TX)

On-site
USD 95,000 - 140,000
Senior AI Performance & Efficiency Engineer - Equity Eligible
Senior AI Performance & Efficiency Engineer - Equity Eligible

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Equity
Competitive benefits