Senior HPC Systems Engineer - AI & Research Clusters

Johns Hopkins University

Baltimore (MD)

On-site

USD 86,000 - 150,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Johns Hopkins University IT@JH Research Computing seeks a Sr. HPC Systems Engineer to design, build, and maintain advanced high‑performance computing environments.

The role focuses on reliable operation, configuration, and optimization of HPC and AI systems, including multi‑node CPU/GPU clusters, InfiniBand networks, and large‑scale storage. You will work with faculty, researchers, and students on ticketed support and project deployments, while implementing secure, reproducible platforms and

Qualifications

  • Bachelor’s degree.
  • Six years of related experience.
  • Additional education may substitute for required experience beyond a high school diploma/graduation equivalent, to the extent permitted by the JHU equivalency formula.

Responsibilities

  • Support and administer production systems used by researchers.
  • Provide technical leadership/project management for system configuration, implementation, management, and user support for both new and existing systems.
  • Research and recommend new functionality for HPC management and administration tools by exploring system-wide impacts.
  • Expertise with architecting, operating, and debugging large scale HPC network and storage infrastructure, including MPI, NCCL, RDMA, Infiniband, and parallel file systems
  • Work with scientific support specialists and assigns tasks and provides oversight to HPC engineering team for researchers using diverse applications
  • Analyze results of server monitoring and implement changes to improve performance, processing, and utilization
  • Propose, maintain, and enforce policies, practices and security procedures
  • Provide break/fix support, setup/installation support, escalation support, and solutions support
  • Collaborate with stakeholders on all aspects of projects
  • Other duties as assigned.

Skills

Six years related experience
HPC systems administration
Linux administration
Automation scripting

Education

Bachelor’s degree

Tools

Slurm
Ansible
Puppet
Salt
GPFS
Lustre
BeeGFS
WekaFS
Infiniband
Docker
Kubernetes

Job description

Johns Hopkins University IT@JH Research Computing seeks a Sr. HPC Systems Engineer to design, build, and maintain advanced high‑performance computing environments.

The role focuses on reliable operation, configuration, and optimization of HPC and AI systems, including multi‑node CPU/GPU clusters, InfiniBand networks, and large‑scale storage. You will work with faculty, researchers, and students on ticketed support and project deployments, while implementing secure, reproducible platforms and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead HPC Systems Engineer for AI and Research Clusters
Lead HPC Systems Engineer for AI and Research Clusters

Johns Hopkins University • Baltimore (MD)

On-site
USD 85,500 - 149,800
Senior HPC & AI Software Engineer
Senior HPC & AI Software Engineer

The Johns Hopkins University • Baltimore (MD)

On-site
USD 80,000 - 120,000
Sr. HPC Systems Engineer (IT@JH Research Computing)
Sr. HPC Systems Engineer (IT@JH Research Computing)

Johns Hopkins University • Baltimore (MD)

On-site
USD 85,500 - 149,800
Sr. HPC Systems Engineer (IT@JH Research Computing) - #Staff
Sr. HPC Systems Engineer (IT@JH Research Computing) - #Staff

Johns Hopkins University • Baltimore (MD)

On-site
USD 86,000 - 150,000
Senior Linux Systems & HPC Automation Engineer
Senior Linux Systems & HPC Automation Engineer

The Johns Hopkins University Applied Physics Laboratory • Laurel (MD)

On-site
USD 100,000 - 245,000
Senior Systems Administrator — Hybrid IT, Automation & Security
Senior Systems Administrator — Hybrid IT, Automation & Security

Johns Hopkins University • Baltimore (MD)

Hybrid
USD 64,000 - 113,000
HPC Scientific Software Engineer (IT@JH Research Computing)
HPC Scientific Software Engineer (IT@JH Research Computing)

The Johns Hopkins University • Baltimore (MD)

On-site
USD 80,000 - 120,000
Senior HPC-AI Cluster Architect (Equity)
Senior HPC-AI Cluster Architect (Equity)

NVIDIA • Santa Clara (CA)

On-site
USD 176,000 - 334,000
Equity
Benefits
Hybrid Sr. Systems Administrator — Secure IT Infrastructure
Hybrid Sr. Systems Administrator — Secure IT Infrastructure

Johns Hopkins University • United States

Hybrid
USD 64,000 - 113,000
Senior Research Systems Engineer - HPC, AI & Secure Data Workflows
Senior Research Systems Engineer - HPC, AI & Secure Data Workflows

University of California • Georgia

On-site
USD 95,200 - 121,400
Health insurance
Tuition Assistance Program
Vacation time
+2