HPC Systems Engineer III - Linux & Automation

Boston University

Boston (MA)

Hybrid

USD 120,000 - 180,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Time Off: PTO and holidays
University retirement plan
Tuition assistance
Wellness information

Job summary

Boston University IS&T seeks a Research Computing Systems Engineer to operate a 1,000+ node HPC cluster, deploy a tiCrypt-based secure environment, and modernize monitoring and automation across the full stack. You will work with SGE, GPFS, Slurm, and virtualization technologies while collaborating with senior engineers and researchers.

The role emphasizes hands-on systems work, end-to-end automation, and a strong voice in infrastructure decisions, with partial remote work and opportunities for

Qualifications

  • 3+ years of relevant work experience with a bachelor's degree, relevant post-secondary education, or a combination of these two.
  • Linux systems administration experience, including OS patching, upgrades, and security maintenance.
  • Experience programming in a Linux environment, preferably Bash and Python.
  • Experience in multiple languages such as Perl or C is a plus.
  • Ability to work with configuration management systems (e.g., Ansible, Puppet, Chef) and revision control systems (e.g., Git).
  • Hands-on experience and understanding of Linux virtualization and containerization technologies (e.g., KVM, OpenStack, OpenNebula, Docker, Kubernetes).
  • Strong communication, problem-solving, and teamwork skills.

Responsibilities

  • Operate, maintain, and optimize the University's production HPC cluster, including a 1,000+ node bare-metal Linux deployment, the Grid Engine (SGE) scheduler, and GPFS (IBM Storage Scale) parallel file system, to keep research computing services reliable and performant.
  • Contribute to the build, deployment, and operations of a new NIST 800-171–compliant secure research computing environment built on tiCrypt, including a Slurm workload manager.
  • Help improve monitoring, observability, and configuration management using CI/CD pipelines, Ansible, Git, and related automation tools.
  • Automate provisioning, deployment, and routine operations across the full stack, from bare metal through services.
  • Identify, diagnose, and resolve complex system, storage, network, and performance issues.
  • Contribute to technical documentation and share your expertise with colleagues and the research community.

Skills

Linux systems administration
Bash
Python
Perl
C
Ansible
Puppet
Chef
Git
Docker
Kubernetes
OpenStack
OpenNebula
communication
teamwork

Education

Bachelor's degree

Tools

Ansible
Puppet
Chef
Git
Docker
Kubernetes
OpenStack
OpenNebula

Job description

Boston University IS&T seeks a Research Computing Systems Engineer to operate a 1,000+ node HPC cluster, deploy a tiCrypt-based secure environment, and modernize monitoring and automation across the full stack. You will work with SGE, GPFS, Slurm, and virtualization technologies while collaborating with senior engineers and researchers.

The role emphasizes hands-on systems work, end-to-end automation, and a strong voice in infrastructure decisions, with partial remote work and opportunities for

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Systems Engineer III – Cloud & Automation (Remote)
HPC Systems Engineer III – Cloud & Automation (Remote)

Boston University • Boston (MA), Northern (KY)

Hybrid
USD 100,000 - 125,000
Time Off
Retirement plan
Tuition assistance
+1
Senior Research HPC Systems Engineer – Cloud & Automation
Senior Research HPC Systems Engineer – Cloud & Automation

Boston University • Boston (MA)

Hybrid
USD 100,000 - 125,000
Time Off & Holidays
Retirement Plan
Tuition Assistance
Research Computing Systems Engineer III, Cloud Technologies IS&T Research Computing Boston, MA
Research Computing Systems Engineer III, Cloud Technologies IS&T Research Computing Boston, MA

Boston University • Boston (MA), Northern (KY)

Hybrid
USD 100,000 - 125,000
Time Off
Retirement plan
Tuition assistance
+1
Research Computing Systems Engineer III, Cloud Technologies IS&T Research Computing
Research Computing Systems Engineer III, Cloud Technologies IS&T Research Computing

Boston University • Boston (MA)

Hybrid
USD 100,000 - 125,000
Time Off & Holidays
Retirement Plan
Tuition Assistance
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux
HPC Systems Engineer — Remote/Hybrid, Slurm/Linux

Strategic Business Systems, Inc (SBS) • Chantilly (VA)

Hybrid
USD 120,000 - 180,000
Flexible work arrangements
HPC Engineer: Slurm, Secure Research Clusters
HPC Engineer: Slurm, Secure Research Clusters

Conditions. Workplace Diversity, LLC. • Boston (MA)

Hybrid
USD 80,000 - 100,000
Medical, dental, and vision health insurance
Generous paid time off
Professional development opportunities
TS/SCI Linux HPC Engineer (Onsite)
TS/SCI Linux HPC Engineer (Onsite)

Jobless • Charlottesville (VA), Northern (KY)

Hybrid
USD 100,000 - 175,000
HPC Systems Engineer - Hybrid/Remote AWS-ready
HPC Systems Engineer - Hybrid/Remote AWS-ready

Strategic Business Systems (SBS) • United States

Hybrid
USD 95,000 - 130,000
HPC Linux Systems Engineer - Automation & SLURM Expert
HPC Linux Systems Engineer - Automation & SLURM Expert

CGG Services (U.S.) Inc. • Houston (TX)

On-site
USD 95,000 - 140,000
HPC Linux Systems Engineer (TS/SCI) – Onsite
HPC Linux Systems Engineer (TS/SCI) – Onsite

Plus3 IT Systems • Charlottesville (VA)

On-site
USD 100,000 - 175,000
Health benefits
401(k) matching
Parental leave