Senior HPC/Linux Engineer -

Infoplus Technologies UK Ltd

Stevenage

Hybrid

GBP 65,000 - 90,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Infoplus Technologies UK Ltd is seeking a Senior HPC Engineer to manage, secure and maintain the organisation's scientific computing environment. The role focuses on RHEL‑based HPC infrastructure, Slurm workload management and scientific application support, while collaborating with researchers and technical teams to deliver reliable, secure, high‑performing computing services.

The candidate will administer multiple RHEL environments, deploy Slurm, monitor cluster health and work closely with

Qualifications

  • Minimum of 10 years' enterprise IT experience, including 3–5 years in an HPC or research‑computing role.
  • Extensive hands‑on administration and troubleshooting with RHEL 7, 8 and 9.
  • Proven experience managing HPC clusters and deploying Slurm.
  • Experience supporting scientific or research applications in a Linux HPC environment.
  • Strong troubleshooting across hardware, OS, schedulers and applications.
  • Working knowledge of ServiceNow or equivalent ITSM platform.
  • Ability to translate computational requirements into practical technical solutions.
  • Excellent communication and stakeholder‑management skills.
  • Ability to work onsite for a minimum of three days per week and attend at short notice when needed.

Responsibilities

  • Administer, patch, secure and maintain RHEL 7, 8 and 9 across HPC clusters and high‑end workstations.
  • Deploy, configure and manage Slurm, including queues, partitions and fair‑share policies.
  • Monitor and optimise cluster health, resource use, storage, networking and job throughput.
  • Install and support scientific software, compilers, libraries and MPI environments.
  • Collaborate with scientists to understand requirements, optimise workloads and resolve issues.
  • Troubleshoot hardware, OS, scheduler and application problems with root‑cause analysis.
  • Manage incidents, problems and service requests through ServiceNow.
  • Coordinate with networking, storage, security, DevOps and external vendors to deliver HPC services.

Skills

HPC administration
RHEL administration
Slurm workload manager
Linux troubleshooting
ServiceNow
Stakeholder management
Communication
Onsite collaboration

Tools

Docker
Ansible
MPI libraries (OpenMPI/MPICH)

Job description

Mode of working

Hybrid/office based

Hybrid

If Hybrid, how many days are required in office?

3 days

Number of positions

1

The Role

We are seeking an experienced Senior HPC Engineer to manage, secure and maintain the organisation's scientific computing environment. The role will focus on RHEL-based HPC infrastructure, Slurm workload management and scientific application support, while working closely with research scientists and technical teams to deliver reliable, secure and high-performing computing services.

Your responsibilities
  • Administer, patch, secure and maintain RHEL 7, 8 and 9 environments across HPC clusters and high-end workstations.
  • Deploy, configure and manage Slurm, including scheduling, queues, partitions and fair-share policies.
  • Monitor and optimise cluster health, resource utilisation, storage, networking and job throughput.
  • Install and support scientific software, compilers, libraries and MPI environments.
  • Work directly with scientists to understand computational requirements, optimise workloads and resolve application issues.
  • Troubleshoot complex hardware, operating system, scheduler and application problems, including root-cause analysis.
  • Manage incidents, problems and service requests through ServiceNow.
  • Collaborate with networking, storage, security, DevOps teams and external vendors to deliver reliable HPC services.
Your Profile
Essential skills/knowledge/experience
  • Minimum of 10 years' enterprise IT experience, including at least 3 to 5 years in an HPC or research-computing role.
  • Extensive hands‑on administration and troubleshooting experience with RHEL 7, 8 and 9.
  • Proven experience managing HPC clusters and deploying, configuring and operating Slurm.
  • Experience supporting scientific or research applications in a Linux HPC environment.
  • Strong troubleshooting skills across hardware, operating systems, schedulers and applications.
  • Working knowledge of ServiceNow or an equivalent ITSM platform.
  • Ability to work collaboratively with research scientists and translate computational requirements into practical technical solutions.
  • Excellent communication and stakeholder‑management skills.
  • Ability to work onsite for a minimum of three days per week and attend at short notice when physical‑system support is required.
Desirable skills/knowledge/experience
  • Experience with Docker or other container technologies in an HPC environment.
  • Knowledge of Ansible or similar configuration‑management and automation tools.
  • Experience supporting GPU computing, CUDA and GPU‑accelerated workloads on RHEL.
  • Understanding of MPI libraries, particularly OpenMPI or MPICH, and their integration with Slurm.
  • Familiarity with InfiniBand, high‑speed Ethernet and HPC networking concepts.
  • Experience with web server configuration and SSL certificate management.
  • Red Hat certifications such as RHCSA or RHCE, or equivalent qualifications.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Engineer - Hybrid - Inside IR35
Senior HPC Engineer - Hybrid - Inside IR35

Hamilton Barnes • Stevenage

Hybrid
GBP 9,963,000 - 16,605,000
Senior HPC Engineer
Senior HPC Engineer

Gazelle Global Consulting Limited • Stevenage

On-site
GBP 60,000 - 90,000
Senior HPC/Linux Engineer — Slurm & Research Compute
Senior HPC/Linux Engineer — Slurm & Research Compute

Infoplus Technologies UK Ltd • Stevenage

Hybrid
GBP 65,000 - 90,000
HPC Engineer
HPC Engineer

EUROPEAN SOFTWARE SOLUTIONS LIMITED • Cambridge

Hybrid
GBP 105,000 - 129,000
Senior HPC Engineer: Slurm & RHEL Expert (Stevenage)
Senior HPC Engineer: Slurm & RHEL Expert (Stevenage)

Hamilton Barnes • Stevenage

Hybrid
GBP 9,963,000 - 16,605,000
HPC Support Engineer
HPC Support Engineer

LinuxRecruit • Greater London

On-site
GBP 60,000 - 90,000
Linux HPC Architect
Linux HPC Architect

Innovate Recruitment Ltd • Stevenage

Hybrid
GBP 90,000 - 120,000
Hybrid work model
Personal development programme
International exposure
HPC Infrastructure Site Reliability Engineer
HPC Infrastructure Site Reliability Engineer

Radiant • Greater London

On-site
GBP 90,000 - 140,000
Linux System Engineer
Linux System Engineer

Chapman Tate Associates • United Kingdom

Remote
GBP 38,000 - 45,000
Excellent career development opportunities
Work remotely with a forward-thinking team
Be part of innovative projects driving scientific breakthroughs
Senior HPC Linux Engineer - Onsite, Slurm/PBS Platform Lead
Senior HPC Linux Engineer - Onsite, Slurm/PBS Platform Lead

Whitehall Resources • England

On-site
GBP 60,000 - 90,000