HPC Engineer

Monash University

Brixworth

On-site

GBP 55,000 - 75,000

Full time

34 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Monash University in the UK is seeking an experienced HPC Engineer to own the reliability, availability and performance of HPC platforms supporting simulation and engineering workloads. You will improve compute, storage, networking and scheduling services, automate tasks with Bash, Python and Ansible, and provide technical escalation for incidents and capacity issues.

Collaboration with suppliers and internal teams is essential to raise service maturity.

Qualifications

  • Expertise in HPC architecture, parallel workloads and resource scheduling.
  • Strong Linux administration skills and secure configuration practices.
  • Experience delivering upgrades, patches and change with minimal downtime.
  • Ability to analyse capacity, performance and data movement issues.

Responsibilities

  • Own reliability, availability and performance of HPC platforms.
  • Improve compute, storage, networking and scheduling services for scalable workloads.
  • Provide technical escalation for HPC incidents and capacity issues.
  • Automate operational tasks using Bash, Python, Ansible or equivalent.
  • Translate user requirements into service improvements.

Skills

HPC architecture
Linux administration
Schedulers (Slurm, PBS, LSF)
Performance tuning
Capacity planning
Scripting (Bash, Python)
Automation (Ansible)
Networking basics
Troubleshooting complex systems

Education

Linux/HPC certifications or equivalent experience

Tools

Slurm
PBS
LSF
Ansible
Bash

Job description

Job no: 498149
Work type: Permanent
Location: Brixworth
Categories:

Purpose of the role is to…
  • Own the reliability, availability and performance of HPC platforms supporting simulation, analysis and engineering workloads
  • Improve compute, storage, networking and scheduling services to enable efficient, scalable workload delivery
  • Provide technical escalation for HPC incidents, capacity issues, performance bottlenecks and complex user problems
  • Administering Linux-based HPC clusters, including compute nodes, schedulers and shared platform services
  • Troubleshooting issues across hardware, OS, network, storage, applications and user workflows
  • Managing capacity, performance and availability for engineering and simulation workloads
  • Automating operational tasks using Bash, Python, PowerShell, Ansible or equivalent tools
  • Translating technical user requirements into practical service improvements
Have experience of…
  • Supporting Linux-based HPC, scientific computing, simulation or high-throughput compute environments
  • Diagnosing workload, queue, licence, performance, data movement and application issues
  • Operating at a senior technical level in an enterprise or engineering-led environment
  • Delivering maintenance, upgrades, patching and change activity with minimal service impact
  • Working with suppliers and internal teams to resolve platform issues and improve service maturity
Demonstrate knowledge of…
  • HPC architecture, parallel workloads, scheduling, queues and resource allocation
  • Linux administration, scripting, patching and secure configuration
  • Schedulers such as Slurm, PBS, LSF or equivalent
  • Scale-out storage, file systems, backup, archive and data lifecycle management
  • Networking, interconnects, latency, bandwidth and data locality considerations
  • Monitoring, performance tuning, benchmarking and capacity forecasting
  • Security, vulnerability management, access control and compliance for shared platforms
  • Desirable: motorsport, automotive, CFD, simulation or data science experience
Hold these qualifications…
  • Relevant technical certifications, or equivalent experience, in Linux, HPC, storage, networking, automation or ITIL
  • Analytical, curious and comfortable solving complex technical problems
  • Proactive in improving resilience, reducing risk and removing operational friction
  • Structured, communicative and effective across hands-on delivery and change control
  • Collaborative, customer-focused and willing to share knowledge
Success in this role will be if you… (deliverables)
  • Maintain stable, secure and performant HPC services for critical engineering workloads
  • Improve compute and storage utilisation through effective monitoring, queue management and capacity planning
  • Resolve incidents quickly and reduce repeat issues through automation, documentation and service improvement
  • Deliver upgrades, maintenance and project work safely with clear communication and change control
  • Improve simulation throughput, data availability and user productivity
  • Define and guide strategic direction on HPC related topics
  • The role combines operational support and project delivery, including planned maintenance, capacity improvement, lifecycle management and occasional out-of-hours activity

Advertised: 18 Aug 2026 GMT Daylight Time
Applications close: 18 Sep 2026 GMT Daylight Time

In August we will open our application window for our next intake of Early Careers vacancies! We have over 50 opportunities available, across all disciplines, for exceptional students to join our Eight-time World Championship winning team at our state–of-the‑art Technology Centre in Northamptonshire.

Want to find out more about what our F1 Trackside Team do here at HPP, view the video below from Barney our Head of Product Validation to get an in-depth insight into the world of the F1 Engineering.

HPP team members and their partners gathered for an evening of appreciation and connection at Rushton Hall for our Long Service Celebration.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Operations Trainee - ERS Build
Operations Trainee - ERS Build

Monash University • Brixworth

On-site
GBP 18,000 - 24,000
HPC Senior Hardware Engineer
HPC Senior Hardware Engineer

CGG Services (UK) Limited • Bolney

On-site
GBP 90,000 - 120,000
Competitive salary
Bonus scheme
Relocation sponsorship
+1
HPC Senior Technology Consultant
HPC Senior Technology Consultant

Hewlett Packard Enterprise • Winnersh

On-site
GBP 60,000 - 90,000
Health & wellbeing benefits
Career development programs
Inclusive work environment
HPC Engineer
HPC Engineer

Amentum • West of England

On-site
GBP 25,000 - 36,000
HPC Engineer
HPC Engineer

Mercedes AMG High Performance Powertrains • Brixworth

On-site
GBP 65,000 - 90,000
Senior HPC Engineer
Senior HPC Engineer

Amentum • West of England

On-site
GBP 55,000 - 75,000
Free medical cover
Digital GP service
Enhanced parental leave pay
+1
Lead HPC Engineer
Lead HPC Engineer

LinuxRecruit • Greater London

On-site
GBP 75,000 - 82,000
HPC Engineer (Linux) - Sussex, Onsite
HPC Engineer (Linux) - Sussex, Onsite

Arden Resourcing • Pease Pottage

On-site
GBP 75,000 - 85,000
Annual bonus scheme
Enhanced pension contribution
Private medical and dental options
+2
HPC Engineer
HPC Engineer

LinuxRecruit • Greater London

On-site
GBP 50,000 - 70,000
HPC Platform Engineer Linux - Trading
HPC Platform Engineer Linux - Trading

Client Server • Greater London

Hybrid
GBP 100,000 - 130,000
Salary up to £130k
Pension
25 days holiday
+2