HPC Engineer – IFM

The Chronicle Of Higher Education, Inc.

United Arab Emirates

On-site

AED 120,000 - 180,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

The Chronicle Of Higher Education, Inc. is looking for an HPC Engineer to support large-scale GPU computing infrastructure at the Institute for Foundation Models (IFM) at MBZUAI. This position is ideal for recent graduates eager to work with Linux systems and distributed computing.

As an HPC Engineer, you will maintain GPU clusters, assist researchers with troubleshooting, and develop automation tools. Candidates should have a relevant Bachelor's degree and some familiarity with Linux, Python, and networking concepts.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering or related disciplines required.
  • Linux administration experience preferred.
  • Essential programming skills include Python, Bash, and C/C++.

Responsibilities

  • Support operation and maintenance of GPU computing clusters.
  • Assist in job submission, troubleshooting, and resource utilization.
  • Monitor cluster health and performance.

Skills

Linux administration
Python programming
Bash programming
C/C++ programming
Networking fundamentals
Cloud platforms (Azure, AWS, GCP)
Containers (Docker, Apptainer, Enroot)
Git and software development workflows
AI/ML infrastructure exposure
HPC experience

Education

Bachelor’s degree in relevant disciplines

Job description

Application Open

Full-Time

The Institute for Foundation Models (IFM) at MBZUAI operates someof the world’s largest AI supercomputing environments, supportingfrontier AI research and foundation model development acrossthousands of GPUs.

We are seeking an HPC Engineer to join our growing infrastructureteam. This role is suitable for recent graduates and early-careerengineers who are passionate about Linux systems, large-scalecomputing, distributed systems, and AI infrastructure.

Key Responsibilities
  • Support operation and maintenance of large-scale GPU computing clusters.
  • Assist researchers with job submission, troubleshooting, and resource utilization.
  • Monitor cluster health, performance, and availability.
  • Troubleshoot Linux, hardware, storage, networking, and software issues.
  • Support Slurm administration and user management.
  • Assist with cluster deployment, upgrades, and validation.
  • Develop scripts and automation tools.
  • Maintain technical documentation and operational procedures.
  • Participate in incident response and operational support.
  • Collaborate with researchers, vendors, and internal teams.
Academic Qualifications
  • Bachelor’s degree in Computer Science, Computer Engineering, Electrical Engineering, Software Engineering, Information Technology, Mathematics, Physics, or related disciplines.
Professional Experience Required

Preferred:

  • Linux administration experience.
  • Python, Bash, Go, or C/C++ programming.
  • Networking fundamentals.
  • Cloud platforms (Azure, AWS, GCP).
  • Containers (Docker, Apptainer, Enroot).
  • Git and software development workflows.
  • AI/ML infrastructure exposure.
  • HPC, distributed systems, or research computing experience.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Engineer – IFM
Senior HPC Engineer – IFM

The Chronicle Of Higher Education, Inc. • United Arab Emirates

On-site
AI Infrastructure HPC Engineer for Large-Scale GPU Clusters
AI Infrastructure HPC Engineer for Large-Scale GPU Clusters

The Chronicle Of Higher Education, Inc. • United Arab Emirates

On-site
AED 120,000 - 180,000
Project Manager - High Performance Computing (HPC)
Project Manager - High Performance Computing (HPC)

Institute of Foundation Models • Abu Dhabi

On-site
High Performance Computing Software Engineer - Supercomputing
High Performance Computing Software Engineer - Supercomputing

Institute of Foundation Models • Abu Dhabi

On-site
AED 350,000 - 700,000
Senior Fullstack Engineer
Senior Fullstack Engineer

Institute of Foundation Models • Abu Dhabi

On-site
AED 180,000 - 250,000
Machine Learning Engineer
Machine Learning Engineer

Institute of Foundation Models • Abu Dhabi

On-site
AED 257,000 - 331,000
AI Infrastructure Engineer
AI Infrastructure Engineer

APPIT Software Inc. • Abu Dhabi

On-site
AED 180,000 - 240,000
GPU-HPC AI Infrastructure Engineer
GPU-HPC AI Infrastructure Engineer

APPIT Software Inc. • Abu Dhabi

On-site
AED 180,000 - 240,000
Senior HPC Engineer: Build & Optimize AI Clusters
Senior HPC Engineer: Build & Optimize AI Clusters

Core42 • Abu Dhabi

On-site
AED 420,000 - 650,000
Competitive Salary
Yearly Bonus
Discount Cards Esaad and Fazaa
+2
AI Performance Engineer
AI Performance Engineer

MBR Partners • Dubai

On-site
AED 300,000 - 500,000
Relocation tickets (incl. family)
Visa support
Insurance