System Engineer – Infrastructure (GPU - HPC Systems)

Neuron Solutions Sdn. Bhd.

Johor Bahru

On-site

MYR 120,000 - 200,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Neuron Solutions Sdn. Bhd. in Johor Bahru, Malaysia, seeks a hands-on System Engineer – Infrastructure to deploy, optimise and maintain large-scale GPU and CPU infrastructure powering AI workloads and HPC environments.

You will deploy and troubleshoot GPU/CPU servers, tune BIOS and OS configurations, monitor health, and work with DevOps, networking and storage teams to ensure end-to-end performance. Strong Linux experience and HPC knowledge are essential.

Qualifications

  • Bachelor's degree in Computer Science, Electrical Engineering or a related technical field.
  • 3+ years hands-on experience managing server infrastructure in HPC, AI, GPU clusters or data centre environments.
  • Strong experience with Linux systems and performance optimisation.
  • Hands-on experience with GPU/CPU servers and bare-metal infrastructure.
  • Knowledge of IPMI, PXE, Redfish or BMC provisioning.
  • Familiarity with monitoring tools such as Prometheus and Grafana.
  • Basic knowledge of Kubernetes or containerised environments.

Responsibilities

  • Deploy, configure and maintain GPU and CPU servers across large-scale compute clusters.
  • Optimise BIOS, firmware and OS configurations for AI and HPC workloads.
  • Perform server health monitoring, diagnostics, firmware upgrades and lifecycle management.
  • Coordinate with hardware vendors and system integrators for cluster deployments.
  • Manage provisioning, OS deployment, GPU/NIC driver installation and hardening.
  • Conduct validation, burn-in testing and workload benchmarking.
  • Monitor system health, diagnose failures and perform root-cause analysis.
  • Collaborate with networking, storage and DevOps teams for end-to-end performance.
  • Develop scripts/automation for deployment, monitoring and remediation.
  • Maintain documentation of configurations, rack layouts, cabling, and procedures.
  • Provide L2/L3 support and participate in on-call activities.

Skills

Linux systems
GPU/CPU servers
HPC/AI infrastructure
On-call experience

Education

Bachelor's degree in Computer Science or Electrical Engineering

Tools

IPMI
PXE
Redfish/BMC
Prometheus
Grafana
Kubernetes (basic)

Job description

Jora Malaysia will close on 16th September 2026. Thank you for being with us, we are cheering you on as you continue your career journey.

Neuron Solutions Sdn. Bhd. – Johor Bahru, Johor

We are looking for a hands-on System Engineer – Infrastructure to support and optimise large-scale GPU and CPU infrastructure powering AI workloads, large model training and high-performance computing environments.

You will be responsible for deploying, maintaining and troubleshooting GPU/CPU servers while ensuring infrastructure reliability, performance and operational readiness.

Key Responsibilities
  • Deploy, configure and maintain GPU and CPU servers across large-scale compute clusters.
  • Optimise BIOS, firmware and operating system configurations for AI and HPC workloads.
  • Perform server health monitoring, hardware diagnostics, firmware upgrades and lifecycle management.
  • Support cluster deployments across multiple racks and coordinate with hardware vendors and system integrators.
  • Manage server provisioning, OS deployment, GPU/NIC driver installation and system hardening.
  • Conduct system validation, burn-in testing and workload benchmarking.
  • Monitor system health, investigate failures and perform root cause analysis.
  • Work closely with networking, storage and DevOps teams to ensure end-to-end infrastructure performance.
  • Develop scripts and automation to improve infrastructure deployment, monitoring and remediation.
  • Maintain technical documentation including system configurations, rack layouts, cabling and operational procedures.
  • Provide L2/L3 support and participate in on-call activities.
Requirements
  • Bachelor's degree in Computer Science, Electrical Engineering or a related technical field.
  • At least 3 years of hands-on experience managing server infrastructure in HPC, AI, GPU cluster or data centre environments.
  • Strong experience with Linux systems, system tuning and performance optimisation.
  • Hands-on experience with GPU/CPU servers and bare-metal infrastructure.
  • Knowledge of server provisioning technologies such as IPMI, PXE, Redfish or BMC.
  • Familiarity with monitoring tools such as Prometheus and Grafana.
  • Basic knowledge of Kubernetes or containerised environments.
  • Experience with server hardware, GPU platforms and infrastructure troubleshooting.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

System Engineer – Infrastructure (AI & HPC Systems)
System Engineer – Infrastructure (AI & HPC Systems)

Neuron Solutions Sdn. Bhd. • Johor Bahru

On-site
MYR 90,000 - 150,000
Monetary compensation
Data Centre GPU Infrastructure Engineer (Based KL)
Data Centre GPU Infrastructure Engineer (Based KL)

CloudEngine Digital Co., Ltd • Kuala Lumpur

On-site
MYR 67,000 - 112,000
AI Infra Systems Engineer – HPC & GPU Clusters
AI Infra Systems Engineer – HPC & GPU Clusters

Neuron Solutions Sdn. Bhd. • Johor Bahru

On-site
MYR 120,000 - 200,000
Senior Data Centre Operations Engineer
Senior Data Centre Operations Engineer

Oxydata Software Sdn Bhd • Malaysia

On-site
MYR 120,000 - 180,000
Infrastructure Systems Engineer for AI & HPC Clusters
Infrastructure Systems Engineer for AI & HPC Clusters

Neuron Solutions Sdn. Bhd. • Johor Bahru

On-site
MYR 90,000 - 150,000
Monetary compensation
GPU Hardware Field Service Engineer
GPU Hardware Field Service Engineer

Oxydata Software Sdn Bhd • Kulai

On-site
MYR 60,000 - 120,000
GPU Hardware Field Service Engineer
GPU Hardware Field Service Engineer

Oxydata Software Sdn Bhd • Malaysia

On-site
MYR 90,000 - 150,000
GPU Data Center Infrastructure Engineer — Kuala Lumpur
GPU Data Center Infrastructure Engineer — Kuala Lumpur

CloudEngine Digital Co., Ltd • Kuala Lumpur

On-site
MYR 67,000 - 112,000
Technical Manager - GPU Cloud & AI Infrastructure
Technical Manager - GPU Cloud & AI Infrastructure

Risewave Consulting, Inc. • Kuala Lumpur

On-site
MYR 180,000 - 280,000
Senior AI Network & Security Engineer (Johor Bahru)
Senior AI Network & Security Engineer (Johor Bahru)

Techstreet • Johor Bahru

On-site
MYR 180,000 - 300,000