Shift Lead – DC Systems Operations Engineer

Neuron Solutions Sdn. Bhd.

Johor Bahru

On-site

MYR 120,000 - 190,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Neuron Solutions Sdn. Bhd. in Malaysia seeks an experienced Shift Lead for Data Centre Systems Operations to lead on-shift GPU cluster and infrastructure activities.

You will manage incidents, ITSM governance and vendor coordination while ensuring service levels and stakeholder communications. The role requires a strong background in data centre operations, GPU hardware familiarity, and experience with Jira Service Management.

Qualifications

  • 5+ years of experience in IT infrastructure or data centre operations.
  • Bachelor's degree in Computer Science, IT, Electrical Engineering or a related field, or equivalent practical experience.

Responsibilities

  • Lead daily data centre and infrastructure operations during the assigned shift.
  • Oversee the stable operation of GPU clusters, networks, storage and supporting infrastructure.
  • Act as the primary escalation point for L1 engineers.
  • Coordinate P1/P2 and major incidents and ensure timely resolution.
  • Ensure incidents, tickets and escalations meet required SLA and ITSM standards.
  • Review shift handovers and ensure operational continuity.
  • Coordinate with hardware, network and data centre vendors.
  • Track vendor cases, SLA performance and hardware replacement activities.
  • Review operational dashboards and ensure data accuracy.
  • Ensure activities are properly documented and supported by evidence.
  • Prepare shift reports covering incidents, SLA risks, vendor cases and data centre activities.
  • Coach and guide L1 engineers on troubleshooting and operational processes.
  • Develop and maintain SOPs and operational runbooks.

Skills

ITSM
Incident management
Shift leadership
Vendor coordination
Customer communication
GPU/AI HPC familiarity
Monitoring tools

Education

Bachelor's degree in Computer Science/IT/Electrical Engineering or related field

Tools

Jira Service Management
Prometheus
Grafana
Zabbix

Job description

Shift Lead – DC Systems Operations Engineer

We are looking for an experienced Shift Lead – Data Centre Systems Operations to lead on-shift operations supporting large-scale GPU clusters and data centre infrastructure.

This role will be responsible for operational execution, incident coordination, ITSM governance, vendor coordination and customer-facing communication during the assigned shift.

Key Responsibilities

Lead daily data centre and infrastructure operations during the assigned shift.

Oversee the stable operation of GPU clusters, networks, storage and supporting infrastructure.

Act as the primary escalation point for L1 engineers.

Coordinate P1/P2 and major incidents and ensure timely resolution.

Ensure incidents, tickets and escalations meet required SLA and ITSM standards.

Review shift handovers and ensure operational continuity.

Coordinate with hardware, network and data centre vendors.

Track vendor cases, SLA performance and hardware replacement activities.

Review operational dashboards and ensure data accuracy.

Ensure activities are properly documented and supported by evidence.

Prepare shift reports covering incidents, SLA risks, vendor cases and data centre activities.

Coach and guide L1 engineers on troubleshooting and operational processes.

Develop and maintain SOPs and operational runbooks.

Requirements

Bachelor's degree in Computer Science, IT, Electrical Engineering or a related field, or equivalent practical experience.

5+ years of experience in IT infrastructure or data centre operations.

Previous experience as a Shift Lead, Senior NOC Engineer, Operations Lead or similar role.

Strong ITSM and incident management experience.

Experience with Jira Service Management or similar ticketing platforms.

Experience managing P1/P2 incidents and SLA-driven environments.

Familiarity with NVIDIA GPU hardware and AI/HPC environments is an advantage.

Experience with monitoring tools such as Prometheus, Grafana or Zabbix.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Centre Operations Lead – GPU Clusters
Senior Data Centre Operations Lead – GPU Clusters

Neuron Solutions Sdn. Bhd. • Johor Bahru

On-site
MYR 120,000 - 190,000
Data Centre Shift Leader - Critical Facilities
Data Centre Shift Leader - Critical Facilities

Agensi Pekerjaan RecruitFirst Sdn Bhd • Cyberjaya

On-site
MYR 120,000 - 180,000
Cloud Operations & Support Engineer (L1 / L2)
Cloud Operations & Support Engineer (L1 / L2)

Straitdeer Pte. Ltd. • Cyberjaya

On-site
MYR 70,000 - 100,000
Data Centre Operations Lead
Data Centre Operations Lead

C&W SERVICES (S) PTE. LTD. • Johor Bahru

On-site
MYR 120,000 - 180,000
Shift Lead Engineer
Shift Lead Engineer

YTL DC SOUTH SDN. BHD. • Kulai

On-site
MYR 60,000 - 120,000
Senior Data Centre Operations Engineer
Senior Data Centre Operations Engineer

Oxydata Software Sdn Bhd • Malaysia

On-site
MYR 120,000 - 180,000
Data Centre Operations Manager / Engineer
Data Centre Operations Manager / Engineer

Agensi Pekerjaan Genie Hunt Talent • Petaling Jaya

On-site
MYR 120,000 - 180,000
Shift Lead
Shift Lead

DayOne • Johor

On-site
MYR 89,000 - 134,000
System Engineer – Infrastructure (AI & HPC Systems)
System Engineer – Infrastructure (AI & HPC Systems)

Neuron Solutions Sdn. Bhd. • Johor Bahru

On-site
MYR 90,000 - 150,000
Monetary compensation
Senior Infrastructure Engineer
Senior Infrastructure Engineer

MTAI Sdn. Bhd. • Kuala Lumpur

On-site
MYR 180,000 - 300,000