Senior Linux Platform Engineer

CLOUDENGINE DIGITAL SDN. BHD.

Kuala Lumpur

On-site

MYR 180,000 - 240,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

CLOUDENGINE DIGITAL SDN. BHD. in Kuala Lumpur is seeking a hands-on Senior Linux Platform Engineer with strong Linux systems, networking, storage, automation, and production troubleshooting experience.

The engineer will own the reliability and operation of Linux-based platforms and must be able to diagnose complex issues, restore services safely, automate repetitive work, and improve platform standards.

Qualifications

  • Years of hands-on Linux system administration or platform engineering experience in production environments.
  • Strong Ubuntu/Debian or RHEL-based Linux administration and troubleshooting skills.
  • Strong Linux boot, kernel, systemd, filesystem, and OS recovery knowledge.
  • Strong Linux networking fundamentals including TCP/IP, routing, DNS, VLAN, bonding, firewall, and Layer 2 / Layer 3 troubleshooting.
  • Strong storage knowledge including LVM, mdadm RAID, NFS, filesystem management, and disk troubleshooting.
  • Hands-on Bash and Python scripting experience for operational automation.
  • Hands-on experience with Ansible, Git, and configuration management practices.
  • Experience operating and troubleshooting Docker or other container technologies.
  • Able to diagnose production incidents independently using logs, metrics, system state, and structured root cause analysis.

Responsibilities

  • Operate, maintain, and troubleshoot production Linux server and platform environments.
  • Administer Ubuntu, Debian, RHEL, Rocky Linux or similar distributions, including systemd, packages, kernel settings, users, permissions, and system services.
  • Troubleshoot Linux OS, kernel, process, CPU, memory, filesystem, package, and service issues using system logs and diagnostic tools.
  • Troubleshoot the Linux boot and recovery process including BIOS / UEFI, GRUB, Kernel, initramfs / initrd, root filesystem, and systemd.
  • Manage and troubleshoot storage including partitions, LVM, mdadm RAID, filesystems, NFS, mounts, capacity, and I/O performance.
  • Troubleshoot Linux networking including TCP/IP, routing, VLAN, bonding / LACP, DNS, DHCP, firewall, MTU, and connectivity issues.
  • Build and operate core infrastructure services such as DNS, DHCP, NTP, NFS, SSH, LDAP / SSSD, and reverse proxy services where required.
  • Develop Bash and Python scripts to automate operational checks, maintenance, troubleshooting, and repetitive administration tasks.
  • Use Ansible and Git to standardize configuration, automate deployments, track changes, and maintain repeatable operating procedures.
  • Support Docker, containerd, and container-based applications, with working knowledge of Kubernetes platform troubleshooting.
  • Implement and maintain monitoring, logging, alerting, health checks, and capacity visibility for Linux platforms.
  • Own production incidents from diagnosis through recovery, validation, root cause analysis, corrective action, and closure.
  • Create SOPs, Runbooks, troubleshooting guides, and technical reports; mentor junior engineers and review high-risk technical changes.

Skills

Linux admin
Ubuntu/Debian/RHEL
Kernel & OS recovery
Networking fundamentals
Storage management
Bash scripting
Python scripting
Ansible
Git
Docker
Kubernetes
Incident diagnosis

Tools

Ansible
Git
Docker
Kubernetes
Terraform

Job description

We are looking for a hands-on Senior Linux Platform Engineer with strong Linux systems, networking, storage, automation, and production troubleshooting experience.

The engineer will own the reliability and operation of Linux-based platforms and must be able to diagnose complex issues, restore services safely, automate repetitive work, and improve platform standards.

Key Responsibilities

Operate, maintain, and troubleshoot production Linux server and platform environments.

Administer Ubuntu, Debian, RHEL, Rocky Linux or similar distributions, including systemd, packages, kernel settings, users, permissions, and system services.

Troubleshoot Linux OS, kernel, process, CPU, memory, filesystem, package, and service issues using system logs and diagnostic tools.

Troubleshoot the Linux boot and recovery process including BIOS / UEFI, GRUB, Kernel, initramfs / initrd, root filesystem, and systemd.

Manage and troubleshoot storage including partitions, LVM, mdadm RAID, filesystems, NFS, mounts, capacity, and I/O performance.

Troubleshoot Linux networking including TCP/IP, routing, VLAN, bonding / LACP, DNS, DHCP, firewall, MTU, and connectivity issues.

Build and operate core infrastructure services such as DNS, DHCP, NTP, NFS, SSH, LDAP / SSSD, and reverse proxy services where required.

Develop Bash and Python scripts to automate operational checks, maintenance, troubleshooting, and repetitive administration tasks.

Use Ansible and Git to standardize configuration, automate deployments, track changes, and maintain repeatable operating procedures.

Support Docker, containerd, and container-based applications, with working knowledge of Kubernetes platform troubleshooting.

Implement and maintain monitoring, logging, alerting, health checks, and capacity visibility for Linux platforms.

Own production incidents from diagnosis through recovery, validation, root cause analysis, corrective action, and closure.

Create SOPs, Runbooks, troubleshooting guides, and technical reports; mentor junior engineers and review high-risk technical changes.

Required Skills

Years of hands-on Linux system administration or platform engineering experience in production environments.

Strong Ubuntu / Debian or RHEL-based Linux administration and troubleshooting skills.

Strong Linux boot, kernel, systemd, filesystem, and OS recovery knowledge.

Strong Linux networking fundamentals including TCP/IP, routing, DNS, VLAN, bonding, firewall, and Layer 2 / Layer 3 troubleshooting.

Strong storage knowledge including LVM, mdadm RAID, NFS, filesystem management, and disk troubleshooting.

Hands-on Bash and Python scripting experience for operational automation.

Hands-on experience with Ansible, Git, and configuration management practices.

Experience operating and troubleshooting Docker or other container technologies.

Able to diagnose production incidents independently using logs, metrics, system state, and structured root cause analysis.

Preferred Background (Not Mandatory)

Experience operating enterprise Linux platforms, data centers, cloud infrastructure, or large-scale server environments.

Experience with Kubernetes administration and troubleshooting in production environments.

Experience with Prometheus, Grafana, Netdata, OpenTelemetry or similar observability tools.

Experience with Terraform, CI/CD pipelines, or infrastructure automation practices.

Experience designing or operating high availability, disaster recovery, backup, and platform resilience solutions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Linux Platform Engineer
Senior Linux Platform Engineer

CloudEngine Digital Co., Ltd • Kuala Lumpur

On-site
MYR 180,000 - 240,000
Senior Linux Platform Engineer | Production Reliability
Senior Linux Platform Engineer | Production Reliability

CLOUDENGINE DIGITAL SDN. BHD. • Kuala Lumpur

On-site
MYR 180,000 - 240,000
Senior Engineer, Platform Infrastructure (Containers & Virtualization)
Senior Engineer, Platform Infrastructure (Containers & Virtualization)

Singtel • Kuala Lumpur

On-site
Confidential
Senior Engineer, Platform Infrastructure (Containers & Virtualization)
Senior Engineer, Platform Infrastructure (Containers & Virtualization)

Singtel Group • Kuala Lumpur

On-site
MYR 120,000 - 200,000
Senior Devops Engineer
Senior Devops Engineer

MHA Consultancy Services Sdn Bhd • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

HFG Insurance Recruitment • Putrajaya, Cyberjaya

On-site
MYR 180,000 - 240,000
Lead Compute Infrastructure Services
Lead Compute Infrastructure Services

IBroad Solutions • Kuala Lumpur

On-site
MYR 240,000 - 360,000
Senior Linux Engineer - Hybrid Infra & Automation
Senior Linux Engineer - Hybrid Infra & Automation

Commerz Global Service Solutions Sdn. Bhd. • Petaling Jaya

Hybrid
MYR 120,000 - 180,000
Hybrid work arrangement
Medical benefits
Insurance coverage
+4
Senior Platform Reliability Engineer
Senior Platform Reliability Engineer

HFG (Hong Kong) Limited • Cyberjaya

On-site
MYR 180,000 - 300,000
Senior / Lead Platform Reliability Engineer (PRE)
Senior / Lead Platform Reliability Engineer (PRE)

HFG Insurance Recruitment • Cyberjaya

On-site
MYR 180,000 - 300,000