Jora Malaysia will close on 9th September 2026. Thank you for being with us, we are cheering you on as you continue your career journey.
We are looking for a hands-on Senior Linux Platform Engineer with strong Linux systems, networking, storage, automation, and production troubleshooting experience.
The engineer will own the reliability and operation of Linux-based platforms and must be able to diagnose complex issues, restore services safely, automate repetitive work, and improve platform standards.
Key Responsibilities
- Operate, maintain, and troubleshoot production Linux server and platform environments.
- Administer Ubuntu, Debian, RHEL, Rocky Linux or similar distributions, including systemd, packages, kernel settings, users, permissions, and system services.
- Troubleshoot Linux OS, kernel, process, CPU, memory, filesystem, package, and service issues using system logs and diagnostic tools.
- Troubleshoot the Linux boot and recovery process including BIOS / UEFI, GRUB, Kernel, initramfs / initrd, root filesystem, and systemd.
- Manage and troubleshoot storage including partitions, LVM, mdadm RAID, filesystems, NFS, mounts, capacity, and I/O performance.
- Troubleshoot Linux networking including TCP/IP, routing, VLAN, bonding / LACP, DNS, DHCP, firewall, MTU, and connectivity issues.
- Build and operate core infrastructure services such as DNS, DHCP, NTP, NFS, SSH, LDAP / SSSD, and reverse proxy services where required.
- Develop Bash and Python scripts to automate operational checks, maintenance, troubleshooting, and repetitive administration tasks.
- Use Ansible and Git to standardize configuration, automate deployments, track changes, and maintain repeatable operating procedures.
- Support Docker, containerd, and container-based applications, with working knowledge of Kubernetes platform troubleshooting.
- Implement and maintain monitoring, logging, alerting, health checks, and capacity visibility for Linux platforms.
- Own production incidents from diagnosis through recovery, validation, root cause analysis, corrective action, and closure.
- Create SOPs, Runbooks, troubleshooting guides, and technical reports; mentor junior engineers and review high-risk technical changes.
Required Skills
- Years of hands-on Linux system administration or platform engineering experience in production environments.
- Strong Ubuntu / Debian or RHEL-based Linux administration and troubleshooting skills.
- Strong Linux boot, kernel, systemd, filesystem, and OS recovery knowledge.
- Strong Linux networking fundamentals including TCP/IP, routing, DNS, VLAN, bonding, firewall, and Layer 2 / Layer 3 troubleshooting.
- Strong storage knowledge including LVM, mdadm RAID, NFS, filesystem management, and disk troubleshooting.
- Hands-on Bash and Python scripting experience for operational automation.
- Hands-on experience with Ansible, Git, and configuration management practices.
- Experience operating and troubleshooting Docker or other container technologies.
- Able to diagnose production incidents independently using logs, metrics, system state, and structured root cause analysis.
Preferred Background (Not Mandatory)
- Experience operating enterprise Linux platforms, data centers, cloud infrastructure, or large-scale server environments.
- Experience with Kubernetes administration and troubleshooting in production environments.
- Experience with Prometheus, Grafana, Netdata, OpenTelemetry or similar observability tools.
- Experience with Terraform, CI/CD pipelines, or infrastructure automation practices.
- Experience designing or operating high availability, disaster recovery, backup, and platform resilience solutions.