Kuala Lumpur, Malaysia | Posted on 07/02/2026
Job Description
We are rapidly scaling global so1ware company specialising in next-gen Warehouse Management and Supply Chain Execu=on Systems. We are seeking a Senior DevOps Engineer to join the Cloud Hosting & Operations team, working as a hands-on individual contributor reporting to the Infrastructure Manager.
Key Responsibilities
- Deploy, operate, and maintain cloud infrastructure — ensuring consistent service levels and availability across all plaJorms.
- Manage infrastructure instances including routine patching, capacity monitoring, and hardware lifecycle upkeep.
- Configure and maintain network components including VPNs, firewalls, load balancers, direct connect and VPC configurations across all cloud platforms.
- Monitor resource utilisation and flag capacity trends or anomalies; support cloud cost optimisation actions such as right-sizing, scheduling, and resource clean-up.
- Automate provisioning, backup, recovery, patching, and system updates; write Python, Shell or Bash scripts to reduce manual operational toil.
- Provide hands-on infrastructure support to delivery teams during customer deployment phases and go-live milestones, ensuring environments are stable and handover-ready in WMS application.
- Manage and tune database infrastructure suppor2ng WMS: MySQL DB, MS SQL Server, or PostgreSQL — including HA, backup, and failover. This including middleware service (Redis, MongoDB, KaKa, Zookeeper, Nacos).
- Execute system upgrades, plaJorm migrations, and infrastructure scaling with minimal service downtime.
- Apply cloud security controls across all environments in line with ISO 27001, SOC 2, PDPA, and GDPR requirements; manage IAM policies, firewall rules, and access controls.
- Support the Infrastructure Manager during compliance reviews and audits by implementing required controls and providing environment evidence.
- Set up and maintain monitoring and alerting coverage environments using CloudWatch, Grafana, Prometheus and Huawei Cloud eye.
- Track and report on infrastructure availability, SLA performance, RTO/RPO metrics, and incident trends to the Infrastructure Manager.
- Maintain accurate, up-to-date documentation for all cloud environments, runbooks, and architecture diagrams.
- Liaise with cloud provider support channels for technical issue escalation and resolution, keeping the Infrastructure Manager informed on status and outcomes.
- Collaborate with R&D, Product, Delivery, and Customer Success teams across regions on infrastructure requirements and deployment coordination.
Requirements
- Minimum 5 years of hands-on experience in IT infrastructure, cloud operations, or DevOps engineering.
- Practical hands-on experience with AWS, Huawei Cloud, or GCP — experience across two or more platforms is a strong advantage.
- Strong background in deployment automation, CI/CD pipelines, monitoring, and cloud security implementation. Experience supporting enterprise so1ware environments (WMS, ERP, or Supply Chain platforms) is a strong advantage.
- Experience in application environments like Nginx, tomcat, DB (MYSQL, MSSQL or PostgreSQL), Middleware (Redis, MongoDB, Kada, Zookeeper, Nacos).
- Linux administration, VPC networking, routing, VPNs, and firewall configuration.
- Docker and Kubernetes (EKS / CCE / GKE); CI/CD pipelines using Jenkins, Bit Bucket, GitLab CI, or GitHub Ac=ons.
- HA/DR implementation, cloud cost optimisation support, and cloud security frameworks.
- Experience in Infrastructure-as-Code (Terraform, CloudFormation, Ansible) will be advantages.
- Strong ownership mindset, self-directed work style, and solid problem-solving abilities.
- Excellent written and spoken English; Mandarin proficiency is a strong plus for regional team and client collaboration.