Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Vuly Play in Kuala Lumpur (KLCC) is seeking a Lead DevOps Engineer to own our AWS-centric cloud ecosystem and deployment pipelines for a PHP monolith and Laravel apps. This hands-on role reports to the CTO and collaborates with the Tech Lead to enhance security, reliability, and performance of global platforms.
You will manage CI/CD, AWS infrastructure (Elastic Beanstalk, EC2, RDS, S3, IAM, Route 53), implement IaC, improve monitoring, and drive cost efficiencies.
We create premium play experiences that get families outside and spark imagination. We are expanding into new regions and launching bold products.
As part of our growing KL hub, we are seeking a Lead DevOps Engineer to serve as the primary technical owner of our AWS-centric cloud ecosystem and deployment pipelines for a custom PHP monolith and some sister Laravel applications. Reporting to the Chief Technology Officer (CTO) and working closely with our Tech Lead, this is a hands-on, high-impact technical role.
We need a dedicated specialist who knows how to achieve operational security and infrastructure resilience. As the primary domain owner, you will be responsible for building secure CI/CD pipelines, optimising AWS infrastructure, driving cost efficiencies, and establishing robust incident readiness to keep our global digital platforms performing seamlessly.
CI/CD & Deployment Management: Build, test, release, rollback, and deployment workflows across Docker and AWS Elastic Beanstalk.
Cloud Infrastructure Administration: Manage AWS infrastructure including Elastic Beanstalk, EC2, load balancers, auto scaling, RDS, S3, CloudWatch, CloudTrail, EventBridge, IAM, Route 53, Certificate Manager, VPC, and related services.
Environment Configuration & Parity: Oversee development, staging, testing, UAT, and production environments to ensure configuration consistency, environment parity, and production readiness.
Secure Configuration Management: Manage application configurations, environment variables, secrets, credentials, API keys, service accounts, and certificates securely.
Monitoring & Observability: Own the system’s health visibility by managing monitoring, logging, metrics, tracing, dashboards, alerts, and uptime checks.
Incident Readiness & Outage Recovery: Develop and maintain incident response plans, runbooks, outage recovery procedures, rollback strategies, and conduct post-incident reviews.
Business Continuity & High Availability: Oversee backup, restore, disaster recovery, failover, high availability, and business continuity processes.
Security & Vulnerability Management: Execute security patching, vulnerability scans, access control management, IAM reviews, MFA/SSO configurations, and enforce minimum-access permissions.
Network & Domain Operations: Manage domains, registrars, Route 53, DNS configurations, subdomains, redirects, and certificate validations.
Cloud Cost Optimisation: Control cloud spend through budget alerts, regular usage reviews, right-sizing resources, scaling rules, and infrastructure waste reduction.
Performance & Capacity Planning: Improve reliability and resilience through autoscaling optimisation, capacity planning, load testing, and removing performance bottlenecks.
Infrastructure as Code (IaC): Implement infrastructure as code and version-controlled infrastructure changes to reduce configuration drift.
Developer Tooling & Workflow Support: Support developer velocity by maintaining local Docker setups, automating manual tasks, writing internal scripts, and assisting with onboarding.
Third-Party & Compliance Support: Manage vendor platform access, track service health, maintain deployment/access logs, and provide evidence for compliance audits and security reviews.
5+ Years of Senior AWS & DevOps Experience: A proven track record as a hands-on engineer managing production cloud infrastructure across core services (EC2, Elastic Beanstalk, RDS, S3, IAM, CloudWatch, Route 53, VPC).
E-Commerce & High-Traffic Infrastructure Experience: Experience supporting high-availability cloud environments for high-traffic e-commerce platforms, particularly during peak promotional events.
CI/CD & Containerisation Mastery: Deep hands-on experience building, maintaining, and automating secure CI/CD pipelines and release workflows using Docker and version control.
Daily Operational Security Focus: Strong practical expertise in managing IAM access controls, handling secrets/certificates, automating vulnerability patching, and auditing system logs.
Incident Readiness & Reliability: Demonstrated experience setting up uptime monitoring, disaster recovery failovers, tested backups, and actionable runbooks to prevent outages.
Disciplined Escalation & Documentation: Exceptional technical communication with a proven ability to work under senior engineering sign-off (CTO), strictly adhering to decision rights, maintaining audit records, and writing clear system documentation.
Clear Communication & Plain-Language Technical Writing: A knack for breaking down complex infrastructure issues, translating technical security risks into clear business impacts, and writing actionable runbooks.
Presence: This is an office-based role in KLCC, Monday to Friday.
Compliance & Audit Readiness: Prior experience collecting deployment records, access logs, and incident tracking data to support compliance audits and formal security reviews.
Cost Governance Frameworks: Experience establishing automated cloud budget tracking alerts and utilising optimisation platforms to systematically uncover infrastructure waste.
Formal Qualifications: AWS Certified DevOps Engineer, AWS Certified Security, Certified Kubernetes Administrator (CKA) or Docker Certified Associate (DCA), Tertiary Degree in Technology.
Opportunity to personally drive the modernisation and architectural evolution of our cloud hosting infrastructure.
Your contributions will directly power our production systems, driving real-time performance, security, and reliability for our global customers and internal engineering teams.
You will partner with a pragmatic, high-performing engineering team that prioritises straight talk and real outcomes.