DevOps Engineer

MTAI Sdn. Bhd.

Kuala Lumpur

On-site

MYR 200,000 - 320,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

MTAI Sdn. Bhd. seeks an experienced DevOps/Platform Engineer to design and implement infrastructure-as-code for our Sovereign AI platform.

You will shape provisioning standards, security practices and deployment workflows in a lean, fast-moving team. The role emphasizes automation, reliability and platform maturity as we scale AI workloads including GPU infrastructure. The candidate will be hands-on with Linux, Terraform, Kubernetes and Docker, and will work closely with leadership to advance

Qualifications

  • 5–10 years of experience in a DevOps, Platform Engineering, SRE or DevSecOps role.
  • Strong, hands-on Linux skills — OS and systems level.
  • Hands-on experience with Terraform, Kubernetes and Docker.
  • Hands-on experience with AWS or GCP in production environments.
  • Experience building and maintaining CI/CD pipelines.
  • Practical understanding of security fundamentals — vulnerability management, container/image hardening, access control, secrets management.

Responsibilities

  • Design, build and maintain infrastructure-as-code for our Sovereign AI platform using Terraform, Kubernetes and Docker.
  • Define processes and standards for provisioning, deployment and security in an early-stage team.
  • Build and maintain CI/CD pipelines for fast, reliable deployments.
  • Operate and improve reliability, scalability, and performance of cloud and on-prem infrastructure.
  • Implement monitoring, logging and alerting for observability and readiness.

Skills

DevOps
Platform engineering
SRE
DevSecOps
CI/CD
Startup experience
Automation
Security fundamentals

Tools

Terraform
Kubernetes
Docker
AWS
GCP

Job description

  • Design, build and maintain infrastructure-as-code for our Sovereign AI platform using Terraform, Kubernetes and Docker.
  • Define and establish processes and standards for infrastructure provisioning, deployment and security — as an early team member, you'll be shaping how things get done, not just following an existing playbook.
  • Build and maintain CI/CD pipelines to support fast, reliable deployment of infrastructure and platform changes.
  • Operate and improve the reliability, scalability and performance of our cloud and on-prem infrastructure.
  • Implement monitoring, logging and alerting to support platform observability and operational readiness.
  • Apply security best practices across the infrastructure lifecycle — from image/container hardening and vulnerability scanning to access control and secrets management.
  • Support incident response and troubleshooting across infrastructure and deployment issues.
  • Work closely with the Head of Sovereign AI Infra and the rest of the team to continuously improve automation, tooling and platform maturity.
  • Contribute to securing AI/ML and high-performance computing (HPC) workloads as the platform evolves, including GPU infrastructure and model-serving environments.
  • 5–10 years of experience in a DevOps, Platform Engineering, SRE or DevSecOps role.
  • Strong, hands-on Linux skills — comfortable operating, troubleshooting and optimising at the OS and systems level.
  • Strong command of the open-source infrastructure ecosystem, with hands‑on experience in Terraform, Kubernetes and Docker.
  • Hands‑on experience with AWS or GCP in a production environment.
  • Experience building and maintaining CI/CD pipelines.
  • Practical understanding of security fundamentals — vulnerability management, container/image hardening, access control, secrets management.
  • Comfortable working in a small, fast‑moving team where you're building and operating, not just designing.
  • Demonstrated ability to define and introduce processes, standards or best practices from scratch, ideally in a startup or early‑stage environment.

Highly Regarded

  • Experience securing or deploying generative AI, AI/ML or high‑performance computing (HPC) workloads.
  • Experience with GPU infrastructure and related tooling.
  • Contribution to open‑source infrastructure or security tooling projects.
  • Experience with observability stacks (e.g. Prometheus, Grafana), open source SIEM (Wazuh) and security scanning tools (e.g. Trivy, Snyk).
  • Prior experience in an infrastructure‑as‑a‑service or platform‑as‑a‑service environment.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Infrastructure Engineer
Senior Infrastructure Engineer

MTAI Sdn. Bhd. • Kuala Lumpur

On-site
MYR 180,000 - 300,000
Senior DevOps Engineer
Senior DevOps Engineer

PSL Group • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Senior Engineer (Infra)
Senior Engineer (Infra)

Pertama Partners • Kuala Lumpur

On-site
MYR 80,000 - 120,000
Senior Devops Engineer
Senior Devops Engineer

MHA Consultancy Services Sdn Bhd • Kuala Lumpur

On-site
MYR 120,000 - 180,000
AI DevOps Engineer (MLOps & Cloud)
AI DevOps Engineer (MLOps & Cloud)

DXC Technology Inc. • Petaling Jaya

On-site
MYR 100,000 - 150,000
Platform Engineer - AI Infra, Kubernetes & CI/CD
Platform Engineer - AI Infra, Kubernetes & CI/CD

MTAI Sdn. Bhd. • Kuala Lumpur

On-site
MYR 200,000 - 320,000
Senior DevOps Engineer
Senior DevOps Engineer

Aventra Group • Kuala Lumpur

On-site
MYR 120,000 - 160,000
Senior DevOps Engineer
Senior DevOps Engineer

P\\S\\L Group • Malaysia

On-site
MYR 180,000 - 300,000
Principal Engineer (Platform Engineering)
Principal Engineer (Platform Engineering)

PayNet (Payments Network Malaysia) • Kuala Lumpur

On-site
MYR 120,000 - 170,000
AI Platform Engineer
AI Platform Engineer

Systems Limited - APAC • Kuala Lumpur

On-site
MYR 180,000 - 300,000