Senior Engineer (Infra)

Pertama Partners

Kuala Lumpur

On-site

MYR 80,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Pertama Partners in Kuala Lumpur is seeking a full-time AI Infrastructure Engineer to deploy and maintain reliable AI systems in production environments. Responsibilities include designing deployment architectures, building CI/CD pipelines, and ensuring compliance with security standards. The ideal candidate has over 5 years of experience in production infrastructure, with expertise in Docker, Kubernetes, and cloud platforms like AWS and GCP. This role offers a technical challenge and opportunity to make significant architectural decisions.

Qualifications

  • 5+ years operating production infrastructure (compute, networking, observability).
  • Deep knowledge of containerization and orchestration (Docker, Kubernetes).
  • Experience with IaC and cloud platforms (Terraform, AWS/GCP).
  • Track record maintaining high-availability services (99.9%+ uptime).

Responsibilities

  • Design and implement deployment architectures for client AI systems.
  • Build CI/CD pipelines and Infrastructure-as-Code configurations.
  • Set up observability: metrics, logs, traces, alerting.
  • Ensure security compliance (SOC 2, ISO 27001, client-specific requirements).
  • Respond to incidents and implement preventive measures.
  • Document runbooks and train client teams on operations.

Job description

Deploy reliable AI infrastructure for enterprise production

full-time

Highly competitive

Overview

Own the infrastructure that runs AI systems in production. You'll design deployment architectures, set up monitoring and alerting, ensure security compliance, and keep everything running smoothly at enterprise scale. This isn't ticket-driven ops work. You're building the platform that enables rapid, reliable delivery across multiple client environments. Expect to write code, design systems, and debug complex distributed failures.

Responsibilities
  • Design and implement deployment architectures for client AI systems
  • Build CI/CD pipelines and Infrastructure-as-Code configurations
  • Set up observability: metrics, logs, traces, alerting
  • Ensure security compliance (SOC 2, ISO 27001, client-specific requirements)
  • Respond to incidents and implement preventive measures
  • Document runbooks and train client teams on operations
Requirements
  • 5+ years operating production infrastructure (compute, networking, observability)
  • Deep knowledge of containerization and orchestration (Docker, Kubernetes)
  • Experience with IaC and cloud platforms (Terraform, AWS/GCP)
  • Track record maintaining high-availability services (99.9%+ uptime)
Preferred Qualifications
  • Built CI/CD pipelines from scratch
  • Responded to production incidents at 3am and improved alerting so it doesn't happen again
  • Experience with service mesh, API gateways, or distributed tracing
  • Open-source contributions to infrastructure projects
Nice to Have

Python AWS

A Day in the Life

Morning: Review overnight alerts (none, because you built good alerts). Mid-morning: Design review for multi-region deployment. Afternoon: Implement automated failover for critical service. Evening: Write runbook for new deployment pattern.

Why This Role

Build infrastructure for systems that millions depend on. Make architectural decisions that matter. Work with modern tools and patterns. No legacy baggage to maintain.

Technical Challenge

This role requires completing a technical challenge as part of the application process. Challenge: Medium: High-Availability Service

Rotating on-call (one week per month). Incidents are rare because we invest in reliability. Compensation for on-call time and incident response.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer (AI & Production Infrastructure)
Senior DevOps Engineer (AI & Production Infrastructure)

Deriv • Cyberjaya

On-site
Senior Engineer (Data)
Senior Engineer (Data)

Pertama Partners • Kuala Lumpur

On-site
MYR 100,000 - 150,000
Staff Applied AI Engineer
Staff Applied AI Engineer

Deriv.com • Cyberjaya

On-site
MYR 180,000 - 300,000
AI DevOps Engineer (MLOps & Cloud)
AI DevOps Engineer (MLOps & Cloud)

DXC Technology Inc. • Petaling Jaya

On-site
MYR 100,000 - 150,000
AI Engineer (Full Stack Engineering)
AI Engineer (Full Stack Engineering)

MetLife • Kuala Lumpur

On-site
MYR 120,000 - 210,000
AI Infrastructure & Orchestration Lead
AI Infrastructure & Orchestration Lead

SNS Network (M) Sdn. Bhd. • Petaling Jaya

On-site
MYR 180,000 - 260,000
AI Engineer SNS Network Right Choice with the Right People
AI Engineer SNS Network Right Choice with the Right People

SNS Network (M) Sdn. Bhd. • Petaling Jaya

On-site
MYR 180,000 - 260,000
Senior DevOps Engineer
Senior DevOps Engineer

Involve Asia • Kuala Lumpur

On-site
MYR 180,000 - 300,000
Senior DevOps Engineer
Senior DevOps Engineer

Aventra Group • Kuala Lumpur

On-site
MYR 120,000 - 160,000
DevOps Engineer
DevOps Engineer

Ad Astra Consultants • Kuala Lumpur

On-site
MYR 150,000 - 210,000