Senior DevOps Engineer: AI-Powered Production Infra

Deriv

Cyberjaya

On-site

MYR 66,960 - 133,920

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Deriv is seeking an experienced production infrastructure and reliability engineer to own end-to-end systems that run millions of trades. You’ll design, build, and operate cloud, container, and database environments with AI-assisted tooling, automate delivery, and harden security alongside a dedicated security team.

You’ll lead incident response, implement self-healing patterns, and continuously improve availability while advancing autonomous AI-based operational capabilities.

Qualifications

  • 4+ years of DevOps, SRE, or production operations experience.
  • Hands-on experience shipping and supporting real production infrastructure.
  • Strong background in incident response, on-call operations, escalation handling, and high-availability web services.
  • Infrastructure automation and configuration management experience: Terraform, CloudFormation, Ansible, or equivalent.
  • Practical containerization experience with Docker and Kubernetes.
  • Cloud infrastructure experience with AWS, GCP, Azure, or similar, at scale.
  • Working knowledge of Linux and Windows Server environments, networking, databases, security, monitoring, and CI/CD.
  • Scripting or programming ability in Bash, Python, Go, PowerShell, or similar.
  • AI-first operating style using tools to build, troubleshoot, document, and improve infrastructure.
  • Ability to move fast while protecting reliability: prototype, test with real workloads, deploy carefully, learn from incidents.

Responsibilities

  • Production Infrastructure — Cloud, container, database, monitoring, and CI/CD environments for high-availability services; design it, not just patch it.
  • AI-Native Delivery — Automation design, infrastructure-as-code generation, testing, refactoring, documentation, and runbook creation with AI tooling.
  • Incident Response & Resilience — Alerting logic, remediation scripts, self-healing patterns, circuit breakers, fault-tolerant architecture.
  • Security Operations — Hardening, intrusion detection, configuration audits, and AI-enhanced threat detection with security team.
  • End-to-End Ownership — Define problem, architect, implement, deploy safely, monitor, and iterate.
  • Monitoring & Observability — Maintain Datadog, Grafana, and custom observability systems with AI-assisted analysis.
  • Incident Diagnosis — Diagnose production incidents, coordinate with developers, and drive durable infra improvements.
  • Autonomous Operations — Prototype and deploy autonomous AI systems for operations, including always-on agents for remediation.

Skills

DevOps
SRE practices
Incident response
On-call operations
Automation

Tools

Terraform
CloudFormation
Ansible
Docker
Kubernetes
AWS
GCP
Azure
Datadog
Grafana
PostgreSQL
Redis
Linux
Windows Server
Python
Go
PowerShell

Job description

Deriv is seeking an experienced production infrastructure and reliability engineer to own end-to-end systems that run millions of trades. You’ll design, build, and operate cloud, container, and database environments with AI-assisted tooling, automate delivery, and harden security alongside a dedicated security team.

You’ll lead incident response, implement self-healing patterns, and continuously improve availability while advancing autonomous AI-based operational capabilities.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer (AI & Production Infrastructure)
Senior DevOps Engineer (AI & Production Infrastructure)

Deriv • Cyberjaya

On-site
Staff AI Systems Engineer — Production-Scale
Staff AI Systems Engineer — Production-Scale

Deriv.com • Cyberjaya

On-site
MYR 180,000 - 300,000
Senior Engineer (Infra)
Senior Engineer (Infra)

Pertama Partners • Kuala Lumpur

On-site
MYR 80,000 - 120,000
Senior DevOps Engineer — AI-Powered Cloud Platform
Senior DevOps Engineer — AI-Powered Cloud Platform

Respond.io • Kuala Lumpur

On-site
MYR 60,000 - 120,000
Mental health allowance
Flexible working hours
Competitive compensation
Senior Offensive Security Engineer
Senior Offensive Security Engineer

Deriv • Cyberjaya

Hybrid
MYR 180,000 - 280,000
Senior DevOps Engineer - AWS, Kubernetes & CI/CD Architect
Senior DevOps Engineer - AWS, Kubernetes & CI/CD Architect

Involve Asia • Kuala Lumpur

On-site
MYR 180,000 - 300,000
DevOps Engineer
DevOps Engineer

WIZ.AI • Kuala Lumpur

On-site
MYR 90,000 - 140,000
Senior DevOps Engineer | Kubernetes & Cloud Automation
Senior DevOps Engineer | Kubernetes & Cloud Automation

Pride Global • Selangor

On-site
MYR 102,000 - 122,000
Senior Infra Engineer — Automation & Incident Stabilization
Senior Infra Engineer — Automation & Incident Stabilization

Cognizant • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Senior DevOps Engineer
Senior DevOps Engineer

Hyppies • Kuala Lumpur

On-site
MYR 120,000 - 180,000