DevOps / Operation Engineer

Scicom MSC Berhad

Kuala Lumpur

On-site

MYR 66,960 - 133,920

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical Insurance provided

Job summary

Scicom MSC Berhad is seeking an experienced Operation Engineer to oversee our cloud-native infrastructure, AI-related services, and database systems, ensuring high availability and top performance of our online services. You will deploy and operate Kubernetes clusters, develop automation scripts in Python, Java, or Go, and manage databases with routine maintenance, backups, and recovery.

You will build observability for logs and metrics, optimize CI/CD workflows with development teams, and

Qualifications

  • Hands-on experience deploying and managing Kubernetes clusters.
  • Proficiency in Python, Java, or Go.
  • Experience operating or maintaining AI/ML infrastructure.
  • Fluent Mandarin for daily communication and strong English for documentation.
  • Good understanding of Linux, networking, and cloud-native architecture.
  • Experience with OpenSearch, Grafana, ELK Stack, APM tools, and Argo ecosystem preferred.

Responsibilities

  • Deploy, operate, monitor, and troubleshoot Kubernetes clusters and containerized workloads.
  • Develop automation scripts using Python, Java, or Go to reduce manual workload.
  • Manage and maintain databases including maintenance, performance tuning, backups, and recovery.
  • Build and maintain observability systems for log collection and metrics monitoring.
  • Collaborate with development teams to optimize CI/CD workflows and improve delivery efficiency.
  • Perform daily system checks, incident handling, and root cause analysis.

Skills

DevOps
Kubernetes
Python
Java
Go
AI/ML infra
Mandarin
English
Linux
Networking
Cloud-native
Root Cause Analysis

Tools

OpenSearch
Grafana
ELK Stack
APM tools
Argo ecosystem

Job description

We are looking for an experienced Operation Engineer to manage, maintain, and optimize our cloud-native infrastructure, AI-related services, and database systems. The successful candidate will collaborate with cross-functional teams to ensure the high availability, stability, and performance of our online services.

Job Description
  • Deploy, operate, monitor, and troubleshoot Kubernetes clusters and containerized workloads.
  • Develop automation scripts and internal tools using Python, Java, or Go to reduce manual workload.
  • Manage and maintain databases, including routine maintenance, performance tuning, backup, recovery, and fault resolution.
  • Build and maintain observability systems for log collection, metrics monitoring, and performance tracking.
  • Collaborate with development teams to optimize CI/CD workflows and improve delivery efficiency.
  • Perform daily system checks, incident handling, and root cause analysis.
  • Complexity of Cloud-Native Infrastructure: Managing and troubleshooting Kubernetes clusters and containerized workloads is inherently complex, requiring constant vigilance to ensure stability in production environments.
  • Maintaining Observability: Building and maintaining systems for log collection, metrics monitoring, and performance tracking across a distributed environment is a significant ongoing technical task.
  • Database Management: Beyond just maintenance, the role involves complex performance tuning, backups, and recovery, which are critical to data integrity and system availability.
  • Optimizing CI/CD Workflows: Balancing the need for rapid software delivery with stability is a core DevOps challenge. The engineer must collaborate with development teams to ensure pipelines are efficient, reliable, and do not introduce errors into production.
  • Skill Breadth & Adaptation: The need to maintain AI/ML infrastructure alongside traditional databases and cloud services requires a broad, high-level technical skillset and the ability to stay updated with rapidly evolving technology.
  • Root Cause Analysis: Moving beyond "patching" issues to performing deep root cause analysis (RCA) is required to implement long-term optimization plans rather than just treating symptoms.
Requirements
  • Strong experience in DevOps, system operations, or cloud-native environments.
  • Hands‑on experience deploying, managing, and troubleshooting Kubernetes clusters.
  • Proficiency in at least one programming language: Python, Java, or Go.
  • Experience operating or maintaining AI/ML systems and related infrastructure.
  • Fluent Mandarin for daily communication and strong English for documentation.
  • Good understanding of Linux, networking, and cloud-native architecture.
  • Strong problem‑solving and troubleshooting skills.
  • Experience with OpenSearch, Grafana, ELK Stack, APM tools, and the Argo ecosystem are preferred.
The Package
  • Attractive Salary: RM6000 up to RM 12000
  • Performance related allowance for confirmed staff
  • Medical Insurance provided
  • Working location: Bangsar
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DevOps / Operation Engineer (English and Mandarin)
DevOps / Operation Engineer (English and Mandarin)

Scicom MSC Berhad • Kampung Cendana

On-site
MYR 67,000 - 134,000
Medical insurance provided
Senior DevOps Engineer
Senior DevOps Engineer

Pride Global • Selangor

On-site
MYR 102,000 - 122,000
Junior Devops Engineer
Junior Devops Engineer

Pride Global • Selangor

On-site
MYR 76,000 - 91,000
Devops Engineer- Cloud CI/CD
Devops Engineer- Cloud CI/CD

Accion Labs • Kuala Lumpur

On-site
MYR 120,000 - 180,000
DevOps Engineer
DevOps Engineer

Lavu Tech Solutions Sdn Bhd • Cyberjaya

On-site
MYR 120,000 - 240,000
Senior DevOps Engineer
Senior DevOps Engineer

Pride Global • Shah Alam

On-site
MYR 67,000 - 112,000
DevOps Engineer
DevOps Engineer

TAT IT Technolgies • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Junior Devops Engineer
Junior Devops Engineer

Pride Global • Shah Alam

On-site
MYR 50,000 - 84,000
Junior Devops Engineer - CI-CD Monitoring
Junior Devops Engineer - CI-CD Monitoring

Pride Global • Kuala Lumpur

On-site
MYR 50,000 - 84,000
IT Operations Engineer - (English and Mandarin)
IT Operations Engineer - (English and Mandarin)

Scicom MSC Berhad • Kampung Cendana

On-site
MYR 61,000 - 73,000
Medical insurance