Site Reliability Engineer(A66694)

Xiaomi Technology

Kuala Lumpur

On-site

MYR 385,000 - 578,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Xiaomi Technology in Singapore seeks a DevOps/Platform Reliability Engineer to ensure stable operation of the Mobile Business Group and overseas sales-service platforms. You will implement and monitor core O&M tasks, respond to incidents, manage capacity, and drive automation to cut manual work.

You will review architecture, identify bottlenecks, and propose cost-effective improvements while staying hands-on with Python/Go scripts and cloud platforms.

Qualifications

  • Bachelor's degree in Computer Science or related field.
  • Proficient in Python, Go or Shell scripting; capable of independently developing modules or platforms.
  • Hands-on experience with cloud-computing services; multi-cloud/hybrid-cloud platform management preferred.
  • Solid understanding of internet-system architectures, networking, load balancing, middleware, and HA disaster-recovery.
  • Willing to work on-site in Singapore; strong teamwork, responsibility, self-motivation, and proactive attitude.
  • Strong analytical thinking and business sense; propose stability and architecture optimization aligned with objectives.

Responsibilities

  • Ensure stability, reliability and efficient operation of the Mobile Business Group and overseas sales-service platforms.
  • Provide technical support for overseas businesses, covering core O&M tasks including resource delivery, incident handling, capacity management, resource governance, monitoring, and quality analysis.
  • Own and review technical architecture designs and identify risks; drive risk-mitigation initiatives.
  • Analyze system deficiencies, pinpoint bottlenecks and optimization opportunities, and formulate actionable solutions for cost-effective, highly-available operations.
  • Perform 7×24 on-call duties to respond to and resolve incidents for continuous stability.
  • Design and build automated platforms and services to improve O&M and delivery efficiency and reduce manual work.

Skills

Python
Go
Shell scripting
Cloud platforms
Multi-cloud
Networking
High availability

Education

Bachelor's degree in CS

Tools

Alibaba Cloud
Azure
AWS

Job description

  • Ensure the stability, reliability and efficient operation of the Mobile Business Group and overseas sales‑service business, and maintain high availability of services at all times.
  • Provide technical support for overseas businesses, covering core O&M tasks including resource delivery, incident handling, capacity management, resource governance, monitoring management and quality analysis.
  • Own and review technical architecture designs, evaluate the rationality of business architectures, proactively identify and assess risks, and drive or lead risk‑mitigation initiatives.
  • Conduct in‑depth analysis of system deficiencies, pinpoint system bottlenecks and optimization opportunities, and formulate actionable solutions to enhance system stability and enable cost‑effective, highly‑available system operations.
  • Perform 7×24‑hour On‑call duties to respond, track and resolve online incidents in a timely manner for continuous business stability.
  • Design and build automated platforms and services to improve O&M and delivery efficiency and reduce repetitive manual work.
Job Requirements
  • Bachelor’s degree in Computer Science or related disciplines.
  • Proficient in Python, Go or Shell scripting; capable of independently developing modules or platforms.
  • Hands‑on experience with cloud‑computing services; prior experience in multi‑cloud or hybrid‑cloud platform management (e.g. Alibaba Cloud, Azure, AWS) is preferred.
  • Solid understanding of internet‑system architectures and common infrastructure components, including networking, load balancing, middleware, and high‑availability disaster‑recovery architectures.
  • Willing to work on‑site in Singapore; strong teamwork awareness, sense of responsibility, self‑motivation and proactive working attitude.
  • Strong systematic thinking and business acumen; capable of proposing stability and architecture optimization solutions aligned with business objectives.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer — 24/7 Reliability & Automation
Site Reliability Engineer — 24/7 Reliability & Automation

Xiaomi Technology • Kuala Lumpur

On-site
MYR 385,000 - 578,000
IT Operations Engineer (Data Center)
IT Operations Engineer (Data Center)

BANYANO ASIA SDN BHD • Johor Bahru

On-site
MYR 180,000 - 240,000
Head, Investment Core Technology Services I Dev Ops Investment Banking Group
Head, Investment Core Technology Services I Dev Ops Investment Banking Group

Maybank • Kuala Lumpur

On-site
MYR 180,000 - 360,000
Site Reliability Engineers (Senior)
Site Reliability Engineers (Senior)

Talent Spot Group • Kuala Lumpur

On-site
MYR 240,000 - 360,000
Regional IT Operations Support Executive Mandarin Speaking
Regional IT Operations Support Executive Mandarin Speaking

CEKAP LINK TECH SDN. BHD. • Kuala Lumpur

On-site
MYR 67,000 - 112,000
Site Reliability Engineer
Site Reliability Engineer

Aisling Group • Kuala Lumpur

On-site
IT Application Operation Engineer Level 1 (7x24)
IT Application Operation Engineer Level 1 (7x24)

SF International • Selangor

Hybrid
MYR 67,000 - 112,000
IT Operations Manager | Application
IT Operations Manager | Application

VMM HOLDINGS SDN BHD • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Annual Bonus
Yearly Performance Bonus
Yearly Salary Increments
+4
Senior Site Reliability Engineer - Scale-Up Platform
Senior Site Reliability Engineer - Scale-Up Platform

Aisling Group • Kuala Lumpur

On-site
Site Reliability Engineer
Site Reliability Engineer

Confidential Jobs • Kuala Lumpur

On-site
MYR 150,000 - 230,000