Site Reliability Engineer

LAVU TECH SOLUTIONS SDN. BHD.

Petaling Jaya

On-site

MYR 180,000 - 300,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

LAVU TECH SOLUTIONS SDN. BHD. seeks a seasoned Site Reliability Engineer (SRE) with 7–10 years of experience to uphold system reliability in Petaling Jaya.

You will partner with development and operations to build scalable infrastructure and maintain high availability. Strong automation and cloud/Kubernetes skills are essential. The role focuses on designing reliable services, implementing CI/CD pipelines, and participating in on-call rotations while documenting best practices for team knowledge

Qualifications

  • Strong understanding of Site Reliability Engineering principles and practices.
  • Proficiency with scripting languages such as Python, Bash, or Go.
  • Experience with cloud platforms (AWS, Azure, GCP) and container orchestration (Kubernetes, Docker).
  • Solid understanding of networking, security, and system architecture.
  • Experience with monitoring and logging tools (Prometheus, Grafana, ELK).
  • Familiarity with IaC tools (Terraform or Ansible) and service mesh technologies (Istio, Linkerd).

Responsibilities

  • Design, implement, and manage scalable and reliable systems and services.
  • Monitor system performance and troubleshoot issues to ensure high availability.
  • Develop and maintain automation tools for deployment, monitoring, and incident response.
  • Collaborate with software engineering teams to improve system reliability and performance.
  • Implement and manage CI/CD pipelines to streamline software delivery.
  • Conduct post-incident reviews and implement improvements to prevent recurrence.
  • Participate in on-call rotations and respond to incidents as needed.
  • Document processes, systems, and best practices for knowledge sharing.

Skills

SRE principles
Python
Bash
Go
Cloud platforms
Kubernetes
Docker
Networking
Security
System architecture
Monitoring tools
IaC
Service mesh
SQL NoSQL

Education

Bachelor's degree in Computer Science, Engineering, or related field

Tools

Prometheus
Grafana
ELK stack
Terraform
Ansible
Istio
Linkerd

Job description

We are seeking a highly skilled Site Reliability Engineer (SRE) with 7-10 years of experience to join our dynamic team in Petaling Jaya. The ideal candidate will possess a deep understanding of Site Reliability Engineering principles and practices, ensuring the reliability, availability, and performance of our systems. You will work closely with development and operations teams to build and maintain scalable and resilient infrastructure.

Key responsibilities
  • Design, implement, and manage scalable and reliable systems and services.
  • Monitor system performance and troubleshoot issues to ensure high availability.
  • Develop and maintain automation tools for deployment, monitoring, and incident response.
  • Collaborate with software engineering teams to improve system reliability and performance.
  • Implement and manage CI/CD pipelines to streamline software delivery.
  • Conduct post incident reviews and implement improvements to prevent recurrence.
  • Participate in on call rotations and respond to incidents as needed.
  • Document processes, systems, and best practices for knowledge sharing.
About you
  • Strong knowledge of Site Reliability Engineering (SRE) principles and practices.
  • Proficiency in scripting languages such as Python, Bash, or Go.
  • Experience with cloud platforms (AWS, Azure, GCP) and container orchestration (Kubernetes, Docker).
  • Solid understanding of networking, security, and system architecture.
  • Experience with monitoring and logging tools (Prometheus, Grafana, ELK stack).
  • Strong problem solving skills and the ability to work under pressure.
  • Familiarity with Infrastructure as Code (IaC) tools such as Terraform or Ansible.
  • Experience with service mesh technologies (Istio, Linkerd).
  • Knowledge of database management and optimization (SQL, NoSQL).
  • Bachelor's degree in Computer Science, Engineering, or a related field.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer: Scale, Resilience & Automation
Senior Site Reliability Engineer: Scale, Resilience & Automation

LAVU TECH SOLUTIONS SDN. BHD. • Petaling Jaya

On-site
MYR 180,000 - 300,000
Site Reliability Engineer
Site Reliability Engineer

Aisling Group • Kuala Lumpur

On-site
Site Reliability Engineer: Build Resilient, Scalable Systems
Site Reliability Engineer: Build Resilient, Scalable Systems

Career Wise • Kuala Lumpur

On-site
MYR 80,000 - 120,000
Senior Site Reliability Engineer - Scale-Up Platform
Senior Site Reliability Engineer - Scale-Up Platform

Aisling Group • Kuala Lumpur

On-site
Site Reliability Engineer: Automation & Cloud Ops
Site Reliability Engineer: Automation & Cloud Ops

Pan Asia Group • Kuala Lumpur

On-site
MYR 140,000 - 210,000
Site Reliability Engineer: Build Reliable, Scalable Systems
Site Reliability Engineer: Build Reliable, Scalable Systems

Setel Ventures • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Leisure area with video games
Casual dress (jeans)
Pantry with coffee, tea and snacks
+2
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Esker • Kuala Lumpur

Hybrid
MYR 120,000 - 180,000
Site (Software) Reliability Engineer (SRE)
Site (Software) Reliability Engineer (SRE)

Realtek • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Site Reliability Engineer (SRE) Esker Asia · Kuala Lumpur, Malaysia ·
Site Reliability Engineer (SRE) Esker Asia · Kuala Lumpur, Malaysia ·

Esker, Inc. • Kuala Lumpur

On-site
MYR 180,000 - 240,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Career Wise • Kuala Lumpur

On-site
MYR 80,000 - 120,000