Turn this role into an interview — a resume and cover letter built around what this employer wants.
Lavu Tech Solutions in Cheras, Kuala Lumpur is seeking a highly skilled Site Reliability Engineer with 10-15 years of experience to join our team in Petaling Jaya. The role focuses on ensuring reliability, availability and performance of our systems.
You will work closely with development and operations teams to apply SRE best practices, build automation, and participate in on-call rotations. Strong scripting and cloud experience are essential.
Jora Malaysia will close on 16th September 2026. Thank you for being with us, we are cheering you on as you continue your career journey.
Lavu Tech Solutions – Cheras, Kuala Lumpur
Please do not proceed if you are from overseas.Job Summary:
We are seeking a highly skilled Site Reliability Engineer (SRE) with 10 -15 years of experience to join our dynamic team in Petaling Jaya. The ideal candidate will possess a deep understanding of SRE principles and practices, ensuring the reliability, availability, and performance of our systems. You will work closely with development and operations teams to implement best practices in system reliability and automation.
Responsibilities:
Design, implement, and maintain scalable and reliable systems and services.
Monitor system performance and reliability, proactively identifying and resolving issues.
Develop and maintain automation tools for deployment, monitoring, and incident response.
Collaborate with development teams to ensure that reliability is built into the software development lifecycle.
Implement and manage incident response processes, including post mortem analysis.
Participate in on call rotations and provide support for production systems.
Continuously improve system architecture and operational processes.
Document processes, systems, and best practices for knowledge sharing.
Mandatory Skills:
Strong knowledge of Site Reliability Engineering (SRE) principles and practices.
Proficiency in scripting and programming languages such as Python, Go, or Ruby.
Experience with cloud platforms (AWS, Azure, GCP) and container orchestration (Kubernetes, Docker).
Solid understanding of networking, security, and system architecture.
Experience with monitoring and logging tools (Prometheus, Grafana, ELK stack).
Strong problem solving skills and the ability to work under pressure.
Preferred Skills:
Experience with CI/CD tools and practices.
Familiarity with configuration management tools (Ansible, Puppet, Chef).
Knowledge of database management and optimization (SQL, NoSQL).
Experience in a DevOps environment.
Strong communication and collaboration skills.
Qualifications:
Bachelor's degree in Computer Science, Engineering, or a related field.
10 -15 years of experience in Site Reliability Engineering or a related field.
Relevant certifications (e.g., Google Professional Cloud DevOps Engineer, AWS Certified DevOps Engineer) are a plus.
By creating an email alert, I agree to Jora's Terms and Privacy Policy and can unsubscribe anytime. If I'm below legal age requirements, I have parental consent for Jora to process my data.