Senior Site Reliability Engineer (L3)

CoinGecko

Malaysia

Hybrid

MYR 197,000 - 217,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Flexible Work Arrangement
Flexible Working Hours
Comprehensive Insurance Coverage
Virtual Share Options
Bonus
Parking Allowance
Meal Allowance
Learning Allowance
Social Activity Allowance
Annual Company Offsite

Job summary

CoinGecko is seeking an experienced DevOps/SRE leader to drive system architecture reviews, ensure reliable operations, and maintain security/compliance across platforms. You will own release planning, incident response, and disaster recovery while collaborating with cross-functional teams to scale our data platform globally.

The role emphasizes proactive automation, IaC, and capacity planning, with flexibility for hybrid work arrangements across our Malaysia offices and remote teams.

Qualifications

  • 3–5 years in managing software deployments and instrumentation in production environments with defined SLAs and SLOs.
  • Cloud platforms experience (AWS, CloudFlare, GCP) and IaC tools (Terraform, CloudFormation).
  • Bachelor’s degree in CS, InfoSec or related field, or professional certificates (e.g., AWS/GCP).
  • Scope of work: capable of delivering features as a sole contributor in open-ended projects.
  • Strong problem solving with data-driven reasoning and risk assessment.
  • Good communication and collaboration skills for cross-functional teams.
  • Nice to have startup experience, multiple tech stacks, and interest in decentralized tech.

Responsibilities

  • Review architecture and ensure best practices across all teams.
  • Own and ensure SLOs/SLAs are met; monitor operational metrics.
  • Manage security controls to meet enterprise requirements and compliance.
  • Lead strategic release plans (canary/blue-green) to reduce blast radius.
  • Lead incident response and post-mortems to resolve production issues.
  • Develop and implement disaster recovery plans and data recovery procedures.
  • Perform day-to-day tasks including onboarding-offboarding, patch management, and capacity planning.
  • Develop runbooks and extend technical documentation for audits.

Skills

DevOps
SRE
Cloud Operations
Automation
IaC
Python

Education

Bachelor's degree in CS / InfoSec or equivalent

Tools

AWS
GCP
CloudFlare
Terraform
CloudFormation

Job description

CoinGecko is a global leader in tracking cryptocurrency data. Operating since 2014, CoinGecko has built the world's largest cryptocurrency data platform, tracking over 10,000 tokens across more than 400 exchanges, serving over 300 million page views in more than 100 countries. We are proud to have played a major part in mainstream awareness, adoption, and education of cryptocurrency globally.

We at CoinGecko believe that cryptocurrency and blockchain will define the future of finance, bringing greater financial and economic freedom around the world. In anticipation of that future, CoinGecko is building the foundation to scale cryptocurrency market data to serve billions.

We practice transparent salaries and a level structure at CoinGecko
  • The salary range for the L3 position is RM16,407 - RM18,047.
  • For more junior candidates, we may evaluate you as a Mid-Level (L2), with a salary range of RM11,702 - RM12,872.
  • Learn more about our level structure at CoinGecko's Career Progression.
What You Will Be Doing:
  • System Architecture: Review architecture and software components with software engineers. Ensure best practices are consistent across all teams.
  • Operational Excellence: Own and ensure SLOs and SLAs are met. Monitor operational metrics and lead improvement plans. Develop and maintain tools including infra-as-code resources to scale operations and allow other teams to be autonomous.
  • Security and Compliance: Manage and audit security controls to meet enterprise requirements. Implement and maintain best practices and compliance standards. Collaborate with legal and compliance to assess overall risk management.
  • Release Planning: Lead strategic release plans (e.g., canary or blue-green deployments) to reduce blast radius and allow for faster reversal during release failures. Work closely with developers for pre-release requirements including provisioning test environments. Conduct ad hoc performance tests based on requirements.
  • Incident Management: Lead incident response and post-mortems to resolve production issues, identify root-causes and prevent future occurrences.
  • Disaster Recovery: Develop and implement DR plans and procedures, including data recovery and fault injection simulations on production replica.
  • Daily Operations: Perform and improve day-to-day tasks including access onboarding-offboarding, config and patch management etc. Plan capacity to ensure our systems have sufficient capacity to handle peak demand while optimizing cost.
  • Documentation: Develop and extend runbooks, documentation and other technical assets. Support periodic technical audits as required.
  • Sharpen the Saw: Stay up-to-date with emerging trends and technologies in software development and contribute to knowledge sharing. Learn advanced architecture standards and new tools that improve the team’s code base and productivity. Demonstrate thorough understanding of a subject matter and how to apply it effectively.
  • Team player: Collaborating with cross-functional teams to ensure smooth deployment and operation of software releases. Answer technical questions from other teams or outside the organization.
  • Coaching: Provide feedback on the performance of junior staff and participate in people development initiatives.
  • Support any ad hoc tasks as required by the company.
What We Look For In You:
  • Proven Track Record: 3 to 5 years in managing software deployments and instrumentation in production environments with defined SLAs and SLOs. Strong knowledge of software delivery and devops principles.
  • Cloud Operations: Experience with cloud platforms (e.g., AWS, CloudFlare, GCP) and infrastructure-as-code tools (e.g., Terraform, CloudFormation). Strong programming and scripting skills, preferably in languages such as Python, Go, or Ruby.
  • Accreditation: Bachelor’s degree in Comp Sci., InfoSec or similar fields, or professional certificates e.g. Certified DevOps Professional, Certified Solutions Architect Professional in AWS or GCP.
  • Scope of Work: Fully capable of taking substantial features from concept to shipping as a sole contributor. Works effectively in open-ended projects and is self-sufficient to deep dive and evaluate multiple solutions to a problem.
  • Problem Solving: Solve hard problems with many constraints, using sound judgment to assess risks and present arguments in a well-structured, data-backed, written narrative. Have passion, creativity and empathy for users.
  • Quick Thinking: Able to derive information, think critically and make snap judgements based on measured data in high pressure situations.
  • People Skills: Strong communicator who is able to build positive working relationships between teams and form relationships with key customers. You must have experience supporting on-call rotations for 24x7 services to troubleshoot, perform runbooks or escalate incidents.
  • Nice to have:
  • // Experience working in a growth stage startup.
  • // Experience building applications in different tech stacks.
  • // Keen interest in decentralized technologies and its applications including cryptocurrencies.
Some of the perks while at CoinGecko
  • Flexible Work Arrangement: We offer a flexible work arrangement that supports both remote and in-office collaboration. Our offices at 1Powerhouse (Malaysia) and WeWork (Singapore) are available whenever you want to connect and collaborate with your teammates.
  • Flexible Working Hours: No 9-5 structure, work the hours you need to get your tasks done.
  • Comprehensive Insurance Coverage: We provide life, medical, and critical illness insurance.
  • Virtual Share Options: You'll be entitled to virtual options, with terms and conditions.
  • Bonus: You’ll be entitled to a bonus, with terms and conditions.
  • Parking Allowance: Allocated on a claim basis to ease the cost of travelling.
  • Meal Allowance: You will be given a monthly fixed allowance of RM600 or SGD400 to subsidize the cost of your meals.
  • Learning Allowance: You will be allocated an annual budget of USD500 (claim basis) to help you continuously learn in the pursuit of your professional and personal development.
  • Social Activity Allowance: Want to set a date to watch a movie or play futsal with your colleagues? Get it organized and we subsidize a portion (claim basis) of the cost.
  • Annual Company Offsite: We gather once a year to meet each other in person, reflect on the year, and partake in social activities!

CoinGecko is an equal employment opportunity employer. Qualified candidates are considered for employment without regard to race, religion, gender, gender identity, sexual orientation, national origin, age, military or veteran status, disability, or any other characteristic protected by applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer
Senior DevOps Engineer

CoinGecko • Malaysia

Hybrid
MYR 16,000 - 18,000
Flexible Work Arrangement
Flexible Working Hours
Comprehensive Insurance Coverage
+7
Senior Site Reliability Engineer (L3)
Senior Site Reliability Engineer (L3)

CoinGecko • Kuala Lumpur

On-site
MYR 178,560 - 207,453
Flexible Work Arrangement — remote and
Comprehensive Insurance Coverage
Virtual Share Options
+6
Remote Senior SRE - Cloud Reliability & Automation Lead
Remote Senior SRE - Cloud Reliability & Automation Lead

CoinGecko • Kuala Lumpur

On-site
MYR 178,560 - 207,453
Flexible Work Arrangement — remote and
Comprehensive Insurance Coverage
Virtual Share Options
+6
Senior Site Reliability Engineer (L3) — Remote Malaysia
Senior Site Reliability Engineer (L3) — Remote Malaysia

CoinGecko • Malaysia

On-site
Remote Work Flexibility
Comprehensive Insurance Coverage
Virtual Share Options
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GrabTaxi Holdings Pte. Ltd. • Petaling Jaya

On-site
Term Life Insurance
Comprehensive Medical Insurance
Flexible Work Arrangements
+2
DevOps Enginer
DevOps Enginer

JobCubby • Malaysia

On-site
MYR 120,000 - 180,000
Senior Site Reliability Engineer — Flexible, Remote-Ready
Senior Site Reliability Engineer — Flexible, Remote-Ready

CoinGecko • Malaysia

Hybrid
MYR 197,000 - 217,000
Flexible Work Arrangement
Flexible Working Hours
Comprehensive Insurance Coverage
+7
DevOps Enginer
DevOps Enginer

Linuxconfig • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Senior DevOps Engineer – Remote + Flexible Hours
Senior DevOps Engineer – Remote + Flexible Hours

CoinGecko • Malaysia

Hybrid
MYR 16,000 - 18,000
Flexible Work Arrangement
Flexible Working Hours
Comprehensive Insurance Coverage
+7
Senior DevOps Engineer
Senior DevOps Engineer

Accenture Southeast Asia • Cyberjaya

On-site
MYR 180,000 - 240,000