Manager, Site Reliability Engineering

BNB Chain

Vancouver

On-site

CAD 140,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

LayerZero is seeking a Manager of SRE to lead a team responsible for the reliability, performance, and scalability of blockchain node infrastructure and platform services. You will guide architecture decisions, balance people leadership with hands-on technical judgment, and drive incident response across LayerZero’s services.

You will design reliability strategies, mentor engineers, and collaborate with engineering and product teams to ensure uptime and rapid evolution of systems.

Qualifications

  • 6+ years in SRE, DevOps, or infrastructure engineering
  • 2+ years directly leading or managing a technical team
  • Experience with blockchain node infrastructure (validator/full/archive nodes, RPC optimization)
  • Strong proficiency in TypeScript or Golang
  • Advanced knowledge of Unix/Linux internals and distributed systems/high-availability design
  • 3+ years running Kubernetes in production, including Helm chart authoring at scale
  • Proven track record building or scaling on-call/incident response processes
  • Excellent communication skills and ability to represent the team to leadership and hires/retains engineers

Responsibilities

  • Lead and develop a team of SREs, setting technical direction and growth plans
  • Own the reliability strategy for blockchain node infrastructure with SLOs, capacity planning, and incident response
  • Collaborate with Engineering leadership and Product/Platform teams to align reliability investments with business priorities
  • Drive infrastructure-as-code practices, focusing on Kubernetes and Helm at scale
  • Establish and improve on-call structures, incident detection/triage automation, and postmortem culture
  • Remain hands-on: review designs, tackle complex incidents, and set the technical bar for the team

Skills

Leadership
Communication
SRE/DevOps
TypeScript
Golang
Unix/Linux
Distributed systems

Education

Bachelor's degree in Computer Science or related field

Tools

Kubernetes
Helm

Job description

LayerZero

The Future is Omnichain.

Founded in 2021, LayerZero's vision is to create a community of cross-chain developers, building dApps that are no longer constrained by individual blockchain capabilities. With LayerZero's simple, generic messaging protocol, builders will develop cross-chain dApps designed to unify the power of individual blockchains.

We are funded by the best investors in the world including:

a16z, Sequoia, PayPal, Binance Ventures, Coinbase Ventures, Uniswap Labs, Circle Ventures, Delphi Digital, and many more.

ABOUT THE ROLE

At LayerZero, our Site Reliability Engineering (SRE) team is at the intersection of software and systems engineering, dedicated to crafting and maintaining large-scale, resilient systems. Our goal is to ensure that all LayerZero services - ranging from critical internal systems to those external users interact with - are reliable, meet the uptime expectations of our users, and continuously evolve at a swift pace. Our SRE professionals will monitor our system's capacity and performance to uphold these standards.

As Manager of SRE, you'll lead a team of engineers responsible for the reliability, performance, and scalability of our blockchain node infrastructure and platform services - while staying technically sharp enough to guide architecture decisions and jump into critical incidents. You'll balance people leadership with hands-on technical judgment, shaping how the team works, grows, and scales alongside LayerZero.

WHAT YOU'LL DO
  • Lead and develop a team of SREs - setting technical direction, growth plans, and performance expectations.
  • Own the reliability strategy for blockchain node infrastructure across a variety of DLTs, including SLOs, capacity planning, and incident response.
  • Partner with Engineering leadership and Product/Platform teams to align reliability investments with business priorities.
  • Drive infrastructure-as-code practices, with a focus on Kubernetes and Helm at scale.
  • Establish and continuously improve on-call structure, incident detection/triage automation, and postmortem culture.
  • Stay hands-on: review designs, dig into complex incidents, and set the technical bar for the team.
ABOUT YOU
  • Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience.
  • 6+ years in SRE, DevOps, or infrastructure engineering, including 2+ years directly managing or leading a technical team.
  • Deep familiarity with blockchain node infrastructure (validator/full/archive nodes, RPC optimization, etc.).
  • Strong proficiency in TypeScript or Golang, with the judgment to know when to write code vs. delegate.
  • Advanced knowledge of Unix/Linux internals and distributed systems / high-availability design.
  • 3+ years running Kubernetes in production, including Helm chart authoring at scale.
  • Track record building or scaling an on-call/incident response process.
  • Excellent communication skills - able to represent the team to leadership and hire/retain strong engineers.

Equal Opportunity Employer

LayerZero Labs is committed to fostering a diverse and inclusive workplace.

LayerZero Labs is an equal opportunity employer and does not discriminate on the basis of race, national origin, religion, gender, gender identity, sexual orientation, marital status, protected veteran status, disability, age, or any other legally protected status.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Manager, Site Reliability Engineering Vancouver, BC 9 Sept 2026
Manager, Site Reliability Engineering Vancouver, BC 9 Sept 2026

Hoodl2 • Vancouver

On-site
CAD 120,000 - 150,000
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

LayerZero Labs • Vancouver

On-site
CAD 140,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

King River Capital Group • Vancouver

On-site
CAD 120,000 - 200,000
Comprehensive benefits
Equity options
Site Reliability Engineer
Site Reliability Engineer

BNB Chain • Vancouver

Hybrid
CAD 120,000 - 200,000
Equity
Medical benefits
Manager of Site Reliability Engineering
Manager of Site Reliability Engineering

LayerZero Labs Ltd. • Vancouver

On-site
CAD 140,000 - 210,000
Site Reliability Engineer Vancouver, BC
Site Reliability Engineer Vancouver, BC

LayerZero • Vancouver

On-site
CAD 167,878 - 279,798
Full range of medical and financial benefits
Systems Engineer
Systems Engineer

BNB Chain • Vancouver

On-site
CAD 120,000 - 180,000
Systems Engineer
Systems Engineer

King River Capital Group • Vancouver

On-site
CAD 80,000 - 120,000
Backend Engineer
Backend Engineer

King River Capital Group • Vancouver

On-site
CAD 120,000 - 200,000
Equity options
Comprehensive medical benefits
Flexible working environment
Engineering Operations Lead
Engineering Operations Lead

BNB Chain • Vancouver

On-site
CAD 100,000 - 130,000