Staff Site Reliability Engineer, Fabric

MongoDB

Vancouver

On-site

CAD 144,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Flexible paid time off
20 weeks fully‐paid gender‐neutral parental leave
Health, dental, and vision benefits

Job summary

A leading technology company is seeking a Site Reliability Engineer (SRE) to join the Fabric team based in Vancouver or Toronto. The role focuses on building and maintaining robust infrastructure for secure communication within services. Candidates should have over 10 years of experience in distributed systems and networking, valuing automation. The position offers a competitive salary and various benefits including flexible paid time off and health insurance.

Qualifications

  • 10+ years of experience working on software and operating distributed systems.
  • Deep expertise in networking fundamentals and understanding of internet protocols.
  • Customer-focused mindset with a drive for process efficiency and automation.

Responsibilities

  • Develop a reliable and resilient multi‑cloud globally‐connected network.
  • Collaborate with service-owning teams on service‑to‑service connectivity.
  • Participate in a 24/7 on‑call rotation to resolve network issues.

Skills

Networking fundamentals
Distributed systems
Automation
Service mesh concepts
Cloud-based infrastructure

Job description

The Team

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems.

The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products.

This role can sit in our Toronto or Vancouver offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation.

Role Overview

We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable.

The ideal candidate should
  • Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles.
  • Possess a customer-focused mindset, driving improvements that benefit end‑users.
  • Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”).
  • Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azure, or GCP, e.g. VPCs, subnetting, routing, VPNs, peering, private link / private service connect, and CDNs.
  • Have a strong knowledge of service mesh and load‑balancing concepts, and be eager to implement these in a multi‑cloud environment.
Expectations
  • Participate in the development of a reliable and resilient multi‑cloud globally‑connected network that is crucial for MongoDB’s services.
  • Collaborate with service‑owning teams to provide internal support, addressing technical issues and offering guidance on best practices for service‑to‑service connectivity.
  • Participate in a 24/7 on‑call rotation to swiftly resolve issues related to network architecture and service‑to‑service connectivity, ensuring minimal disruption and high availability.

MongoDB is an equal opportunities employer. MongoDB is committed to providing any necessary accommodations for individuals with disabilities within the application and interview process. To request an accommodation due to a disability, please inform your recruiter.

MongoDB’s base salary range for this role in Canada is: $144,000—$200,000 CAD. Other benefits for eligible employees may include equity, participation in the employee stock purchase program, flexible paid time off, 20 weeks fully‑paid gender‑neutral parental leave, fertility and adoption assistance, Registered Retirement Savings Plan (RRSP) with employer match, mental health counseling, backup child and elder care, and health, dental, and vision benefits.

Req ID: 426185

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineering, Fabric (Mid, Senior, or Staff)
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)

MongoDB • Toronto

Hybrid
CAD 144,000 - 200,000
Equity
Flexible paid time off
Parental leave
+2
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)

MongoDB • Vancouver

Hybrid
CAD 144,000 - 200,000
Equity
Flexible paid time off
20 weeks fully-paid gender-neutral parental leave
+2
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)

AlleyCorp • Toronto

Hybrid
CAD 144,000 - 200,000
Hybrid work accommodation
Disability accommodations
Equal opportunities employer
Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)
Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

MongoDB • Toronto

On-site
CAD 144,000 - 200,000
Equity options
Flexible paid time off
20 weeks paid parental leave
+1
Site Reliability Engineer (Senior or Staff), Deployments
Site Reliability Engineer (Senior or Staff), Deployments

AlleyCorp • Toronto

On-site
CAD 144,000 - 200,000
Senior Technical Services Engineer MongoDB Montreal
Senior Technical Services Engineer MongoDB Montreal

Neura Market • Montreal (administrative region)

Hybrid
CAD 129,000 - 178,000
Software Engineer, Code Generation
Software Engineer, Code Generation

MongoDB • Calgary

On-site
CAD 108,000 - 149,000
Flexible paid time off
20 weeks fully-paid parental leave
Mental health counseling
Senior Engineering Manager, Developer Productivity MongoDB Alberta; British Columbia; Manitoba; Nova Scotia; Ontario; Quebec
Senior Engineering Manager, Developer Productivity MongoDB Alberta; British Columbia; Manitoba; Nova Scotia; Ontario; Quebec

Neura Market • Canada

Hybrid
CAD 191,000 - 265,000
Site Reliability Engineer
Site Reliability Engineer

Hunter Bond • Montreal (administrative region)

On-site
CAD 150,000 - 200,000
Flexible hours/work options
Cutting-edge technology investment
Start-up style environment
Senior Software Engineer (Server Security)
Senior Software Engineer (Server Security)

MongoDB • Toronto

On-site
CAD 137,000 - 189,000
Equity participation
Flexible paid time off
20 weeks fully-paid gender-neutral parental leave
+1