Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

MongoDB

Toronto

On-site

CAD 144,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity options
Flexible paid time off
20 weeks paid parental leave
Health benefits

Job summary

A tech company specializing in database solutions is seeking an experienced Site Reliability Engineer (SRE) to re-architect their cloud storage layer. This role involves enhancing distributed storage services and ensuring operational safety. The ideal candidate will have over 6 years of experience in software development, proficiency in Python or Go, and expertise in cloud platforms like AWS and Azure. The position can be remote within Canada, offering a competitive salary range of CAD 144,000 to CAD 200,000.

Qualifications

  • 6+ years of experience working on software development and operating distributed systems.
  • Proficiency in Python, Go, or a similar language.
  • Experience with stateful storage or database systems at scale.

Responsibilities

  • Work on multi-tenant distributed storage systems.
  • Build reliable services and infrastructure.
  • Identify and configure key metrics for service health.

Skills

Python
Go
Kubernetes
AWS
Google Cloud Platform
Azure
Linux internals
Networking concepts

Job description

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently.

You will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture.

This role can be based out of our Toronto or Montreal office or remotely in the Canada while physically based in an Eastern or Central time zone location.

The ideal candidate should
  • Have 6+ years of experience working on software development and operating distributed systems
  • Proficiency in Python, Go, or a similar language
  • Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs.
  • Possess a customer-focused mindset
  • Value efficiency in processes and operations
  • Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil
  • Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market
  • Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure
  • Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing)
Responsibilities
  • Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs
  • Build for reliability, making services and infrastructure available, resilient, fault-tolerant, and self-healing
  • Identify and configure key metrics to detect incidents and quantify service health, availability, and performance
  • Participate in a 24/7 on-call rotation to resolve issues involving the storage infrastructure
  • Become an expert in infrastructure performance, helping us optimize from the application level all the way to the kernel
Strong candidates may also have experience with
  • Leading major architectural shifts, such as moving from legacy storage stacks to new multi-tenant storage architectures, including planning and executing large-scale data and workload migrations with tight availability and durability requirements
  • Managing and scaling infrastructure across multi-cloud environments (AWS, GCP, or Azure)
  • Designing secure, multi-tenant runtime environments at scale

MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter.

MongoDB, Inc. provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type and makes all hiring decisions without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

MongoDB’s base salary range for this role is posted below. Compensation at the time of offer is unique to each candidate and based on a variety of factors such as skill set, experience, qualifications, and work location. Salary is one part of MongoDB’s total compensation and benefits package. Other benefits for eligible employees may include: equity, participation in the employee stock purchase program, flexible paid time off, 20 weeks fully-paid gender-neutral parental leave, fertility and adoption assistance, Registered Retirement Savings Plan (RRSP) with employer match, mental health counseling, backup child and elder care, and health, dental, and vision benefits offerings. Please note, the base salary range listed below and the benefits in this paragraph are only applicable to candidates based in Canada.

MongoDB’s base salary range for this role in Canada is:

$144,000—$200,000 CAD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (Senior or Staff), Deployments
Site Reliability Engineer (Senior or Staff), Deployments

AlleyCorp • Toronto

On-site
CAD 144,000 - 200,000
Senior Technical Services Engineer MongoDB Montreal
Senior Technical Services Engineer MongoDB Montreal

Neura Market • Montreal (administrative region)

Hybrid
CAD 129,000 - 178,000
Software Engineer, Code Generation
Software Engineer, Code Generation

MongoDB • Calgary

On-site
CAD 108,000 - 149,000
Flexible paid time off
20 weeks fully-paid parental leave
Mental health counseling
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)
Site Reliability Engineering, Fabric (Mid, Senior, or Staff)

AlleyCorp • Toronto

Hybrid
CAD 144,000 - 200,000
Hybrid work accommodation
Disability accommodations
Equal opportunities employer
Senior Engineering Manager, Developer Productivity MongoDB Alberta; British Columbia; Manitoba; Nova Scotia; Ontario; Quebec
Senior Engineering Manager, Developer Productivity MongoDB Alberta; British Columbia; Manitoba; Nova Scotia; Ontario; Quebec

Neura Market • Canada

Hybrid
CAD 191,000 - 265,000
Software Engineer 3, Atlas Search Systems
Software Engineer 3, Atlas Search Systems

MongoDB • Toronto

Hybrid
CAD 108,000 - 149,000
Parental leave policy
Fertility assistance
Employee affinity groups
Senior Customer Success Manager
Senior Customer Success Manager

MongoDB • Montreal (administrative region)

Hybrid
CAD 107,000 - 148,000
Equity
Employee stock purchase
Flexible PTO
+6
Enterprise Account Executive (Growth)
Enterprise Account Executive (Growth)

MongoDB • Toronto

On-site
CAD 170,000
Employee stock purchase plan
Generous parental leave
Comprehensive health benefits
+1
Software Engineer, Code Generation
Software Engineer, Code Generation

MongoDB • Calgary

On-site
CAD 108,000 - 149,000
Staff Product Manager, Identity and Security
Staff Product Manager, Identity and Security

MongoDB • Toronto

On-site
CAD 155,000 - 194,000