Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

MongoDB

United States

Remote

USD 127,000 - 249,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

A cloud technology company is seeking an experienced software engineer to work on distributed storage systems. This role involves ensuring reliability and performance of multi-tenant storage services. Candidates should have over 6 years experience in software development, proficiency in Python or Go, and expertise in cloud infrastructure platforms like AWS or GCP. The position can be remote, but the candidate should be located in an Eastern or Central time zone. A competitive salary between $127,000 and $249,000 USD is offered.

Qualifications

  • 6+ years of experience in software development and distributed systems.
  • Customer-focused mindset with a preference for automation.
  • Expertise in cloud infrastructure platforms like AWS, GCP, or Azure.

Responsibilities

  • Work on multi-tenant distributed storage systems.
  • Build reliable and fault-tolerant services.
  • Participate in 24/7 on-call rotation for storage infrastructure issues.

Skills

Software development
Distributed systems
Proficiency in Python or Go
Stateful storage systems
Cloud infrastructure
Containerization technologies (Kubernetes)
Linux operating system internals
Networking concepts

Job description

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently.

You will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture.

This role can be based out of our Boston, New York City, Raleigh, Miami, Pittsburgh or remotely in the United States while physically based in an Eastern or Central time zone location.

The ideal candidate should
  • Have 6+ years of experience working on software development and operating distributed systems
  • Proficiency in Python, Go, or a similar language
  • Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs.
  • Possess a customer-focused mindset
  • Value efficiency in processes and operations
  • Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil
  • Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market
  • Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure
  • Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing)
Responsibilities
  • Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs
  • Build for reliability, making services and infrastructure available, resilient, fault-tolerant, and self-healing
  • Identify and configure key metrics to detect incidents and quantify service health, availability, and performance
  • Participate in a 24/7 on-call rotation to resolve issues involving the storage infrastructure
  • Become an expert in infrastructure performance, helping us optimize from the application level all the way to the kernel
Strong candidates may also have experience with
  • Leading major architectural shifts, such as moving from legacy storage stacks to new multi-tenant storage architectures, including planning and executing large-scale data and workload migrations with tight availability and durability requirements
  • Managing and scaling infrastructure across multi-cloud environments (AWS, GCP, or Azure)
  • Designing secure, multi-tenant runtime environments at scale
Equal Employment Opportunity

MongoDB, Inc. provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type and makes all hiring decisions without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

Accommodation for Disabilities

MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter.

Base Salary Range (U.S.)

$127,000—$249,000 USD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)
Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

MongoDB • Miami (FL)

On-site
USD 127,000 - 249,000
Flexible paid time off
Equity participation
401(k) plan
+2
Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)
Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

MongoDB • Pittsburgh

On-site
USD 127,000 - 249,000
Flexible paid time off
20 weeks fully-paid parental leave
401(k) plan
+2
Manager, Site Reliability Engineering - Storage Layer Service
Manager, Site Reliability Engineering - Storage Layer Service

MongoDB • New York (NY)

Hybrid
USD 157,000 - 270,000
Equity
Flexible paid time off
20 weeks fully-paid parental leave
+4
Team Lead, Site Reliability Engineering - Storage Layer Service
Team Lead, Site Reliability Engineering - Storage Layer Service

MongoDB • Boston (MA)

Hybrid
USD 151,000 - 297,000
Equity
Flexible paid time off
401(k) plan
+2
Site Reliability Engineer 3
Site Reliability Engineer 3

MongoDB • New York (NY)

Hybrid
USD 111,000 - 218,000
Equity
Flexible paid time off
20 weeks fully-paid gender-neutral parental leave
+2
Senior Software Engineer, Storage Layer Services
Senior Software Engineer, Storage Layer Services

MongoDB • New York (NY)

On-site
USD 126,000 - 248,000
Equity
Flexible paid time off
20 weeks fully-paid parental leave
+4
Senior Site Reliability Engineer, Fleet Management
Senior Site Reliability Engineer, Fleet Management

MongoDB • United States

Remote
USD 127,000 - 249,000
Flexible paid time off
20 weeks fully-paid parental leave
Fertility and adoption assistance
Site Reliability Engineer (Senior or Staff), Deployments
Site Reliability Engineer (Senior or Staff), Deployments

MongoDB • New Jersey

On-site
USD 127,000 - 249,000
Employee stock purchase program
Flexible paid time off
20 weeks fully-paid parental leave
+5
Senior Site Reliability Engineer, Fleet Management
Senior Site Reliability Engineer, Fleet Management

MongoDB • Austin (TX)

Hybrid
USD 127,000 - 249,000
Flexible paid time off
20 weeks fully-paid gender-neutral parental leave
Fertility and adoption assistance
Site Reliability Engineer (Senior or Staff), Deployments
Site Reliability Engineer (Senior or Staff), Deployments

MongoDB • Raleigh (NC)

On-site
USD 127,000 - 249,000
Flexible paid time off
20 weeks fully-paid parental leave
Fertility and adoption assistance
+3