Senior Site Reliability Engineer

United States Digital Space LLC

São Paulo

Híbrido

BRL 260 000 - 380 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Vantagens oferecidas por esta oferta de emprego

Competitive compensation including a 4
Equity where offered
Retirement plan
Comprehensive health benefits
Learning stipend and development

Resumo da oferta

Braze is seeking a Senior SRE for the MongoDB Platform to drive reliability, scalability, and observability across the company’s MongoDB deployments. You will own the infrastructure, automate operations, and partner with product engineering to optimize schemas, queries, and pipelines.

The role emphasizes an automation-first approach, on-call ownership, and collaboration with remote teams. Hybrid work ensures flexibility while delivering impact at scale.

Qualificações

  • 5+ years of experience as Software Engineer, DevOps Engineer, or SRE in a production environment.
  • Hands-on MongoDB expertise: replica sets, sharding, indexes, explain plans, and performance tuning under real load.
  • Strong Linux fundamentals and OS-level proficiency.
  • Programming skills in Python, Go, Ruby, or JavaScript with automation focus.
  • Experience with IaC tools: Terraform, Ansible, or equivalent.
  • Experience with Docker and Kubernetes.

Responsabilidades

  • Own MongoDB Reliability at Scale with enterprise SLAs and observability.
  • Improve Developer Experience with schema/index reviews and self-service tooling.
  • Build and automate infrastructure using Kubernetes, Terraform, and Ansible.
  • Lead incident response with on-call rotations and post-incident reviews.

Conhecimentos

MongoDB
Linux
Python/Go/ Ruby/ JavaScript
Terraform/Ansible
Docker/Kubernetes
SRE/On-Call
Automation

Ferramentas

Kubernetes
Terraform
Ansible

Descrição da oferta de emprego

At the company, we have found our people. We’re a genuinely approachable, exceptionally kind, and intensely passionate crew.

We seek to ignite that passion by setting high standards, championing teamwork, and creating work-life harmony as we collectively navigate rapid growth on a global scale while striving for greater equity and opportunity – inside and outside our organization.

To flourish here, you must be prepared to set a high bar for yourself and those around you. There is always a way to contribute: Acting with autonomy, having accountability and being open to new perspectives are essential to our continued success.

Our deep curiosity to learn and our eagerness to share diverse passions with others gives us balance and injects a one-of-a-kind vibrancy into our culture.

If you are driven to solve exhilarating challenges and have a bias toward action in the face of change, you will be empowered to make a real impact here, with a sharp and passionate team at your back. If the company sounds like a place where you can thrive, we can’t wait to meet you.

WHAT YOU'LL DO

the company runs one of the largest MongoDB deployments in the world - powering real-time customer engagement for thousands of the world’s leading brands. We process hundreds of billions of data points each month across more than 3.3 billion monthly active users, with MongoDB at the core of how we store, query, and serve that data at scale.

As a Senior SRE on the MongoDB Platform team, your primary mission is to make MongoDB better for the company - and to do so with the rigor, automation-first mindset, and engineering discipline of a world-class SRE. You won’t just keep the lights on; you’ll architect a more reliable, scalable, and observable MongoDB platform that the entire engineering organization depends on.

Main responsibilities:

1) Own MongoDB Reliability at Scale:

  • Design and operate the company’s MongoDB infrastructure to meet strict enterprise-grade SLAs, with deep ownership of availability, durability, and query performance
  • Build proactive monitoring and alerting that fires on symptoms - before customers feel impact – with rich MongoDB-specific observability (oplog lag, replication health, lock contention, index hit rates, etc)
  • Lead capacity planning and sharding strategy as data volumes and query patterns evolve
  • Drive root-cause analysis on MongoDB incidents and translate findings into permanent system improvements

2) Improve the MongoDB Developer Experience:

  • Partner with product engineering teams to review schema designs, index strategies, and aggregation pipelines - catching scalability anti-patterns before they reach production
  • Build self-service tooling, automation, and runbooks that let engineers interact with MongoDB safely and efficiently without needing to page the platform team
  • Define and enforce connection pool sizing, write-concern defaults, and read-preference standards across the fleet

3) Build and Automate Infrastructure:

  • Manage MongoDB cluster lifecycle (provisioning, upgrades, failovers, decommissions) on Kubernetes using the MongoDB Enterprise Kubernetes Operator, with infrastructure defined as code via Terraform and Ansible
  • Develop and maintain automated backup, restore, and point-in-time recovery workflows - tested regularly against real workloads
  • Contribute to internal platform tooling in Ruby and/or Go that reduces operational toil across the SRE organization

4) Incident Response & On-Call:

  • Participate in a PagerDuty on-call rotation with a clear charter: use every quiet shift to eliminate the next page
  • Lead incident retrospectives with a bias toward systemic fixes, automation, and documentation - not blame
  • Maintain and improve runbooks so that any engineer on the team can respond effectively to MongoDB incidents
WHO YOU ARE

Required:

  • 5+ years of experience as a Software Engineer, DevOps Engineer, or Site Reliability Engineer in a production environment
  • Hands-on MongoDB expertise: replica sets, sharding, index design, aggregation pipelines, explain plans, and performance tuning under real load
  • Strong Linux fundamentals and comfort operating at the OS level (disk I/O, memory, networking, process management)
  • Strong programming skills in one or more of: Python, Go, Ruby, or JavaScript – you write automation, not just scripts (JavaScript/Python experience is a plus for MongoDB shell scripting and aggregation pipeline work)
  • Experience with IaC tools: Terraform, Ansible, or equivalent
  • Experience with container orchestration: Docker and Kubernetes
  • A systems thinker who reasons about interfaces, failure modes, edge cases, and cascading effects across the stack
  • Bias toward documentation and asynchronous collaboration across global remote teams

Nice to have:

  • Experience running MongoDB at multi-terabyte scale or in a sharded topology
  • Familiarity with MongoDB Atlas, Ops Manager, or Cloud Manager
  • Experience with complementary data technologies in the company’s stack: Redis, Kafka, Postgres
  • Prior work on database platform engineering or database reliability engineering (DBRE) teams

#LI-Hybrid

**WHAT WE OFFER**

*the company benefits vary by location, and we encourage you to review our specific benefits offerings for each country here. More details on benefits plans will be provided if you receive an offer of employment.*

From offering comprehensive benefits to fostering hybrid ways of working, we’ve got you covered so you can prioritize work-life harmony. the company offers benefits such as:

  • Competitive compensation that may include equity
  • Retirement and Employee Stock Purchase Plans
  • Flexible paid time off
  • Comprehensive benefit plans covering medical, dental, vision, life, and disability
  • Family services that include fertility benefits and equal paid parental leave
  • Professional development supported by formal career pathing, learning platforms, and a yearly learning stipend
  • A curated in-office employee experience, designed to foster community, team connections, and innovation
  • Opportunities to give back to your community, including an annual company-wide Volunteer Week and donation matching
  • Employee Resource Groups that provide supportive communities within the company
  • Collaborative, transparent, and fun culture recognized as a Great Place to Work
**ABOUT the company

**the company is the leading customer engagement platform that empowers brands to Be Absolutely Engaging. the company helps brands deliver great customer experiences that drive value both for consumers and for their businesses. Built on a foundation of composable intelligence, BrazeAI allows marketers to combine and activate AI agents, models, and features at every touchpoint throughout the the company Customer Engagement Platform for smarter, faster, and more meaningful customer engagement. From cross-channel messaging and journey orchestration to Al-powered decisioning and optimization, the company enables companies to turn action into interaction through autonomous, 1:1 personalized experiences.

The company has been consistently recognized as a Leader in marketing technology by industry analysts, and was named a G2 “Best of Marketing and Digital Advertising Software Product” in 2026. the company was also named a 2026 Best Places to Work by Built In, a 2025 America’s Greenest Com

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Customer Success Manager II, Retail
Customer Success Manager II, Retail

United States Digital Space LLC • São Paulo

Híbrido
BRL 150 000 - 210 000
Competitive compensation
Equity
Retirement and ESPP
+5
Senior Software Engineer II, BrazeAI Operator
Senior Software Engineer II, BrazeAI Operator

United States Digital Space LLC • São Paulo

Híbrido
BRL 180 000 - 320 000
Senior Data Scientist (AI Deployment)
Senior Data Scientist (AI Deployment)

United States Digital Space LLC • São Paulo

Híbrido
BRL 180 000 - 320 000
Equity where applicable
Flexible paid time off
Medical/Dental/Vision coverage
+1
Senior Solutions Architect (Pre-Sales)
Senior Solutions Architect (Pre-Sales)

MongoDB • São Paulo

Híbrido
BRL 320 000 - 420 000
Senior Sales Data Analyst
Senior Sales Data Analyst

United States Digital Space LLC • São Paulo

Híbrido
BRL 150 000 - 210 000
Senior Marketing Data Analyst
Senior Marketing Data Analyst

United States Digital Space LLC • São Paulo

Presencial
BRL 120 000 - 240 000
Competitive compensation
Equity potential
Flexible paid time off
+2
Lifecycle Marketing Senior Specialist
Lifecycle Marketing Senior Specialist

United States Digital Space LLC • São Paulo

Híbrido
BRL 180 000 - 260 000
Competitive compensation that may add/
Equity options
Retirement plans
+8
Account Executive, Scale
Account Executive, Scale

United States Digital Space LLC • São Paulo

Híbrido
BRL 120 000 - 180 000
Flexible PTO
Competitive compensation
Equity
Senior Security Engineer, Enterprise Security
Senior Security Engineer, Enterprise Security

United States Digital Space LLC • São Paulo

Híbrido
BRL 180 000 - 300 000
Competitive compensation incl. equity
Retirement and ESPP
Flexible paid time off
+6
Engagement Manager II
Engagement Manager II

United States Digital Space LLC • São Paulo

Híbrido
BRL 624 000 - 936 000
Equity
Flexible PTO
Health benefits
+2