Senior Site Reliability Engineer

MoEngage

Bengaluru

On-site

INR 1,000,000 - 1,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading customer engagement platform in Bengaluru seeks a Site Reliability Engineer (SRE-2) to enhance performance and reliability of critical services. This role calls for 3-5 years of SRE or DevOps experience, strong skills in AWS/GCP and Kubernetes, and a passion for automation. Responsibilities include optimizing cloud infrastructure, mentoring team members, and leading incident response. Join a dynamic team that drives impactful reliability initiatives to support global consumer brands.

Qualifications

  • 3-5 years of hands-on experience in Site Reliability Engineering or similar.
  • Proven ability to automate complex tasks.
  • Strong command of AWS and/or GCP cloud platforms.
  • Excellent communication and problem-solving skills.

Responsibilities

  • Take ownership of the reliability and performance of critical services.
  • Design and implement automation solutions.
  • Lead troubleshooting for complex production incidents.
  • Contribute to cloud infrastructure design and optimization.
  • Implement advanced monitoring and logging solutions.
  • Collaborate with development teams for reliability improvements.
  • Mentor and guide junior engineers.

Skills

Python
Go
Kubernetes
AWS
GCP
Containerization
Infrastructure as code (Terraform, Ansible)
Monitoring and observability (Prometheus, Grafana)
Cloud Security principles
Networking concepts
Distributed systems (Celery, Kafka)

Job description

About the Company

MoEngage is an insights‑led customer engagement platform trusted by 1,350+ global consumer brands, including McAfee, Flipkart, Domino’s, Nestle, Deutsche Telekom, and OYO. MoEngage combines data from multiple sources to help brands gain a 360‑degree view of their customers.

MoEngage Analytics arms marketers and product owners with insights into customer behavior. Brands can leverage MoEngage Personalize to orchestrate journeys and build 1:1 conversations across the website, mobile, email, social, and messaging channels. MoEngage Inform, the transactional messaging infrastructure, helps unify promotional and transactional communication to a single platform for better insights and lower costs. MoEngage’s AI Suite helps marketers develop winning copies and creatives, optimize campaigns and channels that boost engagement, and help with faster execution.

For over a decade, consumer brands in 60+ countries have been using MoEngage to power digital experiences for over a billion monthly customers. With offices in 15 countries, MoEngage is backed by Goldman Sachs Asset Management, B Capital, Steadview Capital, Multiples Private Equity, Eight Roads, F‑Prime Capital, Matrix Partners, Ventureast, and Helion Ventures.

MoEngage was named a Contender in The Forrester Wave™: Real‑Time Interaction Management, Q1 2024 report, and Strong Performer in The Forrester Wave™ 2023 report. MoEngage was also featured as a Leader in the IDC MarketScape: Worldwide Omni‑Channel Marketing Platforms for B2C Enterprises 2023.

About Role

The Opportunity: Site Reliability Engineer (SRE‑2)

Are you an SRE with a few years under your belt, itching to take on more significant challenges and drive impactful reliability initiatives? Do you have a solid grasp of cloud platforms and container orchestration, and a burning desire to automate everything in sight? As an SRE‑2 at MoEngage, you’ll be a critical member of our SRE team, responsible for the health and performance of key services and contributing directly to the evolution of our infrastructure at a scale that few engineers get to experience. This is your chance to deepen your technical expertise, take on more ownership, and mentor emerging talent while working on a platform that operates at the cutting edge.

Roles and Responsibilities
  • Be a Reliability Champion: Take ownership of the reliability, performance, and efficiency of critical services.
  • Automate, Automate, Automate: Design, develop, and implement robust automation solutions to eliminate toil, streamline operations, and improve system resilience.
  • Battle Incidents (and Win): Lead troubleshooting efforts for complex production incidents, perform in‑depth root cause analysis, and implement sustainable preventative measures.
  • Sculpt Our Infrastructure: Actively contribute to the design, implementation, and optimization of our cloud infrastructure on AWS and GCP, leveraging your expertise in technologies like Kubernetes.
  • Enhance Observability: Implement and refine advanced monitoring, alerting, and logging solutions to gain deep insights into system behavior and predict potential issues.
  • Collaborate for Success: Partner closely with development teams to influence architectural decisions, ensuring reliability, scalability, and security are built in from the start.
  • Strengthen Our Security Posture: Implement and advocate for advanced security practices within our infrastructure and operational workflows.
  • Drive Efficiency: Analyze and optimize cloud infrastructure spend, identifying and implementing cost‑saving opportunities.
  • Guide the Next Wave: Mentor and guide SRE‑1 engineers, contributing to the growth and knowledge sharing within the team.
  • Be Ready for Action: Participate in our on‑call rotation, acting as a key point of escalation and resolution for critical issues.
Requirements
  • 3‑5 years of hands‑on experience in Site Reliability Engineering, DevOps, or a similar role with a strong focus on production systems.
  • Demonstrated expertise in Python or Go – you have a proven track record of automating complex tasks.
  • Strong command of AWS and/or GCP cloud platforms.
  • In‑depth experience with containerization and orchestration using Kubernetes (K8s, ArgoCD, Helm/Kustomize).
  • Experience with infrastructure as code tools like Terraform or Ansible is highly valued.
  • Solid understanding and experience with monitoring and observability stacks (VictoriaMetrics, Prometheus, Grafana, ELK stack, etc.).
  • Deep knowledge of Linux/Unix systems internals and advanced networking concepts.
  • Proven ability to diagnose and resolve complex issues in large‑scale distributed systems.
  • A strong understanding of Cloud Security and Information Security principles and best practices.
  • Experience with cloud cost analysis and optimization techniques.
  • Familiarity with CI/CD pipelines and GitOps methodologies.
  • Experience with messaging queues and distributed systems (Celery, Kafka) is a plus.
  • Excellent communication, collaboration, and problem‑solving skills.
  • A desire to mentor and lead by example.
Equal Opportunity Statement
  • Employment at MoEngage is based solely on professional competence, skills, and experience. We stand firmly against all forms of discrimination and support equal rights and opportunities regardless of gender, ethnicity, abilities, age, identity, orientation or expression, marital status (including pregnancy), religion and beliefs, or any other status protected by law.
  • It is our policy to comply with all applicable national, state, and local laws related to non‑discrimination and equal opportunity. MoEngage is truly a place where everyone can bring their passions, authentic selves, and talents to work, collaborating to drive progress and solve meaningful challenges.
Why Join Us!
  • At MoEngage, we are passionate about our team and technology - see below to know more about us.
  • Life@MoEngage
  • Tech@MoEngage
  • We handle more than a billion messages every day. Rest assured, you will be surrounded by really smart and passionate people as we scale much more to build a world‑class technology team.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - II
Site Reliability Engineer - II

MoEngage • Bengaluru

On-site
INR 2,500,000 - 4,500,000
Lead Software Engineer - DPM
Lead Software Engineer - DPM

MoEngage Inc. • India

On-site
INR 6,000,000 - 9,000,000
Lead Software Engineer - DPM
Lead Software Engineer - DPM

MoEngage Inc. • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Lead Software Engineer - DPM
Lead Software Engineer - DPM

F-Prime Capital • Bengaluru

On-site
INR 4,200,000 - 6,500,000
Equal opportunity employer
Senior Customer Success Manager
Senior Customer Success Manager

F-Prime Capital • Bengaluru

On-site
INR 1,500,000 - 2,300,000
Work at Scale and challenge yourself
Work with a smart team that grew up in
Work on an award‑winning product, tech
+1
Technical Lead
Technical Lead

MoEngage • Bengaluru

On-site
INR 2,500,000 - 3,500,000
Support Engineer
Support Engineer

F-Prime Capital • Bengaluru

On-site
INR 420,000 - 700,000
Senior Customer Success Manager
Senior Customer Success Manager

MoEngage Inc. • Bengaluru

On-site
INR 1,500,000 - 2,300,000
Work at Scale and challenge yourself
Work with a smart team that grew up in
Work on an award-winning product, tech
+1
Support Engineer
Support Engineer

MoEngage Inc. • Bengaluru

On-site
INR 800,000 - 1,200,000
Principal Architect
Principal Architect

MoEngage • Bengaluru

On-site
INR 4,000,000 - 6,000,000