Senior Network Operations Center Engineer

Angel One

Bengaluru

On-site

INR 3,000,000 - 5,400,000

Full time

10 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health insurance
Wellness programs
Learning & development opportunities

Job summary

Angel One seeks an experienced Site Reliability Engineer 2 in Bengaluru to scale and optimize our hybrid cloud infrastructure. You will partner with engineering, data, and platform teams to ensure reliability, performance, and security across containerized services and data pipelines.

The role requires 4–6 years of infrastructure experience, strong Linux skills, and deep knowledge of AWS, observability, and automation. Join a fast-growing fintech with a culture that values velocity and impact.

Qualifications

  • 4-6 years of experience in infrastructure and systems delivering operational excellence.
  • Bachelor's degree in Computer Science or related field.
  • Deep expertise in Linux, observability platforms, AWS hybrid environments, container orchestration, and automation.
  • Strong incident and problem management experience in 24/7 production settings.
  • Experience with enterprise monitoring tools (Grafana, Prometheus, New Relic, Dynatrace).

Responsibilities

  • Drive reliability strategy to ensure uptime, latency targets, and system health across critical platforms.
  • Define SLIs, SLOs, and SLAs across multiple systems and services.
  • Lead major incident management and cross-team coordination for critical production issues.
  • Improve observability frameworks and monitoring strategies across services.
  • Participate in a 24×7 shift rotation, including nights, weekends, and holidays.
  • Mentor engineers and drive SRE best practices across teams.

Skills

Linux
Observability platforms
AWS
Container orchestration
Automation
Incident management

Education

Bachelor's degree in Computer Science or related field

Tools

Grafana
Prometheus
New Relic
Dynatrace

Job description

Angel One is one of India's fastest growing fin-techs, on a bold mission to make investing simple, smart, and inclusive for every Indian. With over 3+ crore clients we're building at scale - and building for impact.

Our Super App helps clients manage their investments, trade seamlessly, and access financial tools tailored to their goals. We are working to build personalized financial journeys for our clients, powered by new-age tech, AI, Machine Learning and Data Science.

We're a builder company at heart. You'll have the space to experiment, the freedom to move with velocity, and the mandate to make bold, user-first decisions - every single day.

The vibe? Think less hierarchy, more momentum. Everyone has a seat at the table and a shot to build something that lasts.

Be part of a team that’s scaling sustainably, thinking big, and building for the next billion.

Why You'll Love Working at Angel One!
  • Tech Systems that run at Scale: From AI to real-time data infra, you'll work on tech that’s ahead of the curve and solve problems that truly matter.
  • Build one of India's Leading Fintech Platform: We're not just disrupting finance - we're shaping how billion Indians access wealth.
  • Own It. Drive It. Scale It: You’ll have the freedom to lead, the resources to build, and the opportunity to leave your mark.
  • Empowered Growth: We invest in your growth and empower you to explore your full potential.
  • Exceptional Benefits: Our comprehensive benefits package includes health insurance, wellness programs, learning & development opportunities, and more.

Job Title: Site Reliability Engineer 2

We are seeking an experienced Site Reliability Engineer (SRE) with deep expertise in both AWS and on-premises environments to support, scale, and optimize our hybrid cloud infrastructure. In this role, you will partner closely with engineering, data, and platform teams to ensure the reliability, performance, and operational excellence of our containerized services, data pipelines, observability platforms, and security systems.

What you will do:

Drive reliability strategy to ensure service uptime, availability, latency targets, and overall system health across critical platforms.

Lead the definition and governance of SLIs, SLOs, and SLAs across multiple systems and services.

Lead major incident management and drive cross-team coordination for critical production issues.

Drive organization-wide RCA processes and implement systemic reliability improvements.

Drive strategic operational excellence programs to improve platform reliability, performance, and scalability.

Enhance observability frameworks and implement monitoring strategies using Grafana, Prometheus, CloudWatch, and log aggregation tools across services.

Participate in on-call rotation; troubleshoot incidents across the stack (network, compute, storage, data pipelines, applications).

Drive large-scale automation initiatives to eliminate manual processes and improve platform reliability.

Design platform-level tools and frameworks to improve reliability and developer productivity across teams.

Participate in a 24×7 shift rotation, including nights, weekends, and holidays.

Drive adoption of AI-driven operations, intelligent monitoring, and auto-healing capabilities.

Mentor engineers and drive SRE best practices across teams.

Enforce best practices for access management, network security, secrets management, patching, and vulnerability remediation.

Collaborate with security teams to ensure compliance with organizational and regulatory standards.

Who you are:

4-6 years of experience in an infrastructure and systems environment delivering operational excellence to highly complex distributed systems.

Bachelor's degree in Computer Science or a related field, or equivalent work experience.

Deep expertise in Linux, observability platforms, AWS hybrid environments, container orchestration, and automation frameworks in large-scale production environments.

Strong expertise in Incident Management & Problem Management, leading major incident resolution and driving long-term reliability improvements.

Extensive experience working in a 24/7 operations support environment and managing critical production systems.

Strong experience with enterprise monitoring and observability tools such as Grafana, Prometheus, New Relic, and Dynatrace.

Extensive experience working with hybrid environments (AWS and on-premises infrastructure).

AWS and CKA certifications and advanced cloud architecture knowledge are highly desirable.

Strong experience working with containerization and orchestration platforms.

Experience driving automation, infrastructure-as-code practices, and platform reliability improvements across engineering teams.

At Angel One, our thriving culture is rooted in Diversity, Equity, and Inclusion (DEI).

As an Equal opportunity employer, we wholeheartedly welcome people from all backgrounds irrespective of caste, religion, gender, marital status, sexuality, disability, class or age to be part of our team. We believe that everyone's unique experiences and viewpoints make us stronger together. Come and be a part of #OneSpace*, where your individuality is celebrated and embraced.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Business Strategy Specialist
Business Strategy Specialist

Angel One • Bengaluru Urban

On-site
INR 1,500,000 - 2,600,000
Health insurance
Wellness programs
Learning & Development opportunities
Senior SRE / DevOps Engineer
Senior SRE / DevOps Engineer

Code1 • Bengaluru

On-site
INR 1,500,000 - 2,000,000
SRE & DevOps Engineer
SRE & DevOps Engineer

Recrew AI • Bengaluru

On-site
INR 1,500,000 - 3,000,000
Senior Manager Analytics
Senior Manager Analytics

Angel One • Bengaluru Urban

On-site
INR 1,200,000 - 2,400,000
Health insurance
Wellness programs
Learning & development opportunities
+1
Site Reliability Engineer -2
Site Reliability Engineer -2

Groww • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Director of Engineering
Director of Engineering

Angel One • Bengaluru

On-site
INR 3,000,000 - 5,500,000
Analytics Manager
Analytics Manager

Angel One • Bengaluru Urban

On-site
INR 1,500,000 - 2,500,000
SRE-1
SRE-1

Keka Technologies Private Limited • Bengaluru

On-site
INR 1,200,000 - 2,200,000
Sr Engineering Manager - SRE
Sr Engineering Manager - SRE

OneAdvanced Limited • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities