SUPPORT LEAD (Engineer) - 24x7 OPERATIONS SUPPORT

RemoteEngine Technologies Private Limited

Bengaluru

On-site

INR 3,000,000 - 4,800,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

RemoteEngine Technologies Private Limited is seeking a Support Lead to steer a 24x7 operations support team across a multi-application platform on Azure. You will personally resolve critical production issues while mentoring engineers and shaping shift coverage.

The role emphasizes hands-on problem solving, AI-assisted diagnostics, and strong coordination with engineering during incidents and deployments.

Qualifications

  • Experience leading 24x7 support or SRE teams.
  • Hands-on Azure expertise with AKS, App Gateway, private endpoints, Traffic Manager.
  • Strong SQL/no-SQL skills: PostgreSQL and Redis performance tuning.
  • Proficient in Azure DevOps, CI/CD pipelines, container registries, and production releases.
  • Familiar with React frontend and .NET/Python backend services.

Responsibilities

  • Build, staff, and manage a 24x7 rotational support schedule with backups and surge capacity.
  • Act as final technical escalation point and drive resolution of complex issues.
  • Own hiring, onboarding, and performance management for the support team.
  • Leverage AI tools for root-cause analysis, log analysis, and diagnostics.
  • Lead major incidents and post-incident RCA with preventive actions.
  • Develop runbooks, monitor coverage, and improve on-call processes.

Job description

# SUPPORT LEAD (Engineer) - 24x7 OPERATIONS SUPPORTRemoteEngine Partner·Bengaluru · Remotefull time5–6 yrsRequired skillsAzure Cloud Infrastructure (AKSApp GatewayAzure Services)Production Support & Incident ManagementAI-Assisted Troubleshooting & DebuggingAzure DevOps & CI/CD Pipeline ManagementPostgreSQL & Redis Performance OptimizationReact.NET & Python Application Support24x7 Support Team Leadership & Escalation ManagementSaaS Platform Operations & Root Cause AnalysisAbout the role## What you'll be working onWe are looking for a Support Lead to build, manage, and be the ultimate escalation point for a 24x7 operations support team covering our platform - comprising multiple customer-facing applications, internal operations tools, and third-party integration points - all running on Azure infrastructure.This is a hands-on leadership role: you will manage a team of engineers across shifts, but you are also expected to personally step in and solve the issue when the team is stuck. You are the technical backstop as much as the people manager.You own the design of the shift rotation - including active coverage, backup/surge capacity, leave management, and ensuring no single point of failure in coverage.We need someone with a proven history of personally solving hard, customer-specific production problems - the kind that don't have a documented fix - and who can lead a team to do the same. We also expect you to actively leverage AI-powered tools (code assistants, AI-assisted log analysis, automated diagnostics) as a force multiplier for faster diagnosis, root cause analysis, and resolution - and to embed that capability across your team.Responsibilities## What you'll do* Build, staff, and manage a 24x7 rotational support schedule with adequate backup for leaves, workload spikes, and shift rotations.* Ensure all engineers remain up to date with the platform and actively participate in rotational support.* Mentor and upskill the team on platform applications, Azure infrastructure, and CI/CD processes.* Act as the final technical escalation point and personally drive complex issues to resolution.* Own hiring, onboarding, and performance management for the support engineering team.* Leverage AI tools for root cause analysis, code-level debugging, incident pattern recognition, and solution validation.* Drive AI-assisted engineering practices to reduce resolution time and improve troubleshooting efficiency.* Identify opportunities to automate repetitive diagnostics and incident triage using AI-powered solutions.* Own the complete incident management lifecycle, including severity classification, SLA adherence, and coordination with engineering teams.* Lead major incident and outage bridges, making timely decisions during critical production issues.* Conduct post-incident root cause analysis and ensure implementation of preventive and corrective actions.* Communicate incident status, business impact, and resolution updates effectively to stakeholders.* Maintain strong technical oversight across Azure services, including AKS, App Gateway, PostgreSQL, Redis, Function Apps, Container Registry, Log Analytics, Key Vault, Private Endpoints, Storage, and Traffic Manager.* Oversee troubleshooting across React frontend, .NET backend, Python AI/ML services, and data ingestion pipelines.* Review and validate production releases and deployments to ensure stability and quality.* Establish and improve monitoring, alerting, and on-call processes to reduce detection and resolution times.* Develop and maintain operational runbooks while enabling the team to resolve new and complex issues beyond documented procedures.* Collaborate with Operations and Customer Success teams to streamline customer onboarding and eliminate recurring operational challenges.* Track and report key support metrics, including SLA adherence, incident volume, MTTR, and recurring issues, to leadership.Requirements## What we're looking for* Prior experience as a Forward Deployed Engineer, Technical Lead, or Support/SRE Lead with direct ownership of production issues.* Strong ability to leverage AI tools, including code assistants, AI-powered debugging, and log analysis, to accelerate production troubleshooting.* Hands-on expertise with Azure services, including AKS, App Gateway, Private Endpoints, Traffic Manager, and Azure resource management.* Strong troubleshooting experience with PostgreSQL and Redis, including performance tuning and scalability optimization.* Experience working with Azure Function Apps and data ingestion pipelines.* Solid background in Azure DevOps, CI/CD pipelines, container registries, and production release management.* Working knowledge of React, .NET, and Python services to effectively lead root cause investigations.* Experience leading or managing a 24x7 shift-based support or Site Reliability Engineering (SRE) team.* Proven ability to independently resolve complex, customer-specific, and production-critical issues.* Experience leading support or forward-deployed engineering teams for SaaS or platform-based products.* Exposure to healthcare or patient-facing technology platforms will be an added advantage.* Experience establishing or scaling a 24x7 support function from the ground up.* Demonstrated experience integrating AI-powered tools into engineering and support workflows.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff DevOps Engineer
Staff DevOps Engineer

Reuters • Hyderabad

Hybrid
INR 4,000,000 - 7,000,000
Hybrid work model
Azure Platform Site Reliability Engineer
Azure Platform Site Reliability Engineer

Foss United • India

On-site
INR 2,800,000 - 4,600,000
Sr Associate Support Engineer (AI/ML & Platform Operations)
Sr Associate Support Engineer (AI/ML & Platform Operations)

Workday, Inc. • Pune District

On-site
INR 1,200,000 - 1,800,000
Senior Azure Administrator / DevOps Engineer
Senior Azure Administrator / DevOps Engineer

Competent Groove Private Limited • Mohali

On-site
INR 3,500,000 - 6,000,000
Lead - SRE / Cloud Operations Support Engineer
Lead - SRE / Cloud Operations Support Engineer

Ninestars Information Technologies • Bengaluru

On-site
INR 1,400,000 - 2,200,000
SRE Lead
SRE Lead

Hdfc Securities • Mumbai

On-site
INR 3,500,000 - 5,500,000
Senior Full Stack Engineer (React + Node)
Senior Full Stack Engineer (React + Node)

Fractal Analytics Inc. • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Senior Staff Engineer Cloud
Senior Staff Engineer Cloud

The Tjx Companies • Hyderabad

On-site
INR 2,500,000 - 6,000,000
Systems Engineer
Systems Engineer

Jobgether SRL • India

Remote
INR 2,000,000 - 3,200,000
Fully remote position within India
Competitive benefits package
Professional development opportunities
Lead Site Reliability Engineer or Platform Engineer
Lead Site Reliability Engineer or Platform Engineer

Weekday 1 • Bengaluru

On-site
INR 5,000,000 - 10,000,000