AIOps Support Lead

BCE Global Tech - A Bell Canada Company

Bengaluru

Hybrid

INR 1,500,000 - 2,100,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health benefits
Flexible hours
Remote work options
Training opportunities

Job summary

BCE Global Tech - A Bell Canada Company in Bengaluru, India seeks an experienced Tier 1 AIOps Manager to lead a 14-person team across 50–320 applications. You will drive triage quality, standardize telemetry data, and push toward proactive, automated triage in a dynamic, multi-application landscape.

You'll collaborate with application teams, own data stewardship, and coordinate onboarding and stakeholder communication while supporting a hybrid on-site/remote work model.

Qualifications

  • 7+ years in application/production support or site reliability.
  • 2+ years directly managing a technical support team.
  • Hybrid environment experience across on-prem and cloud.
  • Hands-on with Dynatrace, New Relic, ManageEngine, and Open Telemetry.

Responsibilities

  • Manage Tier 1 team of 14 AIOps Support Engineers across 50–320 applications.
  • Set standards for triage quality, speed, and MTTR.
  • Drive data stewardship and standardize alarm telemetry.
  • Shift the team toward proactive, automated triage over time.
  • Coordinate onboarding of new apps into Tier 1 coverage.
  • Act as primary liaison for stakeholders on status and data quality.
  • Ensure coverage across required hours and report KPIs.

Skills

AIOps management
Team leadership
Observability tooling
ITIL processes
Linux troubleshooting
Windows troubleshooting
Networking basics
SQL querying
API troubleshooting

Tools

Dynatrace
New Relic
ManageEngine
Glass box
Open Telemetry
JSON
XML
Postman
SOAP UI
Kubernetes
OpenShift

Job description

Back At BCE Global Tech we are on a mission to modernize global connectivity, one connection at a time. We aim to build the highway to the future of communications, media and entertainment, determined to emerge as a powerhouse within the technology landscape in India team in Bengaluru. We bring ambitions to life through design thinking that bridges the gaps between people, devices and beyond, fostering unprecedented customer satisfaction through technology. Our core values support a customer-centric approach and the harnessing of cutting-edge technology to provide business outcomes with positive societal impact. Guided by innovation and a commitment to progress, we’re shaping a brighter future for the generations of today and tomorrow. If you would like to be a part of a team of thought-leaders pioneering advancements in 5G, MEC, IoT and cloud-native architecture, we’d love to hear from you.

What You'll Do
  • Manage the Tier 1 team: directly manage a team of 14 AIOps Support Engineers performing manual triage of alarms and alerts across a diverse, 50-to-320-application portfolio including hiring, coaching, scheduling, and performance management.
  • Own triage quality and speed: set and monitor standards for how quickly and accurately the team detects, classifies, and routes incidents, and drive continuous improvement in mean-time-to-triage.
  • Drive data stewardship: partner with application teams to standardize alarm and alert data across heterogeneous log aggregation tools (Dynatrace, New Relic, ManageEngine, Glass box, and others) into a clean, consistent telemetry backbone built on Open Telemetry.
  • Manage the reactive-to-proactive shift: reduce reliance on reactive, manual triage over time by improving alert quality, correlation, and early-warning signals laying the groundwork for future automated and Agentic triage.
  • Navigate a diverse, moving application landscape: support applications spanning different technology stacks and different architecture dispositions (Invest, Tolerate, Retire, Migrate), reprioritizing team focus as the portfolio shifts.
  • Coordinate onboarding of new apps: run a repeatable process for bringing new applications into Tier 1 coverage as the program scales from 50 to 320 applications, including support group and application owner mapping.
  • Manage stakeholders: act as the primary point of contact for support groups, application owners, and AIOps program leadership on Tier 1 status, incidents, and data-quality issues.
  • Manage shift/roster coverage: ensure the team of 14 provides consistent triage coverage across required hours as the application count grows.
  • Report on outcomes: track and report team KPIs triage time, alert-to-incident accuracy, false-positive rates, coverage growth to program leadership.
What We're Looking For
  • Experience: 7+ years in application/production support (L1/L1.5/L2) or site reliability, with 2+ years directly managing a technical support team.
  • Hybrid environment expertise: proven experience supporting applications across both on-premises and cloud environments, with exposure to modern microservices architectures.
  • Observability tooling: hands-on experience with monitoring and observability platforms such as Dynatrace, New Relic, AWS CloudWatch, ManageEngine, or Glass box; working knowledge of Open Telemetry and distributed tracing concepts.
  • ITIL discipline: strong grounding in incident, problem, and change management practices, with ServiceNow or Jira ticket management experience.
  • Technical range: comfortable with Linux and Windows troubleshooting, basic networking (TCP/IP, DNS, HTTP/HTTPS, SSL, load balancers), SQL/database query analysis, and API/integration troubleshooting.
  • People management: demonstrated ability to hire, coach, and retain a team of 10+ technical support staff through a period of significant scale-up (5x application coverage growth).
  • Analytical mindset: able to turn noisy, inconsistent alert data into clear, actionable insight, and to build repeatable frameworks rather than one-off fixes.
  • Comfort with ambiguity: willing to support a moving target a portfolio spanning Invest, Tolerate, Retire, and Migrate applications and to adapt priorities as the program evolves.
Must-Have Skills
  • Incident Management Lifecycle, and working knowledge of Problem, Change Request, and Service Request concepts (ITIL)
  • CMDB concepts and their use in incident and asset traceability
  • Hands-on experience with log aggregation technologies (Dynatrace, New Relic, ManageEngine, Glass box, or similar)
  • Working knowledge of JSON and XML, and basic file/task automation
  • Understanding of IT infrastructure and basic networking: VMs, firewalls, load balancers, containers, OpenShift (OCP), Kubernetes
  • Unix Shell scripting; Windows batch file creation
  • Basic cloud concepts: compute, storage, and security fundamentals
  • Security fundamentals: TLS, SSL, tokens, and secret management
  • Familiarity with API gateways and API testing toolkits (Postman, SOAP UI, or similar)
  • Outage response management and experience leading cross-functional coordination during major incidents
  • Ability to drive Root Cause Analyses (RCAs) and build reusable knowledge articles/runbooks
  • SLA/SLO management and reporting, including availability calculation
  • Working knowledge of data concepts: data latency, data fragmentation, data lineage, and data marts
Nice-to-Have Skills
  • Familiarity with AI concepts such as prompt engineering, knowledge graphs, and Retrieval-Augmented Generation (RAG)
  • Experience with cloud-native observability on AWS, Azure, or GCP
  • Exposure to Agentic AI or automation-driven triage tooling
Success Looks Like
  • A 14-person Tier 1 team that reliably triages alerts across 50+ applications with clear, standardized data.
  • A measurable, ongoing reduction in reactive manual triage as proactive detection improves.
  • A clean, well-governed Open Telemetry-based data backbone that the future Agentic automation layer can build on.

A repeatable onboarding process ready to scale coverage from 50 to 320 applications.

What We Offer
  • Competitive salaries and comprehensive health benefits
  • Flexible work hours and remote work options.
  • Professional development and training opportunities.
  • A supportive and inclusive work environment
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AIOps Support Lead
AIOps Support Lead

BCE Global Technology Centre • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Senior Application Support Operations Engineer
Senior Application Support Operations Engineer

Autonomize, Inc • Bengaluru

On-site
INR 1,800,000 - 2,800,000
Senior Application Support Operations Engineer
Senior Application Support Operations Engineer

Autonomize AI • Bengaluru

On-site
INR 3,500,000 - 6,000,000
IT Operations Lead
IT Operations Lead

MAQ Software • Dadri

On-site
INR 1,000,000 - 1,500,000
AIops Lead
AIops Lead

Pineswift Technologies • Gurugram District

On-site
INR 5,000,000 - 8,000,000
Senior Engineer II - Data Science / MLOps [T500-29625]
Senior Engineer II - Data Science / MLOps [T500-29625]

Marriott Tech Accelerator • Hyderabad

On-site
INR 4,000,000 - 6,000,000
Application Development Associate Director (Production Support)
Application Development Associate Director (Production Support)

Evernorth Health Services • Hyderabad

On-site
INR 3,500,000 - 6,000,000
Production and Support Engineer ( Java, Node.js)
Production and Support Engineer ( Java, Node.js)

Solugenix • Hyderabad

On-site
INR 900,000 - 1,500,000
Senior AI Engineer
Senior AI Engineer

Airowire Networks PVT LTD • India

On-site
INR 3,500,000 - 5,200,000
CSI Lead
CSI Lead

Opstree Global • Gurugram District

On-site
INR 4,500,000 - 7,500,000