Senior Observability And AIOps Platform Engineer Senior Associate Director Software Engineering

HSBC

Pune District

On-site

INR 6,000,000 - 10,000,000

Full time

45 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

HSBC Software Development India is seeking a Senior Associate Director, Software Engineering to lead the design and implementation of a unified observability stack across metrics, logs, and traces. You will define instrumentation standards, build AI-enabled telemetry pipelines, and develop automated alerting tied to business impact.

You will partner with data scientists and SRE to reduce MTTD, improve signal quality, and deliver scalable, self-service instrumentation.

Qualifications

  • 10+ years overall experience with 5+ years specializing in observability, platform engineering, or SRE.
  • Experience with open-source observability stack: Prometheus, Grafana, OpenTelemetry, Loki, Jaeger; strong SQL/PromQL/LogQL proficiency.
  • Experience designing and operating large-scale log/metric/trace pipelines (Vector, Kafka or equivalents).
  • Strong Python and/or Go for automation; able to build data pipelines feeding ML/AI models from telemetry.
  • Proven ability to reduce alert noise and improve MTTR through signal quality improvements.
  • Understanding distributed systems failure modes and instrumentation for visibility.
  • Experience implementing data governance for telemetry: retention, access controls, PII, audit logging, compliance.
  • Ability to bridge cross-functional teams: translate data scientist needs and SRE runbooks into operational signal architecture.

Responsibilities

  • Design and build a unified observability stack aligned to OpenTelemetry standards: metrics, logs, traces, events with scalable ingestion, normalization, enrichment, and retention pipelines.
  • Define and enforce instrumentation standards across services; provide reusable SDKs, libraries, and auto-instrumentation patterns.
  • Build AIOps-ready telemetry data pipelines for AI/ML consumption: context, change events, topology, history, low-latency delivery.
  • Implement anomaly detection and forecasting models with data scientists to surface early warnings before SLO breaches.
  • Build correlation engines linking alerts across signals to provide contextual reasoning for AI agents.
  • Own SLO/SLI dashboards and error-budget tracking; map signals to business metrics with automated burn-rate alerting.
  • Improve alert quality: reduce false positives, add contextual grouping, controls, and map alerts to runbooks and owners.
  • Create self-service instrumentation guides, onboarding templates, and dashboard standards.

Skills

Observability
Platform engineering
SRE
Telemetry
Prometheus
Grafana
OpenTelemetry
Loki
Jaeger
Python
Go
PromQL
Kafka
Vector
Data pipelines

Tools

Prometheus
Grafana
OpenTelemetry
Loki
Jaeger
Vector
Kafka

Job description

Job Description:

Some careers shine brighter than others.

If youre looking for a career that will help you stand out, join HSBC and fulfil your potential. Whether you want a career that could take you to the top, or simply take you in an exciting new direction, HSBC offers opportunities, support and rewards that will take you further.

HSBC is one of the largest banking and financial services organisations in the world, with operations in 64 countries and territories. We aim to be where the growth is, enabling businesses to thrive and economies to prosper, and, ultimately, helping people to fulfil their hopes and realise their ambitions.

We are currently seeking an experienced professional to join our team in the role of Senior Associate Director, Software Engineering

In this role, you will:

  • Design and build a unified observability stack aligned to OpenTelemetry standards: metrics (Prometheus/OTel), logs (ELK/Loki), traces (Tempo/Jaeger), events with scalable ingestion, normalization, enrichment, and retention pipelines
  • Define and enforce instrumentation standards across services; provide reusable SDKs, libraries, and auto-instrumentation patterns that teams can adopt for self-service observability
  • Build AIOps-ready telemetry data pipelines purpose-built for AI/ML consumption: structured context, change events, topology relationships, historical baselines, and low-latency delivery
  • Implement anomaly detection and forecasting models in partnership with data scientists to surface early-warning signals before SLO breaches, measurably reducing time-to-detect (TTD)
  • Build correlation engines that link alerts across signals (latency errors deployments dependencies) to provide rich contextual reasoning for AI agents
  • Own SLO/SLI dashboards and error-budget tracking; link observability signals to business and customer-impact metrics with automated burn-rate alerting
  • Continuously improve alert quality: reduce false positives, add contextual grouping, implement cardinality controls, and map alerts to runbooks and owners
  • Create self-service instrumentation guides, onboarding templates, and dashboard standards that any engineer can follow without central bottlenecks

To be successful in this role, you should meet the following requirements:

  • 10+ years overall experience with 5+ years specializing in observability, platform engineering, or SRE with deep expertise in telemetry and monitoring infrastructure
  • Hands-on with the open-source observability stack: Prometheus, Grafana, OpenTelemetry, Loki, Jaeger, or enterprise equivalents; strong SQL/PromQL/LogQL proficiency
  • Experience designing and operating large-scale log/metric/trace pipelines: Vector, Kafka, or equivalent ingestors at production scale
  • Strong Python and/or Go for pipeline and tooling automation; ability to build data pipelines and feature stores that feed ML/AI models from operational telemetry
  • Demonstrable track record of reducing alert noise and improving MTTD through signal quality improvements in production environments
  • Understanding of distributed systems failure modes: latency, saturation, dependency fan-out, cascading failures, and how to instrument for visibility
  • Experience implementing data governance for telemetry: retention policies, access controls, PII redaction, audit logging, and compliance requirements
  • Proven ability to bridge cross-functional teams: translating data scientist requirements and SRE runbooks into real operational signal architecture

You’ll achieve more when you join HSBC.
www.hsbc.com/careers

HSBC is committed to building a culture where all employees are valued, respected and opinions count. We take pride in providing a workplace that fosters continuous professional development, flexible working and opportunities to grow within an inclusive and diverse environment. Personal data held by the Bank relating to employment applications will be used in accordance with our Privacy Statement, which is available on our website.

Issued by - HSBC Software Development India

Requirements:
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Observability & AIOps Platform Engineer /Senior Associate Director, Software Engineering
Senior Observability & AIOps Platform Engineer /Senior Associate Director, Software Engineering

HSBC • Hyderabad

Hybrid
INR 4,000,000 - 7,000,000
Microservices Java backend engineer /Associate Director, Software Engineering
Microservices Java backend engineer /Associate Director, Software Engineering

HSBC • Pune District

On-site
INR 2,000,000 - 3,000,000
Associate Director Software Engineering
Associate Director Software Engineering

HSBC • Bengaluru

On-site
INR 3,500,000 - 6,500,000
AI/ML Engineering Lead/(Gen AI,Python,LLM,ETL)Associate Director
AI/ML Engineering Lead/(Gen AI,Python,LLM,ETL)Associate Director

HSBC • Pune District

Hybrid
INR 3,000,000 - 6,000,000
L2/L3 Production Support Lead /Associate Director
L2/L3 Production Support Lead /Associate Director

HSBC • Pune District

On-site
INR 4,500,000 - 6,500,000
Engineering Lead,Microservices,Java Springboot /Senior Associate Director, Technology Management
Engineering Lead,Microservices,Java Springboot /Senior Associate Director, Technology Management

HSBC • Pune District

On-site
INR 3,500,000 - 6,000,000
Software Engineer -Observability
Software Engineer -Observability

1203 Barclays Global Serv. Cent • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Associate Director, Software Engineering
Associate Director, Software Engineering

HSBC • Hyderabad

On-site
INR 1,500,000 - 2,100,000
AI – Lead Service Intelligence Autonomous Controls/Senior Associate Director
AI – Lead Service Intelligence Autonomous Controls/Senior Associate Director

HSBC • Hyderabad

On-site
INR 2,500,000 - 6,000,000
Senior Technical Engineer / Associate Director, Solution Architecture
Senior Technical Engineer / Associate Director, Solution Architecture

HSBC • Maharashtra

On-site
INR 1,500,000 - 2,100,000