Observability Technical Lead

Jeevan Technologies

Chennai District

On-site

INR 3,000,000 - 6,000,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jeevan Technologies is seeking an observability leader to drive monitoring, logging, tracing, and overall operational visibility across critical apps and infrastructure.

You will mentor a team, define standards, and advance AI-powered workflows with Copilot capabilities to improve reliability and efficiency across enterprise platforms and services.

Qualifications

  • 8+ years in infrastructure, operations, SRE, or observability
  • Proven people management of technical teams
  • Strong knowledge of major observability tools
  • Experience with Azure and Google Cloud environments
  • Experience with Generative AI in operations and Copilot workflows

Responsibilities

  • Lead and mentor a team of observability engineers
  • Define and execute the observability strategy, standards, and roadmap
  • Oversee monitoring, logging, alerting, tracing, and dashboarding
  • Drive service reliability, incident response readiness
  • Collaborate with application, infrastructure, cloud, and SRE teams to improve system health and performance
  • Establish KPIs, SLAs, and operational metrics
  • Manage hiring, performance development, resource planning, and stakeholder communications
  • Champion adoption of AI and Copilot-enabled workflows in Observability
  • Evaluate and implement AI-driven monitoring, alert correlation, and incident management capabilities
  • Define and execute AI strategy and intelligent automation within observability and operations ecosystem
  • Drive adoption of AI-powered operations including autonomous incident management and predictive analytics
  • Lead initiatives leveraging Copilot, Agentic AI, and AI agents to automate workflows
  • Partner with engineering to identify Agentic AI use cases to reduce manual effort
  • Build intelligent dashboards and automated remediation solutions

Skills

Observability leadership
Team leadership
SRE/DevOps experience
AI-powered observability
Cloud platforms (Azure,GCP)
Incident management
Automation
Dashboarding & telemetry
OpenTelemetry

Education

Bachelor's degree in Computer Science or Engineering

Tools

Splunk
Datadog
AppDynamics
Dynatrace
Grafana
Prometheus
OpenTelemetry

Job description

Job Summary

This role will lead a team responsible for monitoring, logging, tracing, and operational visibility across critical applications and infrastructure. The candidate will drive operational excellence, reliability, and continuous improvement while partnering closely with SREs, GCC and Security teams.

Responsibilities
  • Lead and mentor a team of observability engineers supporting enterprise platforms and services.
  • Define and execute the observability strategy, standards, and roadmap.
  • Oversee monitoring, logging, alerting, tracing, and dashboarding solutions.
  • Drive service reliability, incident response readiness, and operational excellence initiatives.
  • Collaborate with application, infrastructure, cloud, and SRE teams to improve system health and performance.
  • Establish KPIs, SLAs, and operational metrics to measure platform reliability and team effectiveness.
  • Manage hiring, performance development, resource planning, and stakeholder communications.
  • Ensure adoption of best practices for observability, automation, and proactive problem management.
  • Champion the adoption of AI and Copilot-enabled workflows within the Observability organization.
  • Evaluate and implement AI-driven monitoring, alert correlation, and incident management capabilities.
  • Define and execute the organizations strategy for AI, Agentic AI, and intelligent automation within the observability and operations ecosystem.
  • Drive the adoption of AI-powered operations, including autonomous incident management, intelligent alert correlation, predictive analytics, and self-healing platforms.
  • Lead initiatives leveraging Microsoft Copilot, Agentic AI frameworks, and AI agents to automate operational workflows, knowledge management, problem resolution, and service reliability improvements.
  • Partner with engineering teams to identify and prioritize use cases for Agentic AI that reduce manual effort and improve operational efficiency.
  • Partner with engineering and platform teams to build intelligent operational dashboards and automated remediation solutions.
Qualifications
  • Bachelors degree in Computer Science, Engineering, or a related field.
  • 8+ years of experience in infrastructure, operations, SRE, platform engineering, or observability domains.
  • 3+ years of people management experience leading technical teams.
  • Strong knowledge of observability platforms such as Splunk, Datadog, AppDynamics, Dynatrace, Grafana, Prometheus, OpenTelemetry, or similar tools.
  • Experience working in cloud environments (Azure GCP).
  • Experience leveraging Microsoft Copilot, Generative AI, and AI-powered observability capabilities to improve operational efficiency, incident response, and engineering productivity.
  • Knowledge of AI-assisted troubleshooting, anomaly detection, root cause analysis, and predictive monitoring solutions.
  • Excellent communication, stakeholder management, and leadership skills.
  • Preferred Experience leading globally distributed teams.
  • Strong background in automation, DevOps, and reliability engineering practices.
  • Familiarity with enterprise-scale monitoring and incident management processes.
  • Experience with Generative AI, Agentic AI, Microsoft Copilot, Azure AI, OpenAI technologies, or similar AI platforms.
Job Type

Full time

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Observability Technical Lead
Observability Technical Lead

Long Business Systems, Inc. • Chennai District

On-site
INR 4,000,000 - 7,000,000
SRE Observability Engineer
SRE Observability Engineer

Awign • Hyderabad

On-site
INR 4,200,000 - 6,500,000
AI Observability & Monitoring Engineer
AI Observability & Monitoring Engineer

Elabs Infotech • Bengaluru

On-site
INR 3,000,000 - 5,400,000
AI Observability Principal Architect(18+ Years)
AI Observability Principal Architect(18+ Years)

Everforth Apex Systems • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Staff Observability Engineer
Staff Observability Engineer

GE HealthCare • Bengaluru

On-site
INR 1,700,000 - 2,500,000
Staff Observability Engineer
Staff Observability Engineer

I00M05 Wipro GE Healthcare Private Limited • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Observability COE Lead- IT Consulting - Delhi NCR
Observability COE Lead- IT Consulting - Delhi NCR

Michael Page • Dadri

On-site
INR 4,000,000 - 7,000,000
AI Observability Principal Architect
AI Observability Principal Architect

LTM • Bengaluru

On-site
INR 4,500,000 - 7,500,000
AI Engineer
AI Engineer

Allegis Group • Hyderabad, Bengaluru

Hybrid
INR 1,200,000 - 2,500,000
Consultant_AI Observability
Consultant_AI Observability

HCLSoftware • Dadri

On-site
INR 1,200,000 - 1,800,000