Observability Platform Manager

NCR Corporation

Chennai District

On-site

INR 3,500,000 - 5,200,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NCR Voyix is seeking an experienced Observability Platform Manager to lead strategy, implementation, and continuous evolution of enterprise observability across cloud and hybrid environments. You will define platform standards, roadmaps, and operating models; mentor engineers; and partner with SRE, platform, and cloud teams to ensure reliable, cost-aware observability that scales with the business.

Preferred experience includes OpenTelemetry, containerized workloads, Kubernetes, and tools such

Qualifications

  • 10+ years in observability or related disciplines.
  • Experience leading enterprise observability programs.
  • Hands-on with Splunk, Datadog, Dynatrace, Grafana, Prometheus, Elastic.
  • Strong cloud engineering experience (AWS/Azure/GCP).
  • Knowledge of OpenTelemetry & distributed tracing.

Responsibilities

  • Define and execute the vision and roadmap for the enterprise observability platform.
  • Lead cross-functional teams to build and operate observability capabilities.
  • Select and optimize tools like Splunk, Datadog, Grafana, Prometheus, Elastic.
  • Guide cloud engineering practices across AWS/Azure/GCP.
  • Establish SLIs/SLOs and enable self-service observability patterns.
  • Promote AI-enabled observability and automation.

Skills

Observability leadership
SRE
Platform engineering
Cloud engineering
OpenTelemetry
Kubernetes
CI/CD
Telemetry data management

Tools

Splunk
Datadog
Dynatrace
Grafana
Prometheus
Elasticsearch

Job description

About NCR VOYIX

NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.

Position Overview

We are seeking an experienced Observability Platform Manager to lead the strategy, implementation, and continuous evolution of enterprise observability capabilities across cloud and hybrid environments. The ideal candidate has recent hands-on experience implementing observability tools at scale, a strong foundation in cloud engineering, and the leadership skills to drive platform adoption, standardization, and operational excellence across engineering teams. Familiarity with AI applications and AI-enabled observability capabilities is a strong plus.

Key Responsibilities

Platform Strategy & Leadership: Define and execute the vision, roadmap, and operating model for the enterprise observability platform, including logs, metrics, traces, dashboards, and alerting. Lead and mentor engineers responsible for building, operating, and improving observability capabilities across business-critical platforms and applications. Establish platform standards, governance, onboarding patterns, and success measures to drive consistent adoption at scale. Partner with SRE, platform engineering, cloud engineering, infrastructure, and application teams to align observability strategy with reliability and business goals.

Observability Platform Implementation: Lead recent and large-scale implementations of observability tools and platforms across multi-team or enterprise environments. Evaluate, select, and optimize observability tooling such as Splunk, Datadog, Dynatrace, Grafana, Prometheus, Elasticsearch, or equivalent solutions based on scale, cost, and business needs. Drive implementation of telemetry standards, ingestion pipelines, access models, dashboard frameworks, and alerting practices that improve signal quality and reduce operational noise. Manage vendor relationships, platform lifecycle decisions, and cost/performance tradeoffs for observability capabilities.

Cloud Engineering & Reliability: Bring strong cloud engineering experience across AWS, Azure, or GCP, with an understanding of cloud-native architectures, resilience patterns, and scalable operations. Guide teams on instrumentation and monitoring for microservices, containers, Kubernetes, serverless workloads, and distributed systems using modern telemetry standards such as OpenTelemetry. Partner with engineering teams to embed observability into platform design, CI/CD pipelines, incident response, and service reliability practices.

Operational Excellence & Enablement: Establish service-level indicators, objectives, reporting, and operational health reviews to improve platform reliability and engineering outcomes. Develop enablement programs, best practices, and self-service patterns that help teams adopt observability consistently and effectively. Drive continuous improvement in incident detection, triage, root-cause analysis, and post-incident learning through better telemetry and platform workflows.

Automation & AI Applications: Promote automation-first practices for instrumentation, alert tuning, dashboard provisioning, and operational workflows. Identify opportunities to apply AI-enabled capabilities such as anomaly detection, event correlation, intelligent alerting, and operational insights. Familiarity with AI applications, AI platforms, or AI-supported engineering workflows is a plus.

Required Qualifications

10+ years of experience in observability, SRE, platform engineering, cloud engineering, or related disciplines, including recent experience implementing observability tools at scale. Proven experience leading or managing observability platforms, programs, or engineering teams in complex enterprise environments. Hands-on experience with enterprise observability and monitoring tools such as Splunk, Datadog, Dynatrace, Grafana, Prometheus, Elastic, or comparable platforms. Strong cloud engineering experience with AWS, Azure, or GCP, including cloud-native services, architecture patterns, and operational best practices. Knowledge of distributed systems, microservices, containers, Kubernetes, CI/CD, and infrastructure-as-code practices. Familiarity with OpenTelemetry, distributed tracing, service health models, and telemetry data management. Ability to define meaningful KPIs, SLIs/SLOs, alerting strategies, and executive-ready reporting that tie technical health to business outcomes. Strong collaboration and communication skills, with the ability to influence stakeholders across engineering, operations, and leadership teams.

Preferred Qualifications

Familiarity with AI applications, AI engineering workflows, or AI-enhanced observability capabilities. Knowledge of container ecosystems and orchestration platforms (Kubernetes, AKS/EKS/GKE). Experience working with event-driven architectures and microservices environments. Strong scripting or programming skills (Python, PowerShell, Bash, etc.). Relevant certifications (e.g., Splunk Architect, Dynatrace Professional, Cloud certifications).

Soft Skills

Excellent communication and stakeholder management skills. Ability to lead technical strategy and influence architectural decisions. Strong analytical, troubleshooting, and problem-solving abilities. Adaptability and curiosity about new technologies and evolving observability trends.

EEO Statement

Integrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.

Company Culture & Values

Help us run the world's top brands. At NCR Voyix, we specialize in turning routine transactions into meaningful connections. With a rich history of innovation, we have been at the forefront of problem-solving through technology. Operating globally in over 30 countries, we lead in Retail, Restaurant, Digital banking, and Payments. Our solutions optimize banking operations, streamline restaurant services, enhance retail interactions, and foster trust through secure payment systems. We take pride in our strong culture and a history of providing robust career paths. Come work for a leading technology company where you can grow your career. Join us and be part of revolutionizing transactions across these pivotal industries.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Observability Platform Manager
Observability Platform Manager

NCR Voyix • Chennai District

On-site
INR 4,000,000 - 7,000,000
Observability Platform Manager
Observability Platform Manager

NCR Voyix • Hyderabad

On-site
INR 2,000,000 - 5,000,000
null
Senior SRE – Unified Observability Engineer
Senior SRE – Unified Observability Engineer

NCR Voyix • Hyderabad

On-site
INR 2,000,000 - 3,200,000
Senior Site Reliability Engineer – Unified Observability
Senior Site Reliability Engineer – Unified Observability

NCR Voyix • Hyderabad

On-site
INR 3,500,000 - 7,000,000
ServiceNow Platform Developer
ServiceNow Platform Developer

NCR Corporation • Chennai District

On-site
INR 900,000 - 1,500,000
SW Dev Ops Security Engineer III
SW Dev Ops Security Engineer III

3M HEALTHCARE • Chennai District

On-site
INR 1,800,000 - 3,000,000
BI/Big Data Developer II
BI/Big Data Developer II

NCR Corporation • Chennai District

On-site
INR 1,500,000 - 2,800,000
HR Senior Project Manager
HR Senior Project Manager

NCR Corporation • Chennai District

On-site
INR 1,800,000 - 2,400,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

NCR Voyix • Gurugram District

On-site
INR 1,200,000 - 1,800,000
BI/Big Data Developer II
BI/Big Data Developer II

NCR Voyix • Chennai District

On-site
INR 900,000 - 1,400,000