Senior SRE – Unified Observability Engineer

NCR Voyix

Hyderabad

On-site

INR 2,000,000 - 3,200,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

NCR Voyix seeks a Senior Site Reliability Engineer to design and operate a unified observability platform spanning Azure, Google Cloud, Kubernetes, and hybrid environments. You will establish telemetry standards and build dashboards to provide real-time visibility into infra, apps, and customer health, enabling proactive operations across our Retail, Restaurants, and Payments domains.

The role emphasizes SLIs/SLOs, incident response automation, and AI-driven observability to continuously improve

Qualifications

  • Senior Site Reliability Engineer with hands-on experience in designing and operating unified observability platforms.
  • Experience with Kubernetes (AKS/GKE) across cloud and on-prem environments.
  • Proficiency in cloud platforms (Azure and Google Cloud) and telemetry standards.

Responsibilities

  • Design and implement enterprise observability solutions across Azure, GCP, Kubernetes, and hybrid environments.
  • Develop monitoring, logging, tracing, and telemetry standards using industry best practices.
  • Build enterprise dashboards that provide real-time infrastructure, application, and customer health visibility.
  • Define and implement SLIs, SLOs, error budgets, and operational health metrics.
  • Improve proactive detection, alert quality, and incident response through automation and intelligent alerting.
  • Integrate observability with ServiceNow, automation platforms, and operational workflows.
  • Partner with Product Engineering, Infrastructure, Security, and Operations teams to improve platform reliability and operational readiness.
  • Support enterprise initiatives involving AI-driven observability, event correlation, and operational analytics.

Skills

SRE
Kubernetes AKS/GKE
Azure
GCP
Grafana
Dynatrace
Prometheus
OpenTelemetry
Monitoring & Telemetry
Terraform
Python/Go/PowerShell
CI/CD & DevOps

Tools

Grafana
Dynatrace
Prometheus
OpenTelemetry
Terraform

Job description

About NCR VOYIX

NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.

About NCR VOYIX

NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.

We are seeking a Senior Site Reliability Engineer (Unified Observability) to support the F1 Next Generation Customer Unified Observability initiative. This role will build and operate a unified observability platform that provides end-to-end visibility across NCR Voyix Restaurants, Retail, and Payments environments. The engineer will help establish a single operational view spanning infrastructure, applications, Kubernetes platforms, cloud services, customer experience, and business transactions to enable proactive operations and improve service reliability.

Key Responsibilities
  • Design and implement enterprise observability solutions across Azure, GCP, Kubernetes, and hybrid environments.
  • Develop monitoring, logging, tracing, and telemetry standards using industry best practices.
  • Build enterprise dashboards that provide real-time infrastructure, application, and customer health visibility.
  • Define and implement SLIs, SLOs, error budgets, and operational health metrics.
  • Improve proactive detection, alert quality, and incident response through automation and intelligent alerting.
  • Integrate observability with ServiceNow, automation platforms, and operational workflows.
  • Partner with Product Engineering, Infrastructure, Security, and Operations teams to improve platform reliability and operational readiness.
  • Support enterprise initiatives involving AI-driven observability, event correlation, and operational analytics.
Required Skills
  • Site Reliability Engineering (SRE)
  • Kubernetes (AKS/GKE)
  • Azure and Google Cloud Platform
  • Grafana, Dynatrace, Prometheus, OpenTelemetry (or similar)
  • Monitoring, logging, distributed tracing, and telemetry
  • Infrastructure as Code (Terraform)
  • Python, Go, or PowerShell automation
  • CI/CD and DevOps practices

Offers of employment are conditional upon passage of screening criteria applicable to the job

EEO Statement

Integrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.

Statement to Third Party Agencies

To ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer – Unified Observability
Senior Site Reliability Engineer – Unified Observability

NCR Voyix • Hyderabad

On-site
INR 3,500,000 - 7,000,000
Site Reliability Engineer Shift Manager
Site Reliability Engineer Shift Manager

NCR Voyix • Chennai District

Hybrid
INR 1,500,000 - 2,500,000
Observability Platform Manager
Observability Platform Manager

NCR Voyix • Chennai District

On-site
INR 4,000,000 - 7,000,000
Observability Platform Manager
Observability Platform Manager

NCR Voyix • Hyderabad

On-site
INR 2,000,000 - 5,000,000
null
Observability Platform Manager
Observability Platform Manager

NCR Corporation • Chennai District

On-site
INR 3,500,000 - 5,200,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

NCR Corporation • Chennai District

On-site
INR 3,500,000 - 5,200,000
SW Dev Ops Security Engineer III
SW Dev Ops Security Engineer III

3M HEALTHCARE • Chennai District

On-site
INR 1,800,000 - 3,000,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

NCR Voyix • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Infrastructure Operations Specialist (I)
Infrastructure Operations Specialist (I)

NCR Corporation • Gurugram District

On-site
INR 500,000 - 700,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

NCR Corporation • Gurgaon

On-site
INR 900,000 - 1,300,000