Site Reliability Engineer II

NCR Voyix

Chennai District

On-site

INR 800,000 - 1,400,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

NCR Voyix is seeking a Site Reliability Engineer (II) to ensure the reliability, scalability, and performance of our cloud-native applications and infrastructure. You will design scalable systems, implement automation, monitor health with observability tools, and participate in on-call rotations to support production systems.

This role offers collaboration across teams to improve reliability and drive IaC initiatives, while mentoring junior engineers in a fast-paced environment at NCR Voyix.

Qualifications

  • 0-2 years of experience in a Site Reliability Engineering, DevOps, or similar role with a strong focus on cloud-native environments.
  • Proficient in at least one scripting language for automation.
  • Strong understanding of SRE principles, including SLIs, SLOs, and incident management.
  • Experience with cloud platforms and container orchestration.

Responsibilities

  • Design, implement, and maintain scalable, reliable, and performant cloud infrastructure and applications using SRE best practices.
  • Develop and implement automation tools and processes to streamline operational tasks, reduce manual toil, and improve system efficiency.
  • Monitor system health, performance, and availability using various observability tools, and proactively identify and troubleshoot issues.
  • Participate in on-call rotations to provide 24/7 support for critical production systems, diagnosing and resolving incidents quickly and effectively.
  • Collaborate with development teams to ensure new features and services are designed with reliability, scalability, and maintainability in mind.
  • Conduct post-incident reviews (RCAs) to identify root causes, implement preventative measures, and continuously learn from operational events.
  • Contribute to the development and maintenance of infrastructure as code (IaC) solutions for deploying and managing cloud resources.
  • Evaluate and recommend new technologies and tools to enhance our SRE capabilities and improve overall system reliability.
  • Document operational procedures, system configurations, and troubleshooting guides to facilitate knowledge sharing and improve team efficiency.
  • Mentor junior engineers and contribute to a culture of continuous learning and improvement within the SRE team.

Skills

SLIs/SLOs
Incident management
Communication skills

Education

Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience

Tools

Python
Go
Bash
AWS
Azure
GCP
Docker
Kubernetes
Jenkins
GitHub Actions
GitLab CI
Prometheus
Grafana
ELK
Datadog
Splunk
Terraform

Job description

Job Description
About NCR VOYIX
NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.
Job Summary

As a Site Reliability Engineer (II) at NCR VOYIX Corporation, you will play a critical role in ensuring the reliability, scalability, and performance of our cloud-native applications and infrastructure. You will leverage your expertise in SRE principles, automation, and cloud platforms to build robust systems, proactively identify and resolve issues, and continuously improve our operational excellence.

Job Responsibilities
  • Design, implement, and maintain scalable, reliable, and performant cloud infrastructure and applications using SRE best practices.
  • Develop and implement automation tools and processes to streamline operational tasks, reduce manual toil, and improve system efficiency.
  • Monitor system health, performance, and availability using various observability tools, and proactively identify and troubleshoot issues.
  • Participate in on-call rotations to provide 24/7 support for critical production systems, diagnosing and resolving incidents quickly and effectively.
  • Collaborate with development teams to ensure new features and services are designed with reliability, scalability, and maintainability in mind.
  • Conduct post-incident reviews (RCAs) to identify root causes, implement preventative measures, and continuously learn from operational events.
  • Contribute to the development and maintenance of infrastructure as code (IaC) solutions for deploying and managing cloud resources.
  • Evaluate and recommend new technologies and tools to enhance our SRE capabilities and improve overall system reliability.
  • Document operational procedures, system configurations, and troubleshooting guides to facilitate knowledge sharing and improve team efficiency.
  • Mentor junior engineers and contribute to a culture of continuous learning and improvement within the SRE team.
Job Qualifications
  • Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
  • 0-2 years of experience in a Site Reliability Engineering, DevOps, or similar role with a strong focus on cloud-native environments.
  • Proficient in at least one scripting language (e.g., Python, Go, Bash) for automation and tooling.
  • Strong experience with public cloud platforms (e.g., AWS, Azure, GCP), including understanding of core services (compute, networking, storage, databases).
  • Solid understanding of SRE principles, including SLIs, SLOs, error budgets, and incident management.
  • Experience with containerization technologies (e.g., Docker, Kubernetes) and orchestration.
  • Familiarity with CI/CD pipelines and tools (e.g., Jenkins, GitLab CI, GitHub Actions).
  • Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Datadog, Splunk).
  • Knowledge of network protocols, operating systems (Linux), and system administration.
  • Strong problem-solving skills and the ability to diagnose and resolve complex technical issues in a fast-paced environment.
  • Excellent communication and collaboration skills, with the ability to work effectively with cross-functional teams.

This job description outlines the primary duties and responsibilities of the role and is not intended to be exhaustive. The employee may be required to undertake additional duties and projects that are consistent with their skills, capabilities, and the overall purpose of the position

Offers of employment are conditional upon passage of screening criteria applicable to the job

EEO Statement

Integrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.

Statement to Third Party Agencies

To ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes

“When applying for a job, please make sure to only open emails that you will receive during your application process that come from a @ncrvoyix.comemail domain.”

Requirements
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (II)
Site Reliability Engineer (II)

NCR Corporation • Chennai District

On-site
INR 1,200,000 - 2,000,000
Site Reliability Engineer II (GCP & Azure)
Site Reliability Engineer II (GCP & Azure)

ncr • Chennai District

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer II GCP And Azure
Site Reliability Engineer II GCP And Azure

NCR Voyix • Chennai District

On-site
INR 900,000 - 1,300,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

NCR Corporation • Chennai District

On-site
INR 1,500,000 - 2,100,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

ncr • Chennai District

On-site
INR 900,000 - 1,350,000
Site Reliability Engineer (II)
Site Reliability Engineer (II)

NCR Voyix • Hyderabad

On-site
INR 2,500,000 - 6,000,000
Site Reliability Engineer II (GCP & Azure)
Site Reliability Engineer II (GCP & Azure)

NCR Corporation • Chennai District

On-site
INR 1,200,000 - 2,400,000
Infrastructure Or Platform Manager
Infrastructure Or Platform Manager

NCR Voyix • Chennai District

On-site
INR 4,000,000 - 7,000,000
Sr Site Reliability Engineer
Sr Site Reliability Engineer

NCR Voyix • Gurugram District

On-site
INR 1,200,000 - 1,800,000
SW Dev Ops Security Engineer III
SW Dev Ops Security Engineer III

3M HEALTHCARE • Chennai District

On-site
INR 1,800,000 - 3,000,000