Specialist - System Management

Vivantify

India

On-site

INR 1,300,000 - 2,100,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive salary and benefits
Professional growth opportunities
Collaborative work environment
Challenging projects with leading cll­

Job summary

Vivantify seeks a Senior Engineer / Specialist to support AI datacenter and sovereign compute infrastructure across PAN INDIA. The role operates at an L2 level, handling NVIDIA GPU servers, diagnostics, vendor coordination, and cloud operations support.

You will contribute to reliable infrastructure operations, SLA adherence, automation, and regulatory readiness. Work involves documenting processes, knowledge transfer, and enforcing change/incident management while collaborating with delivery

Qualifications

  • 5–7 years of relevant system management, infrastructure, or datacenter experience.
  • Experience supporting AI datacenter infrastructure and NVIDIA GPU servers.
  • Hands-on experience with hardware installation, diagnostics, and troubleshooting.
  • Experience working in L2 system or cloud operations support environments.
  • Strong vendor coordination and stakeholder collaboration skills.
  • Understanding of change and incident management processes.
  • Experience with documentation, knowledge transfer, and SLA-driven operations.
  • Ability to work within delivery teams to build, run, and scale infrastructure.

Responsibilities

  • Support NVIDIA GPU servers at the L2 system support level.
  • Perform GPU server hardware installation, diagnostics, and troubleshooting.
  • Coordinate with vendors for hardware-related issues and resolution.
  • Support Cloud Operations at the L2/L3 tower as required.
  • Build, operate, and scale sovereign compute and AI infrastructure.
  • Monitor and maintain SLA adherence across infrastructure operations.
  • Follow change and incident management processes.
  • Maintain technical documentation and conduct knowledge transfer.
  • Contribute to automation and continuous improvement initiatives.
  • Support compliance readiness related to DPDP, RBI, and CERT-In requirements.

Skills

NVIDIA GPU servers
L2 support
Cloud operations
Vendor coordination
Documentation
SLA adherence
Automation
Change management
GitHub Copilot
Grafana
Prompt Engineering

Tools

Grafana
GitHub Copilot
Prompt Engineering

Job description

Location: PAN INDIA

Experience: 5–7 Years
Employment Type: Permanent
Openings: 1

About the Role

We are seeking a Senior Engineer / Specialist to support AI Datacenter and sovereign compute infrastructure within a Joint GTM capability enhancement initiative. The role operates at an L2 support level and focuses on NVIDIA GPU server hardware, diagnostics, vendor coordination, and cloud operations support. The ideal candidate will contribute to reliable infrastructure operations, SLA adherence, automation, and continuous improvement.

Job Description

The role involves supporting AI datacenter infrastructure across the system frontline and Cloud Operations Support environment. You will manage NVIDIA GPU servers, including hardware installation, diagnostics, troubleshooting, and coordination with vendors.

Working closely with delivery teams and stakeholders, you will help build, operate, and scale the sovereign compute/AI stack while following established change and incident management processes. The role also includes documentation, knowledge transfer, automation, and regulatory compliance readiness.

Key Responsibilities
  • Support NVIDIA GPU servers at the L2 system support level.
  • Perform GPU server hardware installation, diagnostics, and troubleshooting.
  • Coordinate with vendors for hardware-related issues and resolution.
  • Support Cloud Operations at the L2/L3 tower as required.
  • Build, operate, and scale sovereign compute and AI infrastructure.
  • Monitor and maintain SLA adherence across infrastructure operations.
  • Follow change and incident management processes.
  • Maintain technical documentation and conduct knowledge transfer.
  • Contribute to automation and continuous improvement initiatives.
  • Support compliance readiness related to DPDP, RBI, and CERT-In requirements.
Required Skills
  • 5–7 years of relevant system management, infrastructure, or datacenter experience.
  • Experience supporting AI datacenter infrastructure and NVIDIA GPU servers.
  • Hands-on experience with hardware installation, diagnostics, and troubleshooting.
  • Experience working in L2 system or cloud operations support environments.
  • Strong vendor coordination and stakeholder collaboration skills.
  • Understanding of change and incident management processes.
  • Experience with documentation, knowledge transfer, and SLA-driven operations.
  • Ability to work within delivery teams to build, run, and scale infrastructure.
Mandatory Skills
  • GitHub Copilot
  • Grafana
  • Prompt Engineering
Candidate Profile

The ideal candidate should have strong experience in datacenter or system operations with exposure to NVIDIA GPU infrastructure and L2 support. You should be comfortable with hardware troubleshooting, vendor coordination, operational processes, documentation, and continuous improvement in an AI infrastructure environment.

Benefits Package
  • Competitive salary and benefits package
  • Opportunities for professional growth and development
  • A collaborative and inclusive work environment
  • The opportunity to work on exciting and challenging projects with leading clients
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

T&T | EAD | Senior Consultant/Manager | AI Infra GPU | PAN India
T&T | EAD | Senior Consultant/Manager | AI Infra GPU | PAN India

Deloitte & Touche GmbH Wirtschaftsprüfungsgesellschaft • Bengaluru

On-site
INR 3,000,000 - 5,500,000
Senior Staff Site Reliability Engineer
Senior Staff Site Reliability Engineer

NVIDIA Corporation • Pune District

On-site
INR 4,200,000 - 7,000,000
Systems Administrator, Hybrid Cloud and AI Infrastructure
Systems Administrator, Hybrid Cloud and AI Infrastructure

Lenovo • Bengaluru

Hybrid
INR 2,400,000 - 4,000,000
Senior Solutions Architect, Networking and Compute Infrastructure
Senior Solutions Architect, Networking and Compute Infrastructure

NVIDIA Gruppe • Gurugram District

On-site
INR 3,000,000 - 6,000,000
Senior Solutions Architect, Networking and Compute Infrastructure
Senior Solutions Architect, Networking and Compute Infrastructure

NVIDIA • Gurugram District

On-site
INR 3,500,000 - 7,500,000
NOC Technical Lead (L3)
NOC Technical Lead (L3)

Larsen & Toubro • Chennai District

On-site
INR 4,500,000 - 7,500,000
Manager, AI/HPC Infrastructure Technical Delivery — India
Manager, AI/HPC Infrastructure Technical Delivery — India

NVIDIA Gruppe • Pune District

On-site
INR 3,000,000 - 6,000,000
Manager, AI/HPC Infrastructure Technical Delivery — India
Manager, AI/HPC Infrastructure Technical Delivery — India

NVIDIA • Pune District

On-site
INR 4,500,000 - 7,500,000
Senior AI Compute Engineer
Senior AI Compute Engineer

Neysa • Mumbai

On-site
INR 900,000 - 1,500,000
Systems Administrator, Hybrid Cloud and AI Infrastructure
Systems Administrator, Hybrid Cloud and AI Infrastructure

Lenovo • Bengaluru Urban

On-site
INR 4,000,000 - 7,000,000