Job Title: Nutanix Administrator
Hours of Operation: 24 7 rotational support (weekdays, nights, weekends, and holidays)
Experience: 1-4 Years
Job Summary:
The Nutanix Administrator is responsible for the end-to-end management of Nutanix-based HCI infrastructure across clients primary data centres. This includes server induction, monitoring, capacity/resource management, troubleshooting, incident handling, upgrades, and ensuring high system availability. The role also requires managing and mentoring a 247 operations team, coordinating with vendors, and delivering infrastructure projects end-to-end.
Key Responsibilities:
Nutanix HCI & AHV Administration
- Install, configure, maintain, and troubleshoot Nutanix HCI infrastructure including AOS, AHV, Prism Element, and Prism Central.
- Resolve issues related to hardware, software, configuration, performance, and operational failures across Nutanix environments.
- Manage Nutanix Prism Central for advanced operations, monitoring, RBAC, and policy management.
- Manage VM lifecycle: provisioning, cloning, templates, snapshots, protection domains, scheduling, and replication.
- Perform VM migrations including cross-cluster live migrations.
- Troubleshoot cluster health issues, networking, storage, redundancy factor, metadata, and performance anomalies.
- Manage AHV networking including vSwitches, Bridges, VLANs, and Uplink configurations.
- Implement and maintain Nutanix Data Protection, Replication Factor, and disaster recovery configurations.
- Perform LCM-based upgrades of AOS, AHV, BIOS, Firmware, HBA, and other hardware components.
Hardware & Infrastructure Management
- Hands-on experience with HPE, Dell, Cisco, and Lenovo hardware platforms.
- Perform firmware, BIOS upgrades using Nutanix LCM and OEM tools.
- Troubleshoot hardware alerts, disk failures, node reboots, memory errors, and cluster expansion tasks.
Tools, Automation, & Process Management
- Experience with Nutanix Move for migrations/replications.
- Create and run automation playbooks (where applicable).
- Work with ServiceNow for Incident, Change, and Problem management.
- Prepare root cause analysis (RCA) for major incidents.
- Maintain documentation of SOPs, upgrade procedures, and architecture diagrams.
- Operations & Team Management
- Participate in daily operational calls, major incident bridges, and weekly management meetings.
- Manage 247 operations team including task assignments, escalation handling, and skill development.
- Work with OEM vendors for support cases, hardware faults, and escalation coordination.
- Plan and execute cluster expansion, resource augmentation, and capacity forecasting.
Qualifications:
- Bachelors degree in computer science, Information Technology, or a related field.
- Strong technical expertise in Nutanix HCI, AHV virtualization, and enterprise infrastructure