Systems Administrator, GPU/AI Infrastructure

Lenovo

Morrisville (NC)

On-site

USD 90,000 - 120,000

Full time

10 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Lenovo seeks a Systems Administrator to own end-to-end deployment, support, and administration of the Morrisville AI lab across network, compute, storage, OS, and container orchestration (Docker and Kubernetes) layers, plus the NVIDIA stack on 50+ servers. This is hands-on infrastructure work with rack-and-stack and smart hands.

You will be the primary Lab Operations resource, coordinating with the Network Engineer and Junior Systems Administrator, and providing secondary support for NVIDIA

Qualifications

  • 2+ years of systems administration breadth across network, compute, storage and OS, with GPU/AI focus.
  • Production Kubernetes administration with RBAC, cluster networking and upgrades.
  • Hands-on Docker container environments and NVIDIA driver/CUDA toolkit administration.
  • Familiarity with InfiniBand fabric in multi-node GPU environments.
  • Experience with Run:AI, ClearML or similar GPU orchestration/MLOps.
  • Comfort with rack-and-stack and lab hardware alongside higher-level admin duties.

Responsibilities

  • Deploy, support, and administer lab networking in coordination with the Network Engineer.
  • Deploy, support, and administer physical and virtual compute infrastructure.
  • Deploy, support, and administer lab storage across block and file tiers.
  • Install, patch, and administer operating systems in lab infrastructure.
  • Own production Kubernetes administration, including cluster ops and RBAC.
  • Administer NVIDIA drivers, CUDA toolkit, and related stack across GPUs.
  • Provide day-to-day lab support, rack-and-stack, and smart-hands for hardware.
  • Coordinate with Bangalore site for NVIDIA equipment and with adjacent labs.

Skills

Kubernetes Administration
Docker
NVIDIA administration
GPU orchestration
InfiniBand fabric
Hardware hands-on

Tools

Docker
Kubernetes
NVIDIA CUDA toolkit
InfiniBand fabric
Run:AI
ClearML

Job description

General Information

Req #

WD00103969

Career area:

Hardware Engineering

Country/Region:

United States of America

State:

North Carolina

City:

Morrisville

Date:

Wednesday, September 2, 2026

Working time:

Full-time

Additional Locations
  • United States of America - North Carolina - Morrisville
Why Work at Lenovo

We are Lenovo. We do what we say. We own what we do. We WOW our customers.

Lenovo is a US$83 billion revenue global technology powerhouse, ranked #153 in the Fortune Global 500, and serving millions of customers every day in 180 markets. Focused on a bold vision to deliver Smarter Technology for All, Lenovo has built on its success as the world's largest PC company with a full-stack portfolio of AI-enabled, AI-ready, and AI-optimized devices (PCs, workstations, smartphones, tablets), infrastructure (server, storage, edge, high performance computing and software defined infrastructure), software, solutions, and services. Lenovo's continued investment in world-changing innovation is building a more equitable, trustworthy, and smarter future for everyone, everywhere. Lenovo is listed on the Hong Kong stock exchange under Lenovo Group Limited (HKSE: 992) (ADR: LNVGY).

This transformation together with Lenovo's world-changing innovation is building a more inclusive, trustworthy, and smarter future for everyone, everywhere. To find out more visit www.lenovo.com, and read about the latest news via our StoryHub.

Description and Requirements
Position Overview

Lenovo seeks a Systems Administrator to own end-to-end deployment, support, and administration of the consolidated Morrisville AI lab, across the network, compute, storage, operating system, and container orchestration (Docker and Kubernetes) layers, plus the NVIDIA technology stack across 50-plus servers. This is a hands-on infrastructure role: the same person who designs and administers these layers also does the day-to-day admin, support, and engineering work in the lab, including rack-and-stack and smart hands.

You are the only dedicated support resource for the Lab Operations Manager, and you keep the lab's FY26/27 Net CapEx investment operational at the scale the AI Lab Consolidation requires. You will work alongside the lab's Network Engineer and Junior Systems Administrator roles, who own adjacent, more specialized slices of network engineering and routine rack-and-stack work respectively; you provide the full-stack administration that connects their work end to end. You will also provide secondary technical support for capital equipment hosted in Bangalore and validate lab configurations against NVIDIA reference architecture certification standards. The role sits within Lenovo's Hybrid Cloud and AI Infrastructure Services (HCAIS) Lab Operations organization.

Key Responsibilities
Infrastructure Deployment and Administration

Network: deploy, support, and administer lab networking, in coordination with the dedicated Network Engineer role for switch and fabric configuration.

Compute: deploy, support, and administer physical and virtual compute infrastructure across the lab.

Storage: deploy, support, and administer lab storage systems across block and file tiers.

Operating Systems: install, patch, and administer operating systems across lab infrastructure.

Container Orchestration

Docker: deploy and administer Docker container runtime environments.

Kubernetes: own full production Kubernetes administration, including cluster operations, role-based access control (RBAC), cluster networking, and version upgrades.

NVIDIA Platform Administration

NVIDIA Technology Stack: administer NVIDIA drivers, the CUDA toolkit, InfiniBand fabric, Run:AI orchestration, and ClearML across the lab's GPU infrastructure.

Configuration Validation: validate lab configurations against NVIDIA reference architecture (Lenovo Validated Design) certification requirements.

Lab Operations and Smart Hands

Day-to-Day Support: provide day-to-day administration, support, and engineering functions for the network, compute, and storage layers in the physical lab.

Rack and Stack: perform rack-and-stack, structured cabling, and smart-hands support for lab hardware alongside higher-level administration duties.

Cross-Site Support and Coordination

Bangalore Secondary Support: provide secondary technical support for NVIDIA technology stack capital equipment hosted in Bangalore.

Lab Coordination: coordinate with the adjacent ISG AI COE lab (Tech Marketing, customer proof-of-concepts) and the Morrisville Executive Briefing Center given shared physical proximity.

NVIDIA Relationship: maintain a direct working relationship with NVIDIA field contacts based in Morrisville.

Required Qualifications
Required Experience
  • 2+ years of systems administration experience with hands-on breadth across network, compute, storage, and operating system layers, plus a GPU/AI infrastructure specialization.
  • Production Kubernetes Administration: demonstrated experience with cluster operations, RBAC, cluster networking, and version upgrades in a production environment.
  • Docker: hands-on experience administering Docker container runtime environments.
  • Hands-on NVIDIA administration: demonstrated experience administering NVIDIA drivers and the CUDA toolkit.
  • InfiniBand Fabric: working familiarity with InfiniBand fabric in a multi-node GPU environment.
  • GPU Orchestration: experience with Run:AI, ClearML, or a comparable GPU orchestration and MLOps toolset.
  • Physical Infrastructure: comfortable with hands-on hardware work, including rack-and-stack, structured cabling, and smart hands, in addition to higher-level administration.
Core Professional Skills
  • Incident Response: able to own after-hours incident response for a shared, multi-team infrastructure environment.
  • Cross-Team Coordination: comfortable supporting multiple HC/AI teams across a shared lab environment with competing priorities, and coordinating day to day with the Network Engineer and Junior Systems Administrator roles on adjacent, overlapping infrastructure.
Preferred Qualifications
  • Kubernetes Certification: Certified Kubernetes Administrator (CKA) or equivalent.
  • Multi-Tenant Lab Experience: experience supporting a multi-tenant shared lab environment serving distributed teams.
  • Familiarity with NVIDIA certification and reference architecture (Lenovo Validated Design) processes.
  • Experience with DCIM, IPAM, or monitoring tooling comparable to the lab's stack (Hyperview-class DCIM, BlueCat / Infoblox-class IPAM, NVIDIA DCGM monitoring).

We are an Equal Opportunity Employer and do not discriminate against any employee or applicant for employment because of race, color, sex, age, religion, sexual orientation, gender identity, national origin, status as a veteran, and basis of disability or any federal, state, or local protected class.

Additional Locations
  • United States of America
  • United States of America - North Carolina
  • United States of America - North Carolina - Morrisville
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU/AI Infrastructure Admin – Kubernetes & NVIDIA
GPU/AI Infrastructure Admin – Kubernetes & NVIDIA

Lenovo • Morrisville (NC)

On-site
USD 90,000 - 120,000
Lab Operations Site Supervisor
Lab Operations Site Supervisor

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 60,000 - 102,000
Equity
Benefits
Linux Systems Administrator
Linux Systems Administrator

Relha LLC • Santa Clara (CA), Northern (KY)

Hybrid
USD 112,000 - 219,000
Equity
Benefits
AI Infra Staff Researcher
AI Infra Staff Researcher

Lenovo • North Carolina

Hybrid
USD 140,000 - 190,000
Linux Systems Administrator
Linux Systems Administrator

NVIDIA AI • Santa Clara (CA)

On-site
USD 112,000 - 178,000
Equity
Benefits
Cluster Solution Engineer
Cluster Solution Engineer

Lenovo • Morrisville (NC)

On-site
USD 120,000 - 170,000
Senior MEP Engineer – Datacenter Lab Infrastructure
Senior MEP Engineer – Datacenter Lab Infrastructure

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Senior Staff Site Reliability Operations Technical Lead
Senior Staff Site Reliability Operations Technical Lead

NVIDIA Corporation • Durham (NC)

On-site
USD 184,000 - 265,000
Senior Staff Site Reliability Operations
Senior Staff Site Reliability Operations

NVIDIA • Seattle (WA)

On-site
USD 184,000 - 265,000
Junior Rack Scale Engineer
Junior Rack Scale Engineer

Lenovo • North Carolina

Hybrid
USD 70,000 - 95,000