AI Lab Technical Support Specialist

Veriipro

Milpitas (CA)

On-site

USD 80,000 - 100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Veriipro in Milpitas is looking for a hands-on technical support professional to maintain and troubleshoot high-performance computing (HPC) and AI infrastructure. This role involves working in a lab/data center environment, requiring strong Linux administration skills and experience with GPU hardware.

The ideal candidate will have 3-4+ years of relevant experience, capable of managing servers and troubleshooting networking issues. This position offers an opportunity to work closely with engineers to ensure optimal operations in a rapidly evolving tech landscape.

Qualifications

  • 3–4+ years of experience in data centers or similar roles.
  • Strong hands-on experience with GPU/CPU hardware.
  • Solid understanding of high-performance compute systems.

Responsibilities

  • Maintain and troubleshoot servers and AI lab hardware.
  • Use Linux command-line for system performance checks.
  • Collaborate with engineers for AI infrastructure operations.

Skills

Linux administration
GPU troubleshooting
Server hardware maintenance
Networking diagnostics
Scripting (Bash/Python)

Education

Bachelor’s degree in a related field or equivalent experience

Tools

Git
Jenkins

Job description

Role Overview

Hands-on technical support role focused on maintaining and troubleshooting high-performance compute (HPC) and GPU-based AI infrastructure in a lab/data center environment. This position supports Linux-based systems, hardware infrastructure, and compute clusters.

Roles and Responsibilities
  • Rack, stack, cable, and maintain servers and AI lab/data center hardware.
  • Troubleshoot and repair GPU/CPU servers, compute clusters, and networking (including InfiniBand).
  • Perform hardware diagnostics and replace/install server components in racks.
  • Monitor and troubleshoot Linux-based systems, storage, and networking issues.
  • Use Linux command-line tools for system validation, debugging, and performance checks.
  • Collaborate with engineers to ensure stable operation of AI infrastructure.
  • Manage lab space and support hardware inventory tracking and maintenance.
Required Skills & Experience
  • 3–4+ years of experience in data centers, NOC environments, or similar technical infrastructure roles.
  • Strong hands‑on experience with server, GPU, and compute hardware troubleshooting.
  • Solid understanding of computer architecture and high-performance compute systems.
  • Strong Linux administration and troubleshooting skills.
  • Ability to diagnose issues across hardware, OS, and networking stack.
  • Experience performing component‑level hardware installation and repair.
  • Physically capable of handling equipment (lifting up to 50 lbs / 23 kg, and performing extended physical tasks).
  • Basic scripting ability in Bash or Python.
  • Exposure to Git, Jenkins, or similar tools is a plus.
Preferred Qualifications
  • Experience with HPC or GPU compute clusters.
  • Familiarity with InfiniBand and advanced networking environments.
  • Hands‑on experience building or maintaining custom server systems or PC hardware.
  • Prior exposure to AI/ML infrastructure or research lab environments.
Education
  • Bachelor’s degree preferred; equivalent hands‑on experience accepted.
  • Practical experience is prioritized over formal education.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Lab HPC & GPU Infrastructure Technician
AI Lab HPC & GPU Infrastructure Technician

Veriipro • Milpitas (CA)

On-site
USD 80,000 - 100,000
AI Kernel / Cluster Engineer
AI Kernel / Cluster Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD 150,000 - 210,000
AI Operations & Infrastructure Engineer
AI Operations & Infrastructure Engineer

Invictus International • Geraghty Village (MD)

On-site
USD 100,000 - 130,000
AI Operations & Infrastructure Engineer
AI Operations & Infrastructure Engineer

Invictus International Consulting, LLC • Fort Meade (MD)

On-site
USD 100,000 - 130,000
Cluster Engineer
Cluster Engineer

STN Inc • San Francisco (CA)

On-site
USD 180,000 - 240,000
Data Center Technician - AI Infrastructure
Data Center Technician - AI Infrastructure

Hamilton Barnes Associates Limited • South Carolina

On-site
USD 70,000 - 85,000
Funded training and certifications
Stable full-time schedule: Monday–Friday, 9-5
Healthcare benefits
Senior Solutions Engineer, AI Infrastructure
Senior Solutions Engineer, AI Infrastructure

VAST Data • New York (NY)

On-site
USD 150,000 - 200,000
Lab Administrator
Lab Administrator

Ddn • Columbia (MD)

On-site
USD 90,000 - 120,000
Senior AI Solution Architect
Senior AI Solution Architect

Intel • Hillsboro (OR)

On-site
USD 171,000 - 315,000
HPC AI Systems Administrator
HPC AI Systems Administrator

MRE Consulting • Houston (TX)

On-site
USD 95,000 - 140,000