Remote Infrastructure Support Engineer — GPU/AI Infra

Nscale

San Francisco (CA)

On-site

USD 100,000 - 140,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Remote-first
Flexible work
Annual salary reviews

Job summary

Nscale seeks an Infrastructure Support Engineer to maintain GPU fleets, handle tickets, and drive service reliability across data-centre operations, Linux, and data networks. You’ll own end-to-end tasks, triage hardware, interpret logs, and contribute to runbooks and automation with the team.

Ideal candidates have 3–4+ years in infrastructure support, strong GPU hardware troubleshooting, Bash/Python proficiency, and a growth mindset.

Qualifications

  • 3+ years in infrastructure support or support engineering with customer-facing experience.
  • Hands-on GPU infrastructure knowledge and hardware troubleshooting skills.
  • Solid Linux CLI skills and basic scripting capabilities.
  • Experience with ITSM processes, ticketing discipline, and clear documentation.
  • Willingness to travel for deployments and training as needed.

Responsibilities

  • Join support duty rotation and handle tickets and alerts with proper handovers.
  • Perform GPU node triage, hardware troubleshooting, and log interpretation.
  • Run diagnostics on fabric/links, capture evidence, and prepare vendor handovers.
  • Assist with storage and data-path investigations on high-performance platforms.
  • Document validated steps, contribute to runbooks, and share knowledge.
  • Participate in on-call, learn Platform fundamentals, and travel when required.

Skills

GPU infrastructure
Linux
Networking
Ticketing/ITSM
Scripting basics
Platform/DC fundamentals
Growth mindset
Adaptability

Tools

nvidia-smi
NetBox
Bash
Python
Git
OpenStack
MAAS

Job description

Nscale seeks an Infrastructure Support Engineer to maintain GPU fleets, handle tickets, and drive service reliability across data-centre operations, Linux, and data networks. You’ll own end-to-end tasks, triage hardware, interpret logs, and contribute to runbooks and automation with the team.

Ideal candidates have 3–4+ years in infrastructure support, strong GPU hardware troubleshooting, Bash/Python proficiency, and a growth mindset.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior GPU Infrastructure Support Engineer
Senior GPU Infrastructure Support Engineer

Nscale • San Francisco (CA)

On-site
USD 120,000 - 170,000
Equity
Remote-friendly team
Flexible workplace
Senior GPU Infra Engineer — Remote
Senior GPU Infra Engineer — Remote

Nscale • Seattle (WA)

On-site
USD 120,000 - 170,000
Remote-first culture
Equity plan
Flexible workplace
Senior Data Center Engineer - GPU/AI Infra
Senior Data Center Engineer - GPU/AI Infra

Nebius B.V. • Vineland (NJ)

On-site
USD 85,000 - 140,000
Career growth
Flexibility and ownership
Collaborative culture
+2
AI Data Center Ops Lead — GPU/HPC Infrastructure
AI Data Center Ops Lead — GPU/HPC Infrastructure

Nscale • Town of Norway (WI)

On-site
USD 120,000 - 170,000
Remote GPU Infra NOC Engineer — Automation & AI Ops
Remote GPU Infra NOC Engineer — Automation & AI Ops

Orionplacement • Pittsburgh

On-site
USD 75,000 - 140,000
Bonus and equity opportunities
Medical, dental, and vision insurance
401(k)
AI Infrastructure & Automation Engineer
AI Infrastructure & Automation Engineer

Nscale • New York (NY)

On-site
USD 140,000 - 210,000
Competitive package
Equity
Growth opportunities
Remote GPU Infra NOC Engineer | Automation & AI Ops
Remote GPU Infra NOC Engineer | Automation & AI Ops

Orion Placement • Pittsburgh

On-site
USD 75,000 - 140,000
Dental insurance
Paid time off
Retirement plan
+2
AI Infra Engineer – SRE (Kubernetes)
AI Infra Engineer – SRE (Kubernetes)

Berrybytes • United States

On-site
USD 110,000 - 150,000
Staff Site Reliability Engineer - AI Infrastructure
Staff Site Reliability Engineer - AI Infrastructure

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 297,500 - 402,500
Huge stock options
Company bonus
Unlimited PTO
+1
Infra Engineer - SRE(Kubernetes)
Infra Engineer - SRE(Kubernetes)

GMI Cloud • United States

On-site
USD 100,000 - 130,000