Senior Storage Engineer - AI Infra & GPU Clusters (Remote)

Hamilton Barnes

San Francisco (CA)

Remote

USD 170,000 - 230,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Stock options
Remote work allowance

Job summary

Hamilton Barnes is seeking a Senior Storage Engineer to own the high-performance storage layer for large-scale GPU clusters and AI workloads. You will design, deploy, and operate storage platforms, collaborating with infrastructure, compute, and networking teams to scale performance.

The role focuses on optimizing distributed storage, troubleshooting bottlenecks, and driving resiliency and data protection through automation and observability in petabyte-scale environments.

Qualifications

  • Experience in storage engineering for HPC/AI infra or hyperscale data centers.
  • Hands-on expertise with VAST Data storage platforms preferred.
  • Strong knowledge of high-performance distributed storage architectures and parallel file systems.
  • Experience supporting GPU-intensive AI/ML workloads and high-throughput data environments.
  • Proficient Linux system administration.
  • Experience troubleshooting performance across storage, networking and compute.
  • Familiarity with NFS, RDMA, NVMe-oF, InfiniBand and modern storage networks.
  • Automation using Python, Bash, or Ansible.
  • Focus on scalability, resiliency and data protection in enterprise storage environments.

Responsibilities

  • Design, deploy, and operate large-scale high-performance storage platforms supporting AI and HPC workloads.
  • Manage and optimize distributed storage environments for large GPU training and inference clusters.
  • Work closely with platform, compute, and networking teams to ensure end-to-end infrastructure performance.
  • Troubleshoot storage bottlenecks, latency issues, throughput constraints, and data flow inefficiencies.
  • Contribute to storage architecture strategy, scalability planning, and operational best practices.
  • Automate storage provisioning, monitoring, and lifecycle management processes.
  • Support performance tuning across parallel file systems, object storage, and AI data pipelines.
  • Implement observability and capacity planning solutions for petabyte-scale environments.

Skills

Storage engineering
HPC/AI infra
VAST Data
Distributed storage
GPU AI workloads
Linux administration
NFS/RDMA/NVMe-oF
Python/Bash/Ansible
Scalability & resiliency

Tools

VAST Data
InfiniBand

Job description

Hamilton Barnes is seeking a Senior Storage Engineer to own the high-performance storage layer for large-scale GPU clusters and AI workloads. You will design, deploy, and operate storage platforms, collaborating with infrastructure, compute, and networking teams to scale performance.

The role focuses on optimizing distributed storage, troubleshooting bottlenecks, and driving resiliency and data protection through automation and observability in petabyte-scale environments.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Infra Architect – GPU & Kubernetes
Senior AI Infra Architect – GPU & Kubernetes

Hamilton Barnes ? • United States

Remote
USD 180,000 - 240,000
Storage Engineer (AI Infrastructure) - Hosting
Storage Engineer (AI Infrastructure) - Hosting

Hamilton Barnes • San Francisco (CA)

Remote
USD 170,000 - 230,000
Stock options
Remote work allowance
Senior Storage Production Engineer: Low-Latency AI Cloud
Senior Storage Production Engineer: Low-Latency AI Cloud

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 176,000 - 276,000
Equity opportunities
Comprehensive benefits package
Senior Storage Platform Engineer: Self-Service & IaC
Senior Storage Platform Engineer: Self-Service & IaC

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 168,000 - 334,000
Equity
Benefits
Senior AI Storage Architect — High-Performance Infra
Senior AI Storage Architect — High-Performance Infra

NVIDIA Corporation • Santa Clara (CA)

Remote
AUD 260,000 - 346,000
Senior AI Cloud Storage Systems Engineer
Senior AI Cloud Storage Systems Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 176,000 - 334,000
Senior AI/HPC Storage Architect
Senior AI/HPC Storage Architect

Hydra Host, Inc. • Miami (FL)

On-site
USD 120,000 - 150,000
Senior Network Engineer – AI GPU Infra (400G, Remote)
Senior Network Engineer – AI GPU Infra (400G, Remote)

Hamilton Barnes • San Francisco (CA)

Remote
USD 191,000 - 259,000
Stock options
Remote working options and allowance
Senior Storage Engineer - Kubernetes for AI & GPUs (Remote)
Senior Storage Engineer - Kubernetes for AI & GPUs (Remote)

Bazeta • Northern (KY)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Benefits plan
Conference attendance
Senior AI/HPC Storage Architect
Senior AI/HPC Storage Architect

Hydra Host • Miami (FL)

On-site
USD 120,000 - 150,000