Senior AI Infra Architect – GPU & Kubernetes

Hamilton Barnes ?

United States

Remote

USD 180,000 - 240,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Hamilton Barnes is seeking a Senior Solutions Architect to join its post-sales team, focusing on design and deployment of GPU-accelerated AI/ML workloads on a Kubernetes-native Platform-as-a-Service. You will act as a trusted technical advisor across the customer lifecycle, shaping architectures for enterprise-scale AI infrastructure.

You will collaborate with product and engineering to translate workload requirements into scalable architectures, deliver workshops and PoCs, and mentor junior

Qualifications

  • 8+ years in infrastructure, platform, or solutions engineering, including 3+ years focused on AI/ML infrastructure or MLOps
  • Deep, hands-on Kubernetes expertise (cluster lifecycle, workloads, operators, RBAC) plus direct experience with NVIDIA GPU infrastructure (H100/H200 preferred)
  • Working knowledge of distributed training concepts (NCCL, tensor and pipeline parallelism) and GPU networking (GPU Operator, MIG, SR-IOV, IB/RoCEv2)
  • Experience with LLM inference serving and optimisation (vLLM, NIM, TGI)
  • Confident, credible communicator able to run technical discussions with engineers through to executive stakeholders; strong troubleshooting mindset and customer advisory presence

Responsibilities

  • Design end-to-end AI/ML platform architectures spanning inference, training, and data pipelines for enterprise customers
  • Develop reference architectures for GPU cluster deployment, LLM serving, and multi-tenant ML infrastructure
  • Advise on GPU fabric topology, including NVLink, InfiniBand, and RoCEv2, for distributed training environments
  • Act as the primary technical advisor and escalation point for assigned customers, leading root cause analysis on complex production issues
  • Deliver technical workshops, proof-of-concept engagements, and executive-level presentations on AI infrastructure strategy
  • Design observability strategies across DCGM, OpenTelemetry, eBPF, and GPU metrics pipelines
  • Partner with customer platform, MLOps, and data science stakeholders to translate workload requirements into scalable architecture
  • Feed customer insights back into the product and engineering roadmap, and mentor junior members of the Solutions Architecture team

Skills

Kubernetes expertise
NVIDIA GPU infrastructure
LLM inference serving
Distributed training concepts
Executive stakeholder communication
Root cause analysis

Tools

GPU Operator
MIG / SR-IOV
NCCL
InfiniBand / RoCEv2
OpenTelemetry / eBPF

Job description

Hamilton Barnes is seeking a Senior Solutions Architect to join its post-sales team, focusing on design and deployment of GPU-accelerated AI/ML workloads on a Kubernetes-native Platform-as-a-Service. You will act as a trusted technical advisor across the customer lifecycle, shaping architectures for enterprise-scale AI infrastructure.

You will collaborate with product and engineering to translate workload requirements into scalable architectures, deliver workshops and PoCs, and mentor junior

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Storage Engineer - AI Infra & GPU Clusters (Remote)
Senior Storage Engineer - AI Infra & GPU Clusters (Remote)

Hamilton Barnes • San Francisco (CA)

Remote
USD 170,000 - 230,000
Stock options
Remote work allowance
Senior GPU Compute Solutions Architect
Senior GPU Compute Solutions Architect

Computacenter AG & Co. oHG • Northern (KY)

Hybrid
USD 190,000 - 230,000
Senior AI Infrastructure Architect — Enterprise GPU Clusters
Senior AI Infrastructure Architect — Enterprise GPU Clusters

NVIDIA • California (MO)

On-site
USD 184,000 - 287,500
Equity
Benefits
Platform AI Infra Engineer (Kubernetes & Cloud)
Platform AI Infra Engineer (Kubernetes & Cloud)

Hamilton Barnes Associates Limited • United States

On-site
USD 213,000 - 288,000
Equity
Health care
Senior AI Infrastructure Architect – GPU Clusters
Senior AI Infrastructure Architect – GPU Clusters

NVIDIA • California (MO)

On-site
USD 184,000 - 356,500
Equity
Benefits
Senior AI Infra Engineer - Kubernetes Scale & Performance
Senior AI Infra Engineer - Kubernetes Scale & Performance

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Hybrid work
Senior AI Infra & Kubernetes Architect
Senior AI Infra & Kubernetes Architect

NVIDIA Corporation • Durham (NC)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior AI Infra Architect GPU & Cloud
Senior AI Infra Architect GPU & Cloud

Nvidia Corporation in • Washington

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior AI Infra Architect for Enterprise GPU Clusters
Senior AI Infra Architect for Enterprise GPU Clusters

NVIDIA • New York (NY)

On-site
USD 184,000 - 287,500
Equity
Benefits
Senior AI Infrastructure Engineer - GPU & Kubernetes
Senior AI Infrastructure Engineer - GPU & Kubernetes

HCL Technologies Limited • California (MO)

On-site
USD 120,000 - 180,000
401(k) retirement plan
Paid time off (PTO)
Paid holidays
+1