Senior GPU & AI Infrastructure Architect

Referment

Greater London

On-site

GBP 90,000 - 120,000

Full time

41 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Referment is partnering with a cloud infrastructure provider to shape large-scale GPU platforms for customers across the UK. The role focuses on building enterprise GPU clusters, defining robust architectures and detailed BOMs.

You will translate workloads into Kubernetes-based or bare-metal solutions, lead discovery sessions, and drive proofs of concept with benchmarking to validate performance and cost considerations.

Qualifications

  • Minimum five years in solution architecture, systems engineering or technical pre-sales for HPC/AI infra.
  • Deep knowledge of NVIDIA HGX or DGX systems and GPU interconnects.
  • Experience with InfiniBand/RoCE networks and GPU-direct technologies.
  • Hands-on with Kubernetes for workload orchestration or bare-metal with Slurm/Ansible/Terraform.
  • Proven ability to produce end-to-end designs and diagrams.

Responsibilities

  • Design enterprise GPU clusters with high-level & low-level designs, rack and network diagrams, and detailed bills of materials.
  • Shape scale-up and scale-out architectures across NVIDIA GPU platforms, NVLink and NVSwitch, high-speed InfiniBand and RoCE fabrics, storage, power and cooling.
  • Translate customer workload, performance and commercial requirements into practical bare-metal, Slurm or Kubernetes-based solutions.
  • Lead technical discovery sessions, architecture workshops and executive presentations, and contribute to complex proposals and RFP responses.
  • Plan and oversee proofs of concept, using appropriate benchmarking to validate throughput, latency and distributed training performance.
  • Work closely with enterprise customers, hardware vendors and internal engineering teams, feeding technical insight back into the platform roadmap.

Skills

Solution architecture
Systems engineering
Technical pre Sales
NVIDIA HGX/DGX knowledge
GPU interconnects
InfiniBand/RoCE networking
Kubernetes
Slurm
Ansible
Terraform
HLD/LLD design
Executive communication

Education

Degree in a relevant field

Tools

NVIDIA HGX/DGX systems
InfiniBand switches
RoCE fabrics

Job description

Referment is partnering with a cloud infrastructure provider to shape large-scale GPU platforms for customers across the UK. The role focuses on building enterprise GPU clusters, defining robust architectures and detailed BOMs.

You will translate workloads into Kubernetes-based or bare-metal solutions, lead discovery sessions, and drive proofs of concept with benchmarking to validate performance and cost considerations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Solution Engineer – GPU and AI Infrastructure (F673F3F)
Senior Solution Engineer – GPU and AI Infrastructure (F673F3F)

Referment • Greater London

On-site
GBP 90,000 - 120,000
Senior GPU & AI Infra Architect — Remote, 4-Day Week
Senior GPU & AI Infra Architect — Remote, 4-Day Week

Civo Ltd • United Kingdom

Hybrid
GBP 110,000 - 170,000
4-day week
Uncapped holidays
Remote work environment
Lead GPU Infrastructure Architect for Scalable AI Clusters
Lead GPU Infrastructure Architect for Scalable AI Clusters

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 110,000 - 150,000
Senior Cloud & DevOps Architect — GPU-Accelerated AI/HPC
Senior Cloud & DevOps Architect — GPU-Accelerated AI/HPC

NVIDIA • United Kingdom

On-site
GBP 110,000 - 170,000
Technical Solutions Architect – Investors
Technical Solutions Architect – Investors

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 110,000 - 150,000
Platform Engineer – Scale GPU Infra for AI Platform
Platform Engineer – Scale GPU Infra for AI Platform

Ineffable Intelligence LTD • Greater London

Hybrid
GBP 85,000 - 120,000
GPU Infrastructure Lead: Scale, Certification, and Automation
GPU Infrastructure Lead: Scale, Certification, and Automation

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 140,000 - 170,000
Full Benefits
Lead AI GPU Compute Cluster Architect
Lead AI GPU Compute Cluster Architect

Radiant • Greater London

On-site
GBP 120,000 - 190,000
25 days leave
Medical insurance
Cycle to Work
+3
GPU Infrastructure Lead - Systems Integrator
GPU Infrastructure Lead - Systems Integrator

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 140,000 - 170,000
Full Benefits
AI Infrastructure Architect — Hybrid Cloud & GPU HPC
AI Infrastructure Architect — Hybrid Cloud & GPU HPC

Accenture UK & Ireland • Greater London

On-site
GBP 120,000 - 170,000