Senior GPU Solutions Architect – Enterprise HPC UK

Referment

Greater London

On-site

GBP 90,000 - 150,000

Full time

29 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Referment is partnering with a cloud infrastructure provider to shape large-scale GPU platforms for UK enterprise customers. The role focuses on architecting and delivering scalable GPU clusters, with emphasis on performance, efficiency and cost trade-offs.

You will lead architecture workshops, design high-level and low-level documents, and translate workloads into bare-metal, Slurm or Kubernetes-based solutions across NVIDIA GPU platforms, NVLink/NVSwitch, InfiniBand fabrics and cooling

Qualifications

  • 5+ years in solution architecture, systems engineering or technical pre-sales focused on HPC, AI infrastructure or high-performance cloud platforms.
  • Deep knowledge of NVIDIA HGX or DGX systems, GPU interconnects and modern rack-scale GPU architectures.
  • Strong experience designing InfiniBand and/or RoCE networks, including congestion management and GPU-direct technologies.
  • Practical understanding of GPU workload orchestration through Kubernetes and associated NVIDIA operators, or bare-metal environments using Slurm, Ansible and Terraform.
  • A track record of producing robust HLDs, LLDs, network diagrams and itemised infrastructure designs.

Responsibilities

  • Design enterprise GPU clusters and produce clear high-level and low-level designs, rack and network diagrams, and detailed bills of materials.
  • Shape scale-up and scale-out architectures across NVIDIA GPU platforms, NVLink and NVSwitch, high-speed InfiniBand and RoCE fabrics, storage, power and cooling.
  • Translate customer workload, performance and commercial requirements into practical bare-metal, Slurm or Kubernetes-based solutions.
  • Lead technical discovery sessions, architecture workshops and executive presentations, and contribute to complex proposals and RFP responses.
  • Plan and oversee proofs of concept, using appropriate benchmarking to validate throughput, latency and distributed training performance.
  • Work closely with enterprise customers, hardware vendors and internal engineering teams, feeding technical insight back into the platform roadmap.

Skills

HPC architecture
GPGPU platforms
Kubernetes orchestration
InfiniBand networks
Pre-sales experience

Tools

Kubernetes
Slurm
Ansible
Terraform
NVIDIA HGX/DGX

Job description

Referment is partnering with a cloud infrastructure provider to shape large-scale GPU platforms for UK enterprise customers. The role focuses on architecting and delivering scalable GPU clusters, with emphasis on performance, efficiency and cost trade-offs.

You will lead architecture workshops, design high-level and low-level documents, and translate workloads into bare-metal, Slurm or Kubernetes-based solutions across NVIDIA GPU platforms, NVLink/NVSwitch, InfiniBand fabrics and cooling

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior GPU & AI Infrastructure Architect
Senior GPU & AI Infrastructure Architect

Referment • Greater London

On-site
GBP 90,000 - 120,000
Senior Solution Engineer – GPU and AI Infrastructure (F673F3F)
Senior Solution Engineer – GPU and AI Infrastructure (F673F3F)

Referment • Greater London

On-site
GBP 90,000 - 120,000
Senior Solution Engineer – GPU and AI Infrastructure (CAF673F)
Senior Solution Engineer – GPU and AI Infrastructure (CAF673F)

Referment • Greater London

On-site
GBP 90,000 - 150,000
Senior GPU & AI Infra Architect — Remote, 4-Day Week
Senior GPU & AI Infra Architect — Remote, 4-Day Week

Civo Ltd • United Kingdom

Hybrid
GBP 110,000 - 170,000
4-day week
Uncapped holidays
Remote work environment
Senior Cloud & DevOps Architect — GPU-Accelerated AI/HPC
Senior Cloud & DevOps Architect — GPU-Accelerated AI/HPC

NVIDIA • United Kingdom

On-site
GBP 110,000 - 170,000
Technical Solutions Architect – Investors
Technical Solutions Architect – Investors

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 110,000 - 150,000
Senior Technical Operations & Deployment Engineer (GPU Cloud Infrastructure)
Senior Technical Operations & Deployment Engineer (GPU Cloud Infrastructure)

Xapply • United Kingdom

Remote
GBP 90,000 - 130,000
Lead GPU Infrastructure Architect for Scalable AI Clusters
Lead GPU Infrastructure Architect for Scalable AI Clusters

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 110,000 - 150,000
Senior GPU HPC Engineer: InfiniBand & KVM Optimization
Senior GPU HPC Engineer: InfiniBand & KVM Optimization

Nebius • Greater London

On-site
GBP 90,000 - 130,000
Competitive compensation
Career growth
Flexibility and ownership
+3
HPC Network Engineer - GPU Cloud Infra (Remote UK)
HPC Network Engineer - GPU Cloud Infra (Remote UK)

asobbi • United Kingdom

Remote
GBP 53,000 - 69,000