HPC Networking Architect for GPU Clusters

Fuse Energy

Greater London

On-site

GBP 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary and an equity sign‑
Biannual bonus scheme
Fully expensed tech to match your need
Breakfast and dinner allowance for off

Job summary

Fuse Energy is seeking a senior network engineer to design, deploy and operate the fabric for a multi-tenant AI cluster and data centre. You will own architecture, day-2 operations, and keep the network secure, scalable and observable.

You will implement lossless RDMA fabrics (RoCEv2/InfiniBand), leaf-spine designs, automation with Ansible and NetBox, and deliver clear design documentation while mentoring colleagues across teams.

Qualifications

  • 5+ years of production data centre networking experience.
  • Strong BGP/EVPN/VXLAN design and troubleshooting.
  • Hands-on experience with leaf-spine/Clos fabrics and Linux networking.

Responsibilities

  • Design and operate lossless, RDMA-capable fabrics for GPU compute and storage.
  • Build leaf-spine data centre fabrics with routed underlay and overlay design.
  • Implement per-tenant network isolation across compute, storage, and management.
  • Automate provisioning, configuration, and validation; treat switch config as code.
  • Build telemetry and observability for fabric with dashboards and alerts.
  • Troubleshoot performance end-to-end from optics to NIC/DPU configuration.
  • Operate the out-of-band management network and remote recovery paths.
  • Support tenant onboarding: segmentation, addressing, bandwidth guarantees.
  • Write design documentation detailing decisions and reasoning.
  • Own office network: wired/wireless, firewalling, VPN, and connectivity to data centre.
  • Upskill colleagues through documentation, run-throughs, and pairing.

Skills

Network engineering
BGP
EVPN/VXLAN
Leaf-spine design
Linux networking
Python scripting
Automation
Documentation
Telemetry/Monitoring
Campus networking

Tools

Ansible
NetBox
Prometheus/Grafana
MAAS
PXE

Job description

Fuse Energy is seeking a senior network engineer to design, deploy and operate the fabric for a multi-tenant AI cluster and data centre. You will own architecture, day-2 operations, and keep the network secure, scalable and observable.

You will implement lossless RDMA fabrics (RoCEv2/InfiniBand), leaf-spine designs, automation with Ansible and NetBox, and deliver clear design documentation while mentoring colleagues across teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Network Engineer
HPC Network Engineer

Fuse Energy • Greater London

On-site
GBP 90,000 - 120,000
Competitive salary and an equity sign‑
Biannual bonus scheme
Fully expensed tech to match your need
+1
Founding GPU Engineer — CUDA Performance for HPC
Founding GPU Engineer — CUDA Performance for HPC

Fuse Energy, LLC • Greater London

On-site
GBP 90,000 - 130,000
Competitive salary
Equity sign-on bonus
Biannual bonus
+2
Cloud HPC Network Engineer
Cloud HPC Network Engineer

Allscreens Nationwide Ltd • Cambridge

On-site
GBP 70,000 - 110,000
AI/HPC Data Center Network Architect
AI/HPC Data Center Network Architect

Curo Services • Greater London

On-site
GBP 107,000 - 131,000
HPC Network Engineer - GPU Cloud Infra (Remote UK)
HPC Network Engineer - GPU Cloud Infra (Remote UK)

asobbi • United Kingdom

Remote
GBP 53,000 - 69,000
Senior GPU HPC Engineer: InfiniBand & KVM Optimization
Senior GPU HPC Engineer: InfiniBand & KVM Optimization

Nebius • Greater London

On-site
GBP 90,000 - 130,000
Competitive compensation
Career growth
Flexibility and ownership
+3
HPC Systems Engineer: NVLink GPU Interconnect
HPC Systems Engineer: NVLink GPU Interconnect

CoreWeave • Greater London

On-site
GBP 79,000 - 105,000
Family-level Medical Insurance
Family-level Dental Insurance
Generous Pension Contribution
+6
GPU HPC Architect
GPU HPC Architect

Radiant • Greater London

On-site
GBP 110,000 - 180,000
25 days annual leave
Health & Wellbeing: private medical
Cycle to Work Scheme
+3
Data Center Network Engineer for High-Performance AI
Data Center Network Engineer for High-Performance AI

Era4 • England

On-site
GBP 70,000 - 100,000
Senior AI & HPC Network Architect for Distributed Systems
Senior AI & HPC Network Architect for Distributed Systems

NVIDIA • United Kingdom

On-site
GBP 221,000 - 507,000