HPC Network Engineer

Fuse Energy, LLC

Greater London

Hybrid

GBP 85,000 - 120,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully expensed tech to match your need
Private health insurance
Breakfast and dinner allowance

Job summary

Fuse Energy, a London-based energy startup, is hiring a Senior Network Engineer to design, deploy and operate the company's multi-tenant AI cluster network fabric. You will own architecture through day-2 operations, including RDMA fabrics, leaf-spine design, and per-tenant isolation.

You’ll automate provisioning with Python and Ansible, build observability dashboards, and mentor colleagues while expanding the data-centre and office networks.

Qualifications

  • 5+ years as a network engineer operating production data centre networks.
  • Strong dynamic routing experience (BGP in particular) and overlay/encapsulation design (EVPN/VXLAN).
  • Hands-on experience with leaf-spine / Clos fabric design and operation.
  • Experience with modern data centre network operating systems and Linux networking stack.
  • Practical RDMA fabric experience: lossless Ethernet (RoCEv2) or InfiniBand for GPU workloads.
  • Network automation as a working practice: Python, Ansible, source-of-truth config.
  • Solid Linux administration fundamentals for host and switch debugging.
  • Experience with network telemetry and monitoring (Prometheus/Grafana).
  • Experience running corporate/campus networks: wired/wireless, NAC/802.1X, VPN.

Responsibilities

  • Design and operate lossless, RDMA-capable fabrics for GPU compute and storage traffic, including QoS and buffer tuning.
  • Build and manage leaf-spine data centre fabrics with routed underlay and overlay (BGP, EVPN/VXLAN).
  • Implement and maintain per-tenant network isolation across compute, storage and management planes.
  • Automate network provisioning, configuration and validation, treating switch config as code (Ansible, Python, NetBox as source of truth).
  • Build telemetry and observability for the fabric: dashboards and alerting for congestion and link degradation.
  • Troubleshoot performance end to end from optics to NIC/DPU and host communication.
  • Operate the out-of-band management network and remote recovery paths.
  • Support tenant onboarding: segmentation, addressing, bandwidth and isolation guarantees.
  • Write design documentation capturing decisions and trade-offs.
  • Own and maintain the office network: wired/wireless, firewalling, VPN and office-to-data-centre connectivity.
  • Upskill colleagues through documentation and run-throughs.

Skills

Network engineering
BGP routing
EVPN/VXLAN
RDMA networking
Linux networking
Automation (Python)
Ansible

Tools

NetBox
Prometheus/Grafana
CI/CD pipelines
Packet capture tools

Job description

Fuse Energy is an energy startup on a mission to make energy abundant and affordable, fast. We combine first-principles thinking with cutting-edge technology to build a radically better energy system.

We've raised over $200M from top-tier investors including Balderton, Lakestar, Accel, Creandum, Lowercarbon, Ribbit, 20VC, Hummingbird and Collaborative Fund, alongside strategic angels including Nico Rosberg and GPs behind Meta, Revolut, Spotify and Uber.

We're building a fully integrated energy company: developing our own solar, batteries and other generation projects, building our own hardware, improving and developing grid infrastructure, trading power in real time, using AI across the business, and installing distributed energy in homes. By selling directly to consumers we cut out the middleman, lower costs and pass the savings on to our customers.

You'll design, deploy and operate the network fabric for our multi-tenant AI cluster: the high-performance compute and storage fabrics carrying RDMA traffic between GPUs, the tenant-facing and management networks, firewalling and tenant isolation, and the out-of-band infrastructure that keeps it all recoverable. Beyond the data centre, you'll own the office network and act as the networking authority for the company. You own the fabric from architecture through day-2 operations.

Responsibilities
  • Design and operate lossless, RDMA-capable fabrics (e.g. RoCEv2, InfiniBand) for GPU compute and storage traffic, including QoS, congestion control and buffer tuning at scale
  • Build and manage leaf-spine data centre fabrics, with routed underlay and overlay design (e.g. BGP, EVPN/VXLAN)
  • Implement and maintain per-tenant network isolation across compute, storage and management planes
  • Automate network provisioning, configuration and validation, treating switch config as code (e.g. Ansible, Python, NetBox as source of truth), deployed through CI
  • Build telemetry and observability for the fabric: flow-level and buffer-level visibility, dashboards and alerting that catch congestion and link degradation before tenants do
  • Troubleshoot performance issues end to end, from optics and cabling through switch buffers to NIC/DPU configuration and collective-communication behaviour on the hosts
  • Operate the out-of-band management network, console access and remote recovery paths
  • Support tenant onboarding: segmentation and addressing, bandwidth and isolation guarantees, and capacity planning as the cluster scales
  • Write clear design documentation capturing decisions, rationale and rejected alternatives
  • Own and maintain the office network: wired and wireless infrastructure, firewalling, VPN/remote access and office-to-data-centre connectivity
  • Upskill colleagues on networking through documentation, run-throughs and pairing
  • 5+ years as a network engineer operating production data centre networks
  • Strong dynamic routing experience (BGP in particular), plus overlay/encapsulation design and troubleshooting (e.g. EVPN/VXLAN)
  • Hands-on experience with leaf-spine / Clos fabric design and operation
  • Experience with modern data centre network operating systems and comfortable in the Linux networking stack, not just a vendor CLI
  • Practical RDMA fabric experience: lossless Ethernet (e.g. RoCEv2 with PFC/ECN/DCQCN tuning) or InfiniBand, understanding why lossless behaviour matters for GPU workloads
  • Network automation as a working practice: scripting (e.g. Python), configuration management (e.g. Ansible), config generation from a source of truth, version-controlled changes
  • Solid Linux administration fundamentals: you can debug from the host side as well as the switch side
  • Experience with network telemetry and monitoring (e.g. Prometheus/Grafana, sFlow/IPFIX, streaming telemetry)
  • Experience running corporate/campus networks: wired and wireless, switching, NAC/802.1X, VPN and remote access
  • Clear communicator who enjoys teaching
  • Bonus: GPU cluster networking (NVIDIA Spectrum-X or Quantum InfiniBand, ConnectX/BlueField NICs and DPUs, UFM, SHARP or equivalent); container networking (CNI, BGP integration); collective-communication libraries; multi-tenant VRF-based isolation; enterprise firewalls (FortiGate, Palo Alto); storage networking (NVMe-oF); bare-metal provisioning (MAAS, PXE, Redfish); optical layer at 200/400/800G; greenfield data centre network build; CCNP/CCIE or equivalent
  • Competitive salary and eligibility for equity
  • Biannual bonus scheme
  • Fully expensed tech to match your needs
  • Private health insurance
  • Breakfast and dinner allowance for office-based employees

As we hire globally, benefits vary by location.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Network Engineer
HPC Network Engineer

Fuse Energy • Greater London

On-site
GBP 90,000 - 120,000
Competitive salary and an equity sign‑
Biannual bonus scheme
Fully expensed tech to match your need
+1
Network Engineer
Network Engineer

asobbi • United Kingdom

Remote
GBP 53,000 - 69,000
Highly competitive package with equity
Dynamic progression plan
Human-first flexibility
Network Engineer
Network Engineer

NexGen Cloud • Greater London

On-site
GBP 65,000 - 100,000
Competitive salary
Employee wellbeing benefits
25 days holiday
+3
HPC Network Engineer
HPC Network Engineer

Hamilton Barnes ? • City Of London

Hybrid
GBP 332,000 - 406,000
HPC Networking Architect for GPU Clusters
HPC Networking Architect for GPU Clusters

Fuse Energy • Greater London

On-site
GBP 90,000 - 120,000
Competitive salary and an equity sign‑
Biannual bonus scheme
Fully expensed tech to match your need
+1
HPC Network Engineer
HPC Network Engineer

La Fosse • Greater London

On-site
GBP 70,000 - 100,000
Network Engineer
Network Engineer

G-Research • City of Westminster

On-site
GBP 70,000 - 110,000
Lunch provided
Barista bar
30 days annual leave
+4
Cloud HPC Network Engineer
Cloud HPC Network Engineer

Allscreens Nationwide Ltd • Cambridge

On-site
GBP 70,000 - 110,000
HPC Network Engineer - Banking & Finance
HPC Network Engineer - Banking & Finance

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 450,000 - 550,000
Work on advanced low-latency networkch
Influence over architecture and automa
Technical engineering team
+1
Principal Network Engineer
Principal Network Engineer

Nscale • Greater London

On-site
GBP 120,000 - 170,000
Base + equity
Flexible working
Competitive package
+1