HPC Network Engineer

Autonomai Recruitment

New York (NY)

On-site

USD 500,000 - 800,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation package
Global mobility opportunities
Professional development benefits

Job summary

Autonomai Recruitment is seeking an HPC Network Engineer in New York to design, deploy, and operate low-latency, high-throughput networks powering global trading systems.

You will apply expertise from GPU clusters, RDMA fabrics, and lossless Ethernet to build scalable leaf-spine architectures and observability-driven operations across data centers and colocation facilities.

Qualifications

  • 8+ years designing, deploying, and operating large-scale distributed systems or data-intensive backend platforms at a hyperscaler or major tech company.
  • Deep expertise in RDMA and lossless Ethernet architectures for GPU clusters and HPC workloads.
  • Proven leaf-spine/Clos network designs at scale, including high-speed Ethernet switching platforms and smartNICs.
  • Hands-on scripting/automation experience in Python, Go, or similar languages; familiarity with configuration management at scale.
  • Experience with container platforms and microservices, including host-level and pod networking architectures.
  • Excellent troubleshooting, communication, and stakeholder management skills.

Responsibilities

  • Design and deploy large-scale, low-latency network architectures for trading, risk, and research workloads across global data centers and colocation facilities.
  • Own RDMA-based GPU interconnect fabrics and lossless leaf-spine topologies optimized for high-performance computing workloads.
  • Build and maintain L2/L3 network topologies (routing, multicast, overlay networks) at scale, drawing on experience from AI/ML and distributed systems infrastructure.
  • Automate network provisioning, configuration management, and diagnostics using scripting and infrastructure-as-code practices.
  • Monitor and optimize network performance using observability and packet analysis tools, with a focus on collective communications and congestion control.
  • Partner with systems, platform, and trading teams to troubleshoot complex cross-layer performance issues in production environments.
  • Lead vendor selection, capacity planning, and lifecycle management for switching platforms and high-speed optics.
  • Support colocation builds, cross-connect provisioning, carrier diversity, and timing architectures for trading venues.
  • Contribute to incident response, design reviews, and continuous improvement of observability and resilience across the infrastructure.

Skills

RDMA architectures
Lossless Ethernet
Leaf-spine networks
Python scripting
Go scripting
Container platforms
Microservices networking
Troubleshooting
Stakeholder management

Education

Bachelor's degree in Computer Science, Computer Engineering, or related technical field

Tools

Observability tools
Packet analysis tools
Infrastructure as Code tooling

Job description

We're looking for an HPC Network Engineer with experience at a leading hyperscaler or large-scale tech company to help design and operate the network infrastructure that powers global trading systems.

You’ll apply expertise from building massive GPU clusters, RDMA fabrics, and lossless Ethernet architectures to create some of the lowest-latency, highest-throughput networks in the industry.

This role is ideal for engineers from top-tier tech companies who want to bring their large-scale infrastructure skills into high-performance trading.

What You’ll Do
  • Design and deploy large-scale, low-latency network architectures for trading, risk, and research workloads across global data centers and colocation facilities
  • Own RDMA-based GPU interconnect fabrics and lossless leaf-spine topologies optimized for high-performance computing workloads
  • Build and maintain L2/L3 network topologies (routing, multicast, overlay networks) at scale, drawing on experience from AI/ML and distributed systems infrastructure
  • Automate network provisioning, configuration management, and diagnostics using scripting and infrastructure-as-code practices
  • Monitor and optimize network performance using observability and packet analysis tools, with a focus on collective communications and congestion control
  • Partner with systems, platform, and trading teams to troubleshoot complex cross-layer performance issues in production environments
  • Lead vendor selection, capacity planning, and lifecycle management for switching platforms and high-speed optics
  • Support colocation builds, cross-connect provisioning, carrier diversity, and timing architectures for trading venues
  • Contribute to incident response, design reviews, and continuous improvement of observability and resilience across the infrastructure
What We’re Looking For
  • 8+ years of experience designing, deploying, and operating large-scale distributed systems or data-intensive backend platforms at a hyperscaler or major tech company
  • Deep expertise in RDMA and lossless Ethernet architectures for GPU clusters and HPC workloads
  • Proven experience with leaf-spine/Clos network designs at scale, including high-speed Ethernet switching platforms and smartNICs
  • Hands-on scripting/automation experience in Python, Go, or similar languages; familiarity with configuration management at scale
  • Experience with container platforms and microservices, including host-level and pod networking architectures
  • Excellent troubleshooting, communication, and stakeholder management skills
  • Bachelor's degree in Computer Science, Computer Engineering, or related technical field (or equivalent practical experience)
Nice to Have
  • Experience supporting large-scale AI/ML or HPC workloads
  • Familiarity with streaming systems and high-throughput data pipeline architectures
  • Experience with performance benchmarking and profiling tools for network and system performance
  • Demonstrated influence across organizations (tech lead, architect, principal/IC leadership roles) in a hyperscaler environment
What We Offer
  • Competitive compensation package (base + bonus) with total compensation up to $500K-$800K+ for top candidates
  • Opportunity to apply hyperscaler-grade networking expertise to one of the world's most advanced trading and risk computing environments
  • Collaborative, high-performance culture with direct impact on trading infrastructure and business outcomes
  • Comprehensive benefits, professional development, and global mobility opportunities
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Network Engineer
HPC Network Engineer

Autonomai Recruitment • Austin (TX)

On-site
USD 500,000 - 800,000
Sign-on bonus
Global mobility opportunities
Competitive compensation package
HPC Network Engineer
HPC Network Engineer

Hamilton Barnes ? • New York (NY)

Hybrid
USD 450,000 - 550,000
Hyperscale HPC Network Architect for Trading
Hyperscale HPC Network Architect for Trading

Autonomai Recruitment • Austin (TX)

On-site
USD 500,000 - 800,000
Sign-on bonus
Global mobility opportunities
Competitive compensation package
HPC Network Engineer - Banking & Finance
HPC Network Engineer - Banking & Finance

Hamilton Barnes Associates Limited • New York (NY)

On-site
USD 300,000 - 500,000
Senior Network Engineer - Trading
Senior Network Engineer - Trading

Hamilton Barnes Associates Limited • New York (NY)

On-site
USD 350,000 - 500,000
Elite peers
Cutting-edge hardware
Senior Network Engineer
Senior Network Engineer

Thurn Partners • United States

On-site
USD 120,000 - 160,000
Emerging Network Architect
Emerging Network Architect

NMC2 • Dallas (TX)

On-site
USD 100,000 - 140,000
HPC Network Engineer
HPC Network Engineer

Hudson River Trading • Austin (TX)

On-site
USD 200,000 - 300,000
Medical, dental, vision insurance
20 vacation days and 10 paid holidays
Discretionary performance-based bonuses
Senior Network Engineer
Senior Network Engineer

Citadel Enterprise Americas LLC • City of Rochester (NY)

On-site
USD 175,000 - 350,000
Discretionary incentive compensation
Medical insurance
Retirement plan
+1
Senior Manager Network Engineering
Senior Manager Network Engineering

Request Technology, LLC • Chicago (IL)

On-site
USD 180,000 - 240,000