Senior Network Engineer – HPC & GPU Compute Lab

ProducePay

San Francisco (CA)

On-site

USD 193,000 - 234,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Restricted Stock Units
Paid time off
Comprehensive health, dental & vision
Employer contributions to HSA
Paid parental leave
Life insurance
Disability insurance
Tuition reimbursement
Mental health support
Commuter benefits
Cell phone stipend
401(k) retirement plan with match
Volunteer time off

Job summary

Crusoe Cloud is seeking a high-energy, senior network engineer to lead the deployment and management of our lab network infrastructure supporting GPU compute clusters. You will drive end-to-end builds, fix complex hardware issues, and collaborate with data center operations, engineering, and vendors.

Ideal candidates have 8+ years in network engineering for hyperscale environments, strong experience with Arista/Junos/NVIDIA/Mellanox, and automation using Python and Ansible.

Qualifications

  • Bachelor’s degree in a technical field or equivalent experience.
  • 8+ years in network engineering for large-scale data centers.
  • Expert in structured cabling, optical transceivers, and power/cooling.
  • Strong knowledge of BGP, EVPN-VXLAN.
  • Experience with Arista, Juniper, and NVIDIA/Mellanox platforms.
  • Automation experience with Python and Ansible.

Responsibilities

  • Lead end-to-end deployment of network infrastructure for Crusoe lab.
  • Translate high-level designs into implementation plans and templates.
  • Perform burn-in testing and validation for new clusters.
  • Automate deployment tasks with Python, Ansible, and ZTP.
  • Maintain lab operations across multiple time zones.
  • Diagnose hardware faults and perform FRU repairs with data center teams.
  • Document maintenance activities and SOPs for troubleshooting.
  • Support on-call rotations with Europe team.

Skills

Network engineering
Large-scale data centers
Routing & Switching
BGP & EVPN-VXLAN
Python & Ansible automation
Project management
Troubleshooting
Fiber optics & cabling

Education

Bachelor's degree in a technical field or equivalent

Tools

Arista EOS
Juniper Junos
NVIDIA/Mellanox platforms

Job description

Crusoe Cloud is seeking a high-energy, senior network engineer to lead the deployment and management of our lab network infrastructure supporting GPU compute clusters. You will drive end-to-end builds, fix complex hardware issues, and collaborate with data center operations, engineering, and vendors.

Ideal candidates have 8+ years in network engineering for hyperscale environments, strong experience with Arista/Junos/NVIDIA/Mellanox, and automation using Python and Ansible.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Network Deployment Engineer – Global HPC Infra
Staff Network Deployment Engineer – Global HPC Infra

ProducePay • United States

On-site
USD 193,000 - 234,000
Competitive compensation
Restricted Stock Units
Paid time off & holidays
+10
Senior Manager - AI-Driven Network & Cloud Architecture
Senior Manager - AI-Driven Network & Cloud Architecture

Crusoe • San Francisco (CA)

On-site
USD 250,000 - 300,000
Equity
Health insurance
HSA plan
+3
Senior Network Engineer, (Lab)
Senior Network Engineer, (Lab)

ProducePay • San Francisco (CA)

On-site
USD 193,000 - 234,000
Competitive compensation
Restricted Stock Units
Paid time off
+11
Senior Cloud Support Engineer, HPC & AI Compute
Senior Cloud Support Engineer, HPC & AI Compute

Crusoe • San Francisco (CA)

On-site
USD 125,000 - 145,000
Competitive compensation and equity
401(k) match up to 4%
Health, dental, and vision insurance
+2
Staff Network Production & Reliability Engineer
Staff Network Production & Reliability Engineer

ProducePay • United States

On-site
USD 195,000 - 235,000
Competitive compensation
Equity
Health insurance
+2
Senior Cloud HPC Support Engineer
Senior Cloud HPC Support Engineer

Crusoe • Bellevue (WA)

Hybrid
USD 145,000 - 175,000
Equity
Paid time off
Health insurance
+1
Staff Network Production Engineer: Global Reliability & Ops
Staff Network Production Engineer: Global Reliability & Ops

Crusoe • Sunnyvale (CA)

On-site
USD 195,000 - 235,000
Equity/RSU
Health insurance
Paid time off
+2
Senior Cloud Support Engineer – HPC & GPU Infra
Senior Cloud Support Engineer – HPC & GPU Infra

Crusoe • Sunnyvale (CA)

Hybrid
USD 105,000 - 175,000
Competitive compensation and equity
Restricted Stock Units
Health, dental & vision insurance
+3
Senior HPC Cloud Support Engineer
Senior HPC Cloud Support Engineer

Crusoe • Dallas (TX)

On-site
USD 130,000 - 155,000
Equity
Paid time off
Health insurance
+2
Senior Cloud Support Engineer for HPC & AI Compute
Senior Cloud Support Engineer for HPC & AI Compute

Crusoe • Bellevue (WA)

On-site
USD 125,000 - 145,000
Equity
Paid time off
Health insurance
+12