Network Engineer

Gimlet Labs

San Francisco (CA)

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Gimlet Labs is seeking a Network Engineer to design, build, and scale the network infrastructure powering production‑scale AI and distributed systems. The role covers physical infrastructure, network architecture, deployment workflows, and operational troubleshooting across high‑performance environments.

The ideal candidate will have deep data center networking knowledge, experience with spine‑leaf fabrics, ECMP, BGP, EVPN, VXLAN, and a hands‑on approach to automation, monitoring, and incident

Qualifications

  • Design, deploy and operate production network infrastructure.
  • Strong fundamentals in routing, switching, performance and reliability.
  • Experience with spine-leaf/Clos fabrics, ECMP, BGP, EVPN, VXLAN.

Responsibilities

  • Design, deploy, and scale datacenter network infrastructure supporting AI workloads and high-performance compute environments.
  • Lead network provisioning, device configuration, connectivity validation, deployment testing, and production turn-up activities for new infrastructure builds and hardware expansions.
  • Build and maintain scalable network topology designs, IPAM, deployment standards, and infrastructure readiness processes.
  • Troubleshoot complex networking, routing, hardware, connectivity, and performance issues across physical infrastructure and distributed systems environments.
  • Partner closely with infrastructure, systems, deployment, and operations teams to improve network reliability and deployment velocity.
  • Drive automation and operational improvements across provisioning, configuration management, monitoring, and incident response workflows.

Skills

Network design
Routing & Switching
Data center networking
EVPN/VXLAN
BGP/ECMP
Network automation
RoCE/InfiniBand

Tools

Arista hardware
Cisco hardware
Juniper hardware

Job description

About Us

Gimlet is building the next generation of AI infrastructure: large-scale AI datacenters and the orchestration platform that coordinates them.

The future of AI will require vastly more compute than exists today. But as AI workloads become more complex and new hardware architectures emerge, simply deploying more GPUs isn't enough. The challenge is making increasingly diverse compute work together.

Gimlet's platform intelligently partitions and routes workloads across heterogeneous hardware, enabling step‑function improvements in performance and efficiency. Customers deploy through production‑grade APIs without needing to think about hardware selection, placement, or optimization.

We work with foundation labs, hyperscalers, and AI‑native companies to power production workloads at massive scale and help define the infrastructure layer for the future of AI.

About this Role

Gimlet Labs is seeking a Network Engineer to design, build, and scale the network infrastructure powering production‑scale AI and distributed systems for frontier labs, hyperscalers, and other high‑performance compute environments.

This is an opportunity to build the network foundation for systems serving real production traffic at massive scale, while also shaping the network architecture for the next generation of AI datacenters. You will help determine how future high‑performance compute environments are designed, deployed, interconnected, and operated.

The ideal candidate has deep technical knowledge of modern data center networking and is comfortable operating across physical infrastructure, network architecture, deployment workflows, and operational troubleshooting. We are looking for someone who can operate independently, drive infrastructure improvements, and help build scalable networking foundations for high‑performance compute and AI workloads.

What you will work on
  • Design, deploy, and scale datacenter network infrastructure supporting AI workloads, distributed systems, and high‑performance compute environments.

  • Lead network provisioning, device configuration, connectivity validation, deployment testing, and production turn‑up activities for new infrastructure builds and hardware expansions.

  • Build and maintain scalable network topology designs, IPAM, deployment standards, operational documentation, and infrastructure readiness processes.

  • Troubleshoot complex networking, routing, hardware, connectivity, and performance issues across physical infrastructure and distributed systems environments.

  • Partner closely with infrastructure, systems, deployment, and operations teams to improve network reliability, deployment velocity, operational readiness, and infrastructure scalability.

  • Drive automation and operational improvements across provisioning, configuration management, monitoring, deployment validation, and incident response workflows.

You may be a good fit if
  • Have experience designing, deploying, and operating production network infrastructure.

  • Have strong networking fundamentals across routing, switching, connectivity, performance, and reliability.

  • Have worked with spine‑leaf or Clos fabrics, backbone or WAN networks, ECMP, BGP, EVPN, VXLAN, and routing policies.

  • Understand high‑performance AI/HPC networking concepts such as RoCEv2, InfiniBand, lossless Ethernet, QoS, DSCP, queuing, shaping, LAGs, optical transport, DWDM, coherent optics, and traffic engineering.

  • Can troubleshoot complex issues across hardware, software, network, and distributed systems boundaries.

  • Enjoy building systems, automating workflows, and improving operational processes.

  • Work well across engineering, infrastructure, deployment, and operations teams.

  • Take ownership end‑to‑end and operate effectively in ambiguous, fast‑moving environments.

Strong candidates may also have
  • Experience with AI/HPC, GPU, or large‑scale distributed infrastructure.

  • Knowledge of AI application traffic patterns, including collective operations and workload colocation strategies that optimize network performance.

  • Experience with cloud networking on GCP, AWS, Azure, or similar platforms.

  • Experience with Arista, Cisco, Juniper, or NVIDIA networking platforms, as well as Palo Alto Networks PAN‑OS.

  • Experience with network automation using Python, Ansible, Terraform, or similar tooling.

  • Familiarity with RDMA, RoCE, InfiniBand, or other high‑performance networking environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Engineer
Network Engineer

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 250,000 - 320,000
Network Engineering Lead
Network Engineering Lead

The Consensus • San Francisco (CA)

On-site
USD 230,000 - 290,000
Network Engineering Lead
Network Engineering Lead

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff - Infrastructure
Member of Technical Staff - Infrastructure

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
AI Data Center Network Engineer
AI Data Center Network Engineer

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 180,000
Senior Network Engineer, AI Infrastructure & HPC
Senior Network Engineer, AI Infrastructure & HPC

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 250,000 - 320,000
GPU Network Engineer
GPU Network Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD <240,000
Member of Technical Staff - Distributed Systems
Member of Technical Staff - Distributed Systems

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Network Engineer
Network Engineer

QumulusAI • United States

On-site
USD 100,000 - 130,000
Competitive compensation
Equity participation
Opportunity to work on cutting-edge technology
Data Center Technician
Data Center Technician

Gimlet Labs • Oklahoma City (OK)

On-site
USD 60,000 - 90,000