GPU Networking Engineer - RDMA & Distributed Inference

Baseten

San Francisco (CA)

On-site

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
100% medical, dental, and vision insurance
Generous PTO policy
Paid parental leave
401(k)
Exposure to ML startups

Job summary

A cutting-edge AI infrastructure company in San Francisco seeks an experienced network engineer to optimize high-performance networking protocols for AI models. The ideal candidate will integrate RDMA and InfiniBand into the inference stack, ensuring efficient communication and low latency. A deep understanding of NVIDIA architectures and experience with C++ or Python are essential. This role offers competitive compensation and substantial equity, alongside comprehensive benefits including health insurance and generous PTO.

Qualifications

  • Deep experience with InfiniBand and RoCE v2.
  • Ability to bridge high-level logic with hardware.
  • Experience diving into TensorRT-LLM source code.

Responsibilities

  • Integrate RDMA/RoCE/InfiniBand capabilities into inference stack.
  • Implement networking layers for efficient Disaggregated KV Cache Offload.
  • Design tools for visualizing packet flow and diagnosing system behaviors.

Skills

High-performance networking protocols
C++ or Python
Memory hierarchy in NVIDIA architectures
Debugging NVLink topology

Tools

NCCL
NVSHMEM
UCX

Job description

A cutting-edge AI infrastructure company in San Francisco seeks an experienced network engineer to optimize high-performance networking protocols for AI models. The ideal candidate will integrate RDMA and InfiniBand into the inference stack, ensuring efficient communication and low latency. A deep understanding of NVIDIA architectures and experience with C++ or Python are essential. This role offers competitive compensation and substantial equity, alongside comprehensive benefits including health insurance and generous PTO.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RDMA-First GPU Networking & Distributed Inference Engineer
RDMA-First GPU Networking & Distributed Inference Engineer

Baseten • New York (NY)

On-site
USD 185,000 - 250,000
GPU Networking Engineer for Distributed Inference & RDMA
GPU Networking Engineer for Distributed Inference & RDMA

The Consensus • New York (NY)

On-site
USD 120,000 - 160,000
Competitive compensation, including equity
100% medical, dental, and vision insurance
Flexible PTO policy
+3
GPU Networking Engineer — RDMA & Distributed Inference
GPU Networking Engineer — RDMA & Distributed Inference

Baseten • United States

Remote
USD 180,000 - 240,000
Senior RDMA Networking Engineer for AI Inference Pods
Senior RDMA Networking Engineer for AI Inference Pods

Etched • San Jose (CA)

On-site
USD 150,000 - 275,000
Full medical, dental, and vision packages
Housing subsidy of $2,000/month
Daily lunch and dinner in our office
+1
Lead Architect, AI Networking & Distributed GPU Dataflow
Lead Architect, AI Networking & Distributed GPU Dataflow

NVIDIA • Austin (TX)

On-site
USD 272,000 - 432,000
Senior AI Networking & Performance Engineer
Senior AI Networking & Performance Engineer

NVIDIA • Colorado

On-site
USD 272,000 - 432,000
Principal Engineer - AI Networking
Principal Engineer - AI Networking

Ll Oefentherapie • Seattle (WA)

On-site
USD 100,000 - 130,000
Lead Architect, AI Networking & Distributed GPU Dataflow
Lead Architect, AI Networking & Distributed GPU Dataflow

NVIDIA • Oregon (WI)

On-site
USD 272,000 - 432,000
Senior AI Infra Networking Engineer | High-Perf GPU Cloud
Senior AI Infra Networking Engineer | High-Perf GPU Cloud

Nscale • United States

Remote
USD 100,000 - 200,000
Senior AI Infra Network Engineer (Datacenter/GPU) (Equity)
Senior AI Infra Network Engineer (Datacenter/GPU) (Equity)

QumulusAI • United States

On-site
USD 100,000 - 130,000