Network Engineer: AI Infrastructure & High-Performance Networking

OpenAI

California (MO)

On-site

USD 150,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenAI is seeking engineers to build and operate the networking foundation behind frontier AI systems. You will work on host networking, datacenter fabrics, and global WAN infrastructure to support AI training and inference at scale.

Role involves solving performance-critical infrastructure problems near the hardware/software boundary, with a focus on low latency, high throughput, and reliability across diverse environments.

Qualifications

  • Experience building or operating large-scale networking or distributed systems infrastructure.
  • Comfortable working close to the hardware/software boundary.
  • Experience with Linux networking, kernel systems, NICs, RDMA, or performance-sensitive infra software.

Responsibilities

  • Design, build, and operate networking systems that support large-scale AI training and inference infrastructure.
  • Improve performance, reliability, and scalability across host networking, datacenter fabrics, and WAN systems.
  • Develop automation for provisioning, configuration management, validation, upgrades, and lifecycle management of networking infrastructure.
  • Build tooling and observability systems for network health, performance analysis, debugging, and automated remediation.
  • Optimize network performance across RDMA, RoCE, InfiniBand, Ethernet, and high-performance GPU interconnects.
  • Define and operationalize networking protocols, readiness criteria, and continuous validation systems.
  • Partner with compute, storage, hardware, and infrastructure teams to scale networking with fleet growth.
  • Contribute to architecture decisions around topology design, capacity planning, failure domains, and network reliability.
  • Diagnose complex distributed systems and networking issues across large heterogeneous compute environments.

Skills

Linux networking
RDMA
Performance engineering
C++
Python
Go

Tools

DPDK
InfiniBand
RoCE

Job description

OpenAI is seeking engineers to build and operate the networking foundation behind frontier AI systems. You will work on host networking, datacenter fabrics, and global WAN infrastructure to support AI training and inference at scale.

Role involves solving performance-critical infrastructure problems near the hardware/software boundary, with a focus on low latency, high throughput, and reliability across diverse environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Core Network Engineering
Software Engineer, Core Network Engineering

OpenAI • California (MO)

On-site
USD 150,000 - 230,000
Network Engineer: Secure, Scalable & Automated Infra
Network Engineer: Secure, Scalable & Automated Infra

OpenAI • San Francisco (CA)

Hybrid
USD 293,000 - 385,000
Relocation assistance
Hybrid work model (3 days in office)
AI Infrastructure Network Operations Engineer
AI Infrastructure Network Operations Engineer

OpenAI • San Francisco (CA)

On-site
USD 140,000 - 210,000
Compute Infrastructure Engineer for Frontier AI
Compute Infrastructure Engineer for Frontier AI

OpenAI • California (MO)

On-site
USD 180,000 - 260,000
Software Engineer, Core Network Engineering
Software Engineer, Core Network Engineering

OpenAI • San Francisco (CA)

On-site
USD 230,000 - 342,000
Medical, dental, and vision insurance
401(k) retirement plan with employer match
Paid parental leave
+3
Network Operations Engineer, AI Networking
Network Operations Engineer, AI Networking

OpenAI • San Francisco (CA)

On-site
USD 140,000 - 210,000
Networking Engineer - AI Data Center Fabric
Networking Engineer - AI Data Center Fabric

Tensordyne • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
Comprehensive benefits
Flexible spending options
Recognition program
Lead Front-End Network Engineer, AI Infrastructure
Lead Front-End Network Engineer, AI Infrastructure

Nscale • New York (NY)

On-site
USD 180,000 - 240,000
Equity
Medical, dental, vision insurance
Flexible paid time off
+2
Senior Network Engineer — AI Infra & HPC Fabric Expert
Senior Network Engineer — AI Infra & HPC Fabric Expert

Nscale • Houston (TX)

On-site
USD 150,000 - 210,000
Competitive benefits package
Flexible paid time off
Parental leave
+1
Global HPC Network Engineer for AI Infra
Global HPC Network Engineer for AI Infra

Together • San Francisco (CA)

On-site
USD 190,000 - 280,000
Startup equity
Health insurance
Competitive benefits