Lead GPU Inference & ML Infrastructure Engineer

Avride Inc.

Austin (TX)

On-site

USD 140,000 - 190,000

Full time

23 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Avride is seeking a software engineer with leadership experience to drive the ML infrastructure layer across the company. You will lead GPU inference optimization for onboard and high-throughput offboard scenarios, while guiding broader ML infrastructure across pipelines.

The role requires deep C++ experience, strong multi-threading skills, and collaboration with the applied ML team to align neural model architectures with performance goals.

Qualifications

  • Experience with PyTorch and GPU architectures.
  • Proven ability to diagnose and resolve performance issues in large systems.
  • Strong track record building distributed infrastructure.

Responsibilities

  • Own the GPU inference framework with a focus on performance.
  • Take ownership of broader ML infrastructure across pipelines.
  • Collaborate with the applied ML team on model architecture and deployment.

Skills

PyTorch
GPUs
Performance debugging
Infrastructure
C++
Multi-threading

Job description

Avride is seeking a software engineer with leadership experience to drive the ML infrastructure layer across the company. You will lead GPU inference optimization for onboard and high-throughput offboard scenarios, while guiding broader ML infrastructure across pipelines.

The role requires deep C++ experience, strong multi-threading skills, and collaboration with the applied ML team to align neural model architectures with performance goals.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead AI Infrastructure Engineer
Lead AI Infrastructure Engineer

Avride Inc. • Austin (TX)

On-site
USD 140,000 - 190,000
Engineering Manager: GPU-Accelerated AI Inference
Engineering Manager: GPU-Accelerated AI Inference

NVIDIA • Illinois

On-site
USD 224,000 - 431,250
Equity
Benefits
Senior AI Compiler Engineer - MLIR, GPU Inference, Equity
Senior AI Compiler Engineer - MLIR, GPU Inference, Equity

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits package
Senior DL Inference Engineer — GPU-Optimized LLMs (Remote)
Senior DL Inference Engineer — GPU-Optimized LLMs (Remote)

NVIDIA Corporation • Northern (KY)

Hybrid
USD 152,000 - 288,000
Senior DL Inference Engineer — GPU-Accelerated AI, Equity
Senior DL Inference Engineer — GPU-Accelerated AI, Equity

NVIDIA Gruppe • California (MO)

On-site
USD 152,000 - 288,000
Senior AI Compiler Engineer (MLIR) — Equity & Impact
Senior AI Compiler Engineer (MLIR) — Equity & Impact

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
INFERENCE OPTIMIZATION ENGINEER
INFERENCE OPTIMIZATION ENGINEER

Up Top • United States

On-site
USD 180,000 - 320,000
ML Systems Engineer: AI Infra & GPU Acceleration
ML Systems Engineer: AI Infra & GPU Acceleration

Meta • San Francisco (CA)

On-site
USD 180,000 - 240,000
Bonus
Equity
Senior ML Inference Engineer: High-Performance GPU Systems
Senior ML Inference Engineer: High-Performance GPU Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
ML Infra Tech Lead: High-Performance GPU & K8s
ML Infra Tech Lead: High-Performance GPU & K8s

Reducto • Santa Fe (NM)

On-site
USD 190,000 - 230,000
Unlimited PTO
Daily Lunch
Commuter Reimbursement
+3