Lead AI Inference & Hardware Acceleration Engineer

Figureai

San Jose (CA)

On-site

USD 180,000 - 275,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Figureai in San Jose, CA, is seeking a Staff AI Inference & Acceleration Engineer to own the on-board inference architecture for humanoid robots. The role involves mapping AI workloads efficiently across compute hardware while driving optimization to meet strict latency and reliability demands.

Successful candidates will have a strong background in hardware acceleration and ML systems, with at least 8 years of industry experience. Key responsibilities include optimizing models and defining inference architectures.

Qualifications

  • 8+ years of experience in hardware acceleration and compute architecture.
  • Deep understanding of AI/ML inference model formats and deployment pipelines.
  • Hands-on experience optimizing embedded hardware models.

Responsibilities

  • Own and define the on-board inference architecture for humanoid robots.
  • Optimize inference toolchains for target hardware.
  • Partner with teams to ensure hardware-friendly model architecture.

Skills

Hardware acceleration
ML systems
C++
Python
AI/ML inference
Computational architecture

Education

M.S. or Ph.D. in Computer Engineering, Electrical Engineering, Computer Science or related field

Tools

TVM
MLIR
TensorRT
Torch
CUDA

Job description

Figureai in San Jose, CA, is seeking a Staff AI Inference & Acceleration Engineer to own the on-board inference architecture for humanoid robots. The role involves mapping AI workloads efficiently across compute hardware while driving optimization to meet strict latency and reliability demands.

Successful candidates will have a strong background in hardware acceleration and ML systems, with at least 8 years of industry experience. Key responsibilities include optimizing models and defining inference architectures.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff AI Inference & Acceleration Architect
Staff AI Inference & Acceleration Architect

Figure • San Jose (CA)

On-site
USD 180,000 - 275,000
Staff AI Inference and Acceleration Engineer
Staff AI Inference and Acceleration Engineer

Figure • San Jose (CA)

On-site
USD 180,000 - 275,000
Staff AI Inference and Acceleration Engineer
Staff AI Inference and Acceleration Engineer

Figureai • San Jose (CA)

On-site
USD 180,000 - 275,000
Lead AI Inference Performance Architect
Lead AI Inference Performance Architect

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 160,000
Opportunity to shape next-generation AI inference infrastructure
High-impact technical ownership
Work in a fast-moving engineering environment
Staff AI Systems Engineer — Inference & RL
Staff AI Systems Engineer — Inference & RL

Together • San Francisco (CA)

On-site
USD 200,000 - 280,000
Health insurance
Startup equity
Competitive benefits
Next-Gen AI Accelerator Architect for Inference
Next-Gen AI Accelerator Architect for Inference

The Consensus • San Jose (CA)

On-site
USD 190,000 - 280,000
Medical, dental, and vision packages
Housing subsidy/relo support
Daily lunch + dinner
+1
Video Foundation AI Engineer for Humanoid Autonomy
Video Foundation AI Engineer for Humanoid Autonomy

Figureai • San Jose (CA)

On-site
USD 120,000 - 160,000
Staff Engineer — Inference Runtime Lead
Staff Engineer — Inference Runtime Lead

Anthropic • New York (NY)

Hybrid
USD 405,000 - 485,000
Staff Engineer - Customer-Facing AI Inference Infra
Staff Engineer - Customer-Facing AI Inference Infra

Simplify • San Francisco (CA)

On-site
USD 200,000 - 300,000
Housing stipend
Uber/Waymo rides
Inference Engineer, Robotics
Inference Engineer, Robotics

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Relocation assistance
Hybrid work model (3 days in office)