Inference Systems Engineer, High-Throughput AI Serving

Future Ventures

Palo Alto (CA)

On-site

USD 135,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive medical, vision, dental coverage
401(k) retirement plan
Paid parental leave
Paid vacation and holidays

Job summary

Future Ventures in Palo Alto is seeking an Application Software Engineer to develop scalable AI inference systems. The successful candidate will be responsible for architecture and optimization of distributed systems, working to deliver reliable performance.

A Bachelor's degree in computer science or similar, alongside proficiency in Rust or C++, is required. The position emphasizes problem-solving and teamwork within a fast-paced environment.

Qualifications

  • 2+ years of professional experience building software.
  • Experience in implementing scalable distributed systems.
  • 1+ years of experience with production systems.

Responsibilities

  • Develop high-throughput inference systems.
  • Architect scalable distributed infrastructure.
  • Optimize model inference for low latency.

Skills

Distributed systems design
Full stack development
Rust programming
C++ programming

Education

Bachelor's degree in computer science or related field

Tools

Docker
Kubernetes
PostgreSQL
MongoDB

Job description

Future Ventures in Palo Alto is seeking an Application Software Engineer to develop scalable AI inference systems. The successful candidate will be responsible for architecture and optimization of distributed systems, working to deliver reliable performance.

A Bachelor's degree in computer science or similar, alongside proficiency in Rust or C++, is required. The position emphasizes problem-solving and teamwork within a fast-paced environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Inference Engineer - High-Throughput Distributed Systems
AI Inference Engineer - High-Throughput Distributed Systems

United States Digital Space LLC • Palo Alto (CA)

On-site
USD 135,000 - 210,000
Stock options
Medical Insurance
Vision Insurance
+7
AI Inference Systems Engineer (High-Throughput, Low-Latency)
AI Inference Systems Engineer (High-Throughput, Low-Latency)

SpaceX • Palo Alto (CA)

On-site
USD 135,000 - 210,000
Stock options
Excellent medical coverage
401(k) plan
AI Inference Engineer - Scalable, Low-Latency Systems
AI Inference Engineer - Scalable, Low-Latency Systems

SPACE EXPLORATION TECHNOLOGIES CORP • Palo Alto (CA), Northern (KY)

Hybrid
USD 135,000 - 210,000
401(k)
Medical, vision and dental coverage
Paid parental leave
+4
Senior Inference Systems Engineer for High-Performance AI
Senior Inference Systems Engineer for High-Performance AI

Causal Labs • San Francisco (CA)

On-site
USD 150,000 - 210,000
Inference Systems Engineer for Transformers & Low-Latency HPC
Inference Systems Engineer for Transformers & Low-Latency HPC

Etched • San Jose (CA)

On-site
USD 180,000 - 240,000
Medical, dental, and vision packages
Housing subsidy of $2k per month
Relocation support for those moving to San Jose
+1
Senior AI Infra Engineer: High-Performance Inference
Senior AI Infra Engineer: High-Performance Inference

Ddn • Sacramento (CA)

On-site
USD 140,000 - 200,000
Staff Engineer, Inference Runtime — High-Performance AI Serving
Staff Engineer, Inference Runtime — High-Performance AI Serving

Anthropic • Seattle (WA)

Hybrid
USD 405,000 - 485,000
Staff Inference Systems Engineer — High-Throughput AI
Staff Inference Systems Engineer — High-Throughput AI

Kindredventures • San Francisco (CA)

On-site
USD 180,000 - 240,000
Staff Engineer, Scalable AI Inference Infrastructure
Staff Engineer, Scalable AI Inference Infrastructure

Inferact • San Francisco (CA)

Hybrid
USD 200,000 - 400,000
Staff Software Engineer - Real-Time AI Inference Infra
Staff Software Engineer - Real-Time AI Inference Infra

Cerebras Systems, Inc. • Sunnyvale (CA)

On-site
USD 110,000 - 140,000