Distributed Systems Engineer for Scalable AI Platform

Fal

San Francisco (CA)

On-site

USD 180,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Remote work options for senior levels
Visa sponsorship and relocation
Health, dental, and vision insurance
Equity and competitive salary

Job summary

fal in San Francisco is seeking an experienced Software Engineer, Distributed Systems to build large-scale platform components in Python or Rust. You will own request routing, AI workload orchestration, scheduling, and GPU autoscaling for a rapidly growing service serving millions of users.

This full-time in-person role offers compelling compensation and visa sponsorship, with remote options considered for Senior and Staff levels.

Qualifications

  • 3+ years experience building distributed compute and orchestration platforms in Python or Rust.
  • Strong understanding of distributed systems fundamentals: consensus, scheduling, fault tolerance, capacity planning.
  • Deep understanding of computational complexity and memory allocation.
  • Track record of designing systems that scale under real production load.
  • Experience building and using observability to drive performance and reliability decisions.
  • Excellent communication and ability to drive technical decisions across teams.
  • Self‑starter who executes quickly, takes ownership, and constantly seeks improvement.

Responsibilities

  • Build our core Python/Rust platform: request routing, AI workload orchestration, scheduling, GPU autoscaling, large‑scale file storage, queueing, etc.
  • Produce forward designs for platform evolution as we scale to 100x current traffic and need to provide low latency across the world.
  • Leverage AI to an extreme level to automate the mundane parts of building complex but reliable systems.
  • Profile and tune low‑level CPU and memory performance.

Skills

Distributed systems
Python
Rust
Observability
Communication
Ownership

Tools

Async runtimes
Zero-copy
Memory-safe concurrency

Job description

fal in San Francisco is seeking an experienced Software Engineer, Distributed Systems to build large-scale platform components in Python or Rust. You will own request routing, AI workload orchestration, scheduling, and GPU autoscaling for a rapidly growing service serving millions of users.

This full-time in-person role offers compelling compensation and visa sponsorship, with remote options considered for Senior and Staff levels.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Distributed Systems
Software Engineer, Distributed Systems

Fal • San Francisco (CA)

On-site
USD 180,000 - 250,000
Remote work options for senior levels
Visa sponsorship and relocation
Health, dental, and vision insurance
+1
Software Engineer
Software Engineer

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 350,000
Staff Distributed Systems Engineer — AI Infra, Open Source
Staff Distributed Systems Engineer — AI Infra, Open Source

TrustIn • San Francisco (CA)

On-site
USD 400,000 - 500,000
Distributed Systems Engineer — Open Platform, Equity
Distributed Systems Engineer — Open Platform, Equity

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health insurance
Paid parental leave
+2
Senior Platform Engineer - Cloud, AI & Distributed Systems
Senior Platform Engineer - Cloud, AI & Distributed Systems

Scale AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Infrastructure Engineer - Rust/C++ (Distributed Systems)
Senior Infrastructure Engineer - Rust/C++ (Distributed Systems)

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Relocation assistance
Hybrid work model
Member of Technical Staff - Distributed Systems
Member of Technical Staff - Distributed Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Staff Distributed Systems Engineer - AI Infrastructure
Staff Distributed Systems Engineer - AI Infrastructure

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Distributed AI Infrastructure Engineer
Distributed AI Infrastructure Engineer

Dedalus Labs • San Francisco (CA)

On-site
USD 180,000 - 260,000
Visa sponsorship
Relocation support
Equity
+1
Systems Engineer - Distributed AI Infra, Equity & Visa
Systems Engineer - Distributed AI Infra, Equity & Visa

Dedalus Labs, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive salary
Meaningful equity
Meals and office benefits
+2