Software Engineer, Machine Learning-Backend

fal

San Francisco (CA)

On-site

USD 120,000 - 180,000

Full time

40 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Interesting and challenging work
Learning and growth opportunities
Health, dental, and vision insurance (
Regular team events and offsites

Job summary

fal in San Francisco is seeking a backend software engineer to build high-performance HTTP proxies and serverless endpoints that connect with 3rd party model providers. You’ll work on robust APIs with a focus on reliability and scalability.

You should have 3+ years of Python experience and a strong background in HTTP services, along with writing clean, well-tested code. This role offers growth, hands-on learning, and collaboration with a talented team.

Qualifications

  • 3+ years of demonstrated experience in building HTTP services with Python.
  • Experience designing, analyzing and improving efficiency, scalability, and stability of various system resources
  • Proficiency in version control practices and CI/CD pipelines.

Responsibilities

  • Identify, design, and develop foundational HTTP proxies and fal serverless endpoints for 3rd party model providers
  • Write clear, well-tested, and maintainable software
  • Analyze and improve the robustness and scalability of our existing proxies, APIs and fallback infrastructure
  • Conduct design and code reviews, create developer documentation, and develop testing strategies for robustness and fault tolerance

Skills

Python
HTTP services
Backend design

Tools

Git
CI/CD

Job description

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.

As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.

About this role:

This role is ideal for engineers who want to be on the forefront of the GenAI media revolution. Utilize your deep experience with backend APIs, robust http client and server design to build high-performance, reliable proxies to our partner model providers.

What you’ll do:

  • Identify, design, and develop foundational HTTP proxies and fal serverless endpoints for 3rd party model providers

  • Write clear, well-tested, and maintainable software

  • Analyze and improve the robustness and scalability of our existing proxies, APIs and fallback infrastructure

  • Conduct design and code reviews, create developer documentation, and develop testing strategies for robustness and fault tolerance

Qualifications:

  • 3+ years of demonstrated experience in building HTTP services with Python

  • Experience designing, analyzing and improving efficiency, scalability, and stability of various system resources

  • Proficiency in version control practices and CI/CD pipelines.

What we offer at fal:

  • Interesting and challenging work

  • A lot of learning and growth opportunities

  • Health, dental, and vision insurance (US)

  • Regular team events and offsites

U.S. EQUAL EMPLOYMENT OPPORTUNITY INFORMATION:

fal provides equal employment opportunities to applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, or any other classification protected by applicable law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer, Machine Learning-Backend
Software Engineer, Machine Learning-Backend

Speedrun Talent Network • San Francisco (CA)

On-site
USD 120,000 - 190,000
Health, dental, and vision insurance (
Software Engineer, Applied Machine Learning
Software Engineer, Applied Machine Learning

fal • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health, dental, and vision insurance (
Regular team events and offsites
Software Engineer, Applied Machine Learning
Software Engineer, Applied Machine Learning

Speedrun Talent Network • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health/Dental/Vision
Team offsites
Growth opportunities
Senior Software Engineer, Full Stack (Serverless)
Senior Software Engineer, Full Stack (Serverless)

Speedrun Talent Network • San Francisco (CA)

On-site
USD 140,000 - 180,000
Senior Software Engineer, Full Stack (Serverless)
Senior Software Engineer, Full Stack (Serverless)

fal • San Francisco (CA)

On-site
USD 150,000 - 190,000
Health, dental, and vision insurance
Regular team events
Learning and growth opportunities
Software Engineer, Distributed Systems
Software Engineer, Distributed Systems

fal - Features & Labels • United States

Remote
USD 140,000 - 190,000
Interesting work
Learning opportunities
Team offsites
Senior Software Engineer, Machine Learning Infrastructure & Automation
Senior Software Engineer, Machine Learning Infrastructure & Automation

fal • United States

Remote
USD 140,000 - 190,000
Health, dental, and vision insurance (
Software Engineer, Site Reliability
Software Engineer, Site Reliability

fal - Features & Labels • United States

Remote
USD 120,000 - 180,000
Product Marketer, Developer Platform
Product Marketer, Developer Platform

fal • San Francisco (CA)

On-site
USD 160,000 - 200,000
Visa sponsorship
Relocation to San Francisco
Health insurance
+3
Product Marketer, Infrastructure
Product Marketer, Infrastructure

fal • San Francisco (CA)

On-site
USD 160,000 - 200,000
Visa sponsorship and relocation to San
Health, dental, and vision insurance (