Software Engineer, AI Infra

Sarah Smith Fund

Palo Alto (CA)

Hybrid

USD 140,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Equity
Health benefits
Company retreats

Job summary

The Sarah Smith Fund in Palo Alto is seeking a Staff/Lead Software Engineer to build core infrastructure for AI capabilities. You will design GPU infrastructure, lead the technical direction, and mentor engineers.

Applicants must have over 5 years of experience in systems infrastructure and proficiency in languages such as Python. The role offers competitive salary, equity, and comprehensive benefits while fostering a collaborative office environment.

Qualifications

  • 5+ years experience as a software engineer in systems infrastructure with ML serving and GPU orchestration.
  • Deep knowledge of distributed systems and cloud-native infrastructure.
  • Proficiency in backend languages like Python, Go, or C++.

Responsibilities

  • Design and maintain scalable GPU infrastructure for AI models.
  • Optimize APIs for AI model serving.
  • Lead GPU resource orchestration for high-performance applications.

Skills

Systems infrastructure
GPU orchestration
Distributed systems
API optimization
Collaboration

Tools

Kubernetes
AWS
TensorFlow Serving

Job description

About the Role

We are looking for a Staff/Lead Software Engineer, AI Infrastructure, to play a critical role in building and scaling the core infrastructure that powers Pika’s AI capabilities. In this position, you will lead the design and implementation of GPU infrastructure, AI model serving APIs, and general AI infrastructure execution—enabling cutting‑edge machine learning features that drive our products.

You will be responsible for architecting robust, distributed systems optimized for high‑performance AI workloads, large‑scale GPU orchestration, and low‑latency, reliable API serving. Your work will directly impact the way users experience and interact with generative AI at scale. As a senior technical leader, you’ll also mentor engineers, drive best practices, and set the technical vision for AI infrastructure at Pika.

What You’ll Do
  • Design, develop, and maintain scalable GPU infrastructure for training and serving state‑of‑the‑art AI models
  • Architect and optimize high‑throughput, low‑latency APIs for AI model serving and inference
  • Lead the orchestration, scheduling, and efficient utilization of heterogeneous GPU resources across clusters
  • Build and support robust systems for model deployment, monitoring, scaling, and reliability in production environments
  • Collaborate with ML, backend, and platform engineering teams to deliver seamless AI‑powered product features
  • Drive technical direction, code reviews, and mentorship across the AI Infrastructure team
What We’re Looking For
  • Strong experience (5+ years) as a software engineer working on systems infrastructure, including hands‑on work with ML serving and GPU orchestration
  • Deep knowledge of distributed systems, Kubernetes (or similar orchestration frameworks), and cloud‑native infrastructure (AWS/GCP/Azure)
  • Proven expertise in building and optimizing APIs for large‑scale AI model serving (TensorFlow Serving, Triton, TorchServe, or similar)
  • Familiarity with the challenges of high‑throughput, scalable GPU fleet management, scheduling, and efficient model execution
  • Proficiency in backend languages such as Python, Go, or C++, and experience optimizing for performance and reliability
  • Ownership mentality and the drive to solve complex problems independently in ambiguous, high‑growth environments
  • Excellent communication, collaborative, and mentorship skills
Nice to Have
  • Experience with multi‑modal AI model infrastructure (LLMs, generative models, video/image/speech models)
  • Background in building infra for multi‑tenant SaaS, enterprise AI/ML platforms, or operational automation at scale
  • Previous startup experience or experience leading high‑impact projects through ambiguity and rapid iteration
  • Experience with competitive coding or large‑scale distributed computing environments
What We Offer
  • Competitive salary in the AI industry
  • Equity in a rapidly growing team shaping the future of AI
  • Comprehensive health benefits, monthly stipends, and company retreats
  • A supportive and collaborative office culture—everyone builds, ships, and learns together
About Pika

At Pika, we’re building the infrastructure that empowers everyone to create videos and express ideas through advanced AI. Our team is passionate about removing technical barriers to creativity, and we thrive on working together to solve hard problems. We’re based in Palo Alto, CA, with a collaborative team working in‑office 3–5 days a week.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Backend
Software Engineer, Backend

Pika • Palo Alto (CA)

On-site
USD 130,000 - 160,000
Competitive salary in the AI industry
Equity in a rapidly growing startup
Comprehensive health benefits
+1
Software Engineer, Product
Software Engineer, Product

Pika • Palo Alto (CA)

On-site
USD 130,000 - 170,000
Competitive salary
Equity in startup
Comprehensive health benefits
+2
Staff Software Engineer (AI Infrastructure)
Staff Software Engineer (AI Infrastructure)

DeepRec.ai • Palo Alto (CA)

On-site
USD 180,000 - 320,000
Software Engineer, AI Infra
Software Engineer, AI Infra

Makers Fund • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Monthly stipends
+1
Research Scientist, Data
Research Scientist, Data

AI Chopping Block • Palo Alto (CA)

Hybrid
USD 120,000 - 160,000
Competitive salary
Substantial equity
Full health benefits
+1
Research Scientist, Data
Research Scientist, Data

Pika • Palo Alto (CA)

Hybrid
USD 120,000 - 160,000
Competitive salary
Full health benefits
401k matching
+1
Software Engineer, Product
Software Engineer, Product

Makers Fund • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Comprehensive health benefits
Monthly stipends
+1
Senior Software Engineer, Infrastructure Software for AI (Centralized AI Data Centers & Distrib[...]
Senior Software Engineer, Infrastructure Software for AI (Centralized AI Data Centers & Distrib[...]

Intelliswift - An LTTS Company • Sunnyvale (CA)

On-site
USD 120,000 - 150,000
Competitive salary
Health insurance
Flexible work hours
Senior Software Engineer, ML Infrastructure
Senior Software Engineer, ML Infrastructure

Arena • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive salary
Equity
Health benefits
+2
Member of Technical Staff - GPU Infrastructure
Member of Technical Staff - GPU Infrastructure

Prime Intellect • United States

On-site
USD 120,000 - 150,000