Founding Real-Time AI Platform Engineer

uRun

San Francisco (CA)

On-site

USD 140,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary and equity
Full health, dental, and vision coverage
401(k) retirement savings
Flexible spending accounts
Paid time off
Access to top-tier AI tools
Hardware such as MacBook Pro and AirPods

Job summary

uRun, located in San Francisco, is looking for a skilled engineer to design and operate a scalable low-latency infrastructure for real-time AI inference. You will build platforms that support interactive workloads, handling GPU-constrained and bursty scenarios while ensuring reliability and performance at scale.

With over 7 years of engineering experience, you should have deep Kubernetes expertise and a strong background in cloud services. This role offers competitive salary, health benefits, and equity in an early-stage AI company.

Qualifications

  • 7+ years as an engineer with experience in large-scale production systems.
  • Deep understanding of Kubernetes, particularly with GPU-heavy clusters.
  • Solid knowledge of cloud services like AWS, GCP, or Azure.

Responsibilities

  • Design, operate, and evolve the cloud-native platform for real-time inference.
  • Ensure reliability and performance at scale with observability standards.
  • Collaborate with ML teams to optimize workflows for low-latency tasks.

Skills

Cloud-native platform design
Kubernetes expertise
Observability practices
Latency optimization
Python/Go/TypeScript programming
Mentoring
GPU-optimized workloads
Real-time systems

Tools

Prometheus
Grafana
AWS
GCP
Azure
Terraform
Pulumi

Job description

uRun, located in San Francisco, is looking for a skilled engineer to design and operate a scalable low-latency infrastructure for real-time AI inference. You will build platforms that support interactive workloads, handling GPU-constrained and bursty scenarios while ensuring reliability and performance at scale.

With over 7 years of engineering experience, you should have deep Kubernetes expertise and a strong background in cloud services. This role offers competitive salary, health benefits, and equity in an early-stage AI company.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding ML Inference Performance Engineer
Founding ML Inference Performance Engineer

uRun • San Francisco (CA)

On-site
USD 120,000 - 160,000
Health, dental, and vision
401(k) participation
Flexible spending accounts
+3
Founding Real-Time AI Engineer — Healthcare
Founding Real-Time AI Engineer — Healthcare

Stealth AI Infrastructure Startup • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
401(k)
Founding Engineer - Site Reliability
Founding Engineer - Site Reliability

uRun • San Francisco (CA)

On-site
USD 120,000 - 140,000
Health, dental, and vision – full coverage
401(k) – company-supported retirement savings
FSA/HSA – flexible spending accounts
+3
Founding AI Infra Engineer: GPU Clusters, Kubernetes
Founding AI Infra Engineer: GPU Clusters, Kubernetes

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 300,000
Cash bonus
Founding engineer equity
Benefits
Member of Technical Staff- Distributed Systems
Member of Technical Staff- Distributed Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 350,000
Equity
Founding SRE for Interactive AI Inference Infra
Founding SRE for Interactive AI Inference Infra

uRun • San Francisco (CA)

On-site
USD 120,000 - 140,000
Health, dental, and vision – full coverage
401(k) – company-supported retirement savings
FSA/HSA – flexible spending accounts
+3
Software Engineer, AI Infrastructure
Software Engineer, AI Infrastructure

Harell Data • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Senior Platform Engineer - Cloud, AI & Distributed Systems
Senior Platform Engineer - Cloud, AI & Distributed Systems

Scale AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Engineer, GPU AI Inference & RL Infrastructure
Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, and vision insurance
Fully paid parental leave
+2
Principal Platform Engineer
Principal Platform Engineer

European Recruitment BV • United States

On-site
USD 150,000 - 200,000