ML Infra Engineer — Scale AI Platform with Kubernetes

Physical Intelligence

San Francisco (CA)

On-site

USD 180,000 - 210,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Physical Intelligence in San Francisco is seeking a seasoned Platform Infrastructure Engineer to own and scale AI-native infrastructure. You will operate Kubernetes clusters, build a scalable microservice platform, and drive security, observability, and cost-aware design across multi-cloud environments.

You will collaborate with researchers and engineers to translate fast-moving needs into reusable infrastructure, owning systems end-to-end from design to operation.

Qualifications

  • Strong first-principles thinking and ownership.
  • Experience with cloud platforms (GCP, AWS) and distributed systems.
  • In-depth experience with infrastructure-as-code and containerization.

Responsibilities

  • Own and scale AI-native infrastructure: operate and evolve Kubernetes clusters and service deployment patterns, support safe rollouts, upgrades, and rollback strategies.
  • Drive security, observability and cost-aware infrastructure, surface reliability and performance issues early, improve cost visibility.
  • Harden platform foundations across multi-cloud, design authentication/authorization flows, networking, quota and rate-limiting services.
  • Improve developer experience with clear interfaces and self-serve infrastructure usage.
  • Collaborate with researchers and engineers to own systems end-to-end from design to operation.

Skills

First-principles thinking
Agent infrastructure
Cloud platforms (GCP, AWS)
Distributed systems
Kubernetes
Observability
Security
Sandboxing
Cost optimization
Cross-functional communication

Tools

Kubernetes
Terraform
Docker
GCP
AWS
CI/CD

Job description

Physical Intelligence in San Francisco is seeking a seasoned Platform Infrastructure Engineer to own and scale AI-native infrastructure. You will operate Kubernetes clusters, build a scalable microservice platform, and drive security, observability, and cost-aware design across multi-cloud environments.

You will collaborate with researchers and engineers to translate fast-moving needs into reusable infrastructure, owning systems end-to-end from design to operation.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Infra Platform Engineer — Scale Cloud & Kubernetes
Infra Platform Engineer — Scale Cloud & Kubernetes

LlamaIndex • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Competitive base salary and equity compensation
Comprehensive benefits coverage
Unlimited paid time off
+1
Senior Platform Engineer, AI Scale & Cloud Foundations
Senior Platform Engineer, AI Scale & Cloud Foundations

Scale AI, Inc. • San Francisco (CA)

On-site
USD 216,000 - 270,000
Equity
Health, dental & vision
Retirement benefits
+3
ML Infra Engineer: Scale & Optimize Large-Scale Training
ML Infra Engineer: Scale & Optimize Large-Scale Training

Physical Intelligence • San Francisco (CA)

On-site
USD 180,000 - 240,000
Staff Infra Engineer: AI-Scale Cloud & Kubernetes
Staff Infra Engineer: AI-Scale Cloud & Kubernetes

LlamaIndex, Inc. • San Francisco (CA)

Hybrid
USD 200,000 - 275,000
Competitive base salary and equity
Comprehensive medical/dental/vision
Unlimited paid time off
+1
Senior ML Platform Engineer — Scale Research ML Infra
Senior ML Platform Engineer — Scale Research ML Infra

techire ai • San Francisco (CA)

On-site
USD 270,000 - 330,000
Stock options
AI Platform Engineer — Scale Cloud Infra, Equity
AI Platform Engineer — Scale Cloud Infra, Equity

Harrison Clarke • San Francisco (CA)

On-site
USD 150,000 - 210,000
Platform Engineer – AI Infra, CI/CD & Observability
Platform Engineer – AI Infra, CI/CD & Observability

Outmarket AI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Platform Engineer: AI Infra for Hybrid Cloud & Kubernetes
Platform Engineer: AI Infra for Hybrid Cloud & Kubernetes

Meibel • Washington

On-site
USD 120,000 - 180,000
ML Infra Engineer, Platform
ML Infra Engineer, Platform

Physical Intelligence • San Francisco (CA)

On-site
USD 180,000 - 210,000
ML Infra Engineer — Real-Time, Scalable Data Systems
ML Infra Engineer — Real-Time, Scalable Data Systems

Arena Intelligence, Inc. • San Francisco (CA)

On-site
USD 180,000 - 230,000
Competitive compensation & equity
Health benefits
Cutting-edge AI work
+1