Get more replies from employers
Send a job-specific resume in minutes.
Physical Intelligence in San Francisco is building the core ML infrastructure to scale training from prototype to production-grade runs. The ML Infrastructure team owns training/inference systems, scheduling, checkpointing, and metrics collection.
You will scale JAX-based training across TPU and GPU clusters, profile memory usage, improve throughput, and create abstractions for launching and monitoring experiments in close collaboration with researchers.
Physical Intelligence in San Francisco is building the core ML infrastructure to scale training from prototype to production-grade runs. The ML Infrastructure team owns training/inference systems, scheduling, checkpointing, and metrics collection.
You will scale JAX-based training across TPU and GPU clusters, profile memory usage, improve throughput, and create abstractions for launching and monitoring experiments in close collaboration with researchers.