Get more replies from employers
Send a job-specific resume in minutes.
Perplexity is building a unified, self-serve platform to run training and inference workloads across multiple clouds. You will own the Kubernetes layer for GPU orchestration and drive platform reliability and scalability.
You will design and operate a multi-cloud GPU fleet, implement scheduling to maximize capacity, and ensure low-latency inference while long-running training jobs run smoothly on the same fleet. This role empowers teams to push AI workloads with minimal GPU provisioning.
Perplexity is building a unified, self-serve platform to run training and inference workloads across multiple clouds. You will own the Kubernetes layer for GPU orchestration and drive platform reliability and scalability.
You will design and operate a multi-cloud GPU fleet, implement scheduling to maximize capacity, and ensure low-latency inference while long-running training jobs run smoothly on the same fleet. This role empowers teams to push AI workloads with minimal GPU provisioning.