Get more replies from employers
Send a job-specific resume in minutes.
uRun, located in San Francisco, is looking for a skilled engineer to design and operate a scalable low-latency infrastructure for real-time AI inference. You will build platforms that support interactive workloads, handling GPU-constrained and bursty scenarios while ensuring reliability and performance at scale.
With over 7 years of engineering experience, you should have deep Kubernetes expertise and a strong background in cloud services. This role offers competitive salary, health benefits, and equity in an early-stage AI company.
uRun, located in San Francisco, is looking for a skilled engineer to design and operate a scalable low-latency infrastructure for real-time AI inference. You will build platforms that support interactive workloads, handling GPU-constrained and bursty scenarios while ensuring reliability and performance at scale.
With over 7 years of engineering experience, you should have deep Kubernetes expertise and a strong background in cloud services. This role offers competitive salary, health benefits, and equity in an early-stage AI company.