Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Baseten is hiring distributed systems engineers to build the runtime powering large-scale LLM inference. You’ll deploy and operate APIs and runtimes, focusing on performance, reliability, and cost efficiency across Kubernetes-based deployments.
You’ll own end-to-end projects, drive architectural decisions, and collaborate with performance engineers to deliver scalable solutions for modern AI workloads while enhancing developer experience and platform usability.
Baseten is hiring distributed systems engineers to build the runtime powering large-scale LLM inference. You’ll deploy and operate APIs and runtimes, focusing on performance, reliability, and cost efficiency across Kubernetes-based deployments.
You’ll own end-to-end projects, drive architectural decisions, and collaborate with performance engineers to deliver scalable solutions for modern AI workloads while enhancing developer experience and platform usability.