Get more replies from employers
Send a job-specific resume in minutes.
Baseten, based in San Francisco, seeks a Software Engineer for the Inference Stack to build and operate large-scale LLM inference systems. You will own production-grade infrastructure, spanning deployment orchestration, routing, and observability, and collaborate with model performance engineers to push optimizations to customers.
You will work across the stack, from customer-facing features to low-level components, improving reliability, scalability, and developer experience while owning
Baseten, based in San Francisco, seeks a Software Engineer for the Inference Stack to build and operate large-scale LLM inference systems. You will own production-grade infrastructure, spanning deployment orchestration, routing, and observability, and collaborate with model performance engineers to push optimizations to customers.
You will work across the stack, from customer-facing features to low-level components, improving reliability, scalability, and developer experience while owning