Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Baseten powers mission-critical AI inference at scale. We seek distributed systems engineers to own the runtimes and APIs that run large-scale LLM inference and hosted endpoints.
You’ll work across the stack from developer tooling to orchestration on Kubernetes, driving speed, reliability, and cost-efficiency for customers. You’ll collaborate with the team on model serving, inference performance, and API design while shipping robust, production-grade infrastructure and tools for a rapidly
Baseten powers mission-critical AI inference at scale. We seek distributed systems engineers to own the runtimes and APIs that run large-scale LLM inference and hosted endpoints.
You’ll work across the stack from developer tooling to orchestration on Kubernetes, driving speed, reliability, and cost-efficiency for customers. You’ll collaborate with the team on model serving, inference performance, and API design while shipping robust, production-grade infrastructure and tools for a rapidly