An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Crusoe Energy Systems is seeking an engineering leader to optimize large language model inference workloads in production. You will profile, tune, and deploy end-to-end serving stacks, from frameworks like vLLM to CUDA kernels, collaborating with customers to meet latency and cost targets.
You’ll work hands-on with Python and C++, shipping reliable solutions and driving performance improvements across AI deployments in a customer-facing role.