Get more replies from employers
Send a job-specific resume in minutes.
Modal is seeking a highly skilled researcher to own end-to-end inference bets for our LLM serving platform. You will collaborate with the research lead to prioritize initiatives that reduce cost per token and tail latency, while working directly with customers and deployment engineers.
You will drive collaborations with external labs and shape the research agenda, turning frontier serving techniques into practical products and scalable solutions for production workloads.
Modal is seeking a highly skilled researcher to own end-to-end inference bets for our LLM serving platform. You will collaborate with the research lead to prioritize initiatives that reduce cost per token and tail latency, while working directly with customers and deployment engineers.
You will drive collaborations with external labs and shape the research agenda, turning frontier serving techniques into practical products and scalable solutions for production workloads.