Turn this role into an interview — a resume and cover letter built around what this employer wants.
Runpod is hiring a ML Systems Engineer, Inference to lead the performance end-to-end of LLM serving. You will measure, diagnose, and improve latency and cost across models, hardware generations, and workloads in a remote-first environment.
You will ship fixes that directly affect customer experience, working hands-on on modern inference engines and optimization techniques, with opportunities to influence tooling and deployment strategies across Runpod’s platform.
Runpod is hiring a ML Systems Engineer, Inference to lead the performance end-to-end of LLM serving. You will measure, diagnose, and improve latency and cost across models, hardware generations, and workloads in a remote-first environment.
You will ship fixes that directly affect customer experience, working hands-on on modern inference engines and optimization techniques, with opportunities to influence tooling and deployment strategies across Runpod’s platform.