Get more replies from employers
Send a job-specific resume in minutes.
Morph is looking for candidates with exceptional skills in infrastructure to manage high-uptime systems focused on GPU loads. The role requires deep knowledge of inference engines and experience with various infrastructures, including load balancers.
Your contributions to projects like vLLM or SGLang will be a valuable asset. Join us in tackling complex challenges in custom inference stacks.
Goal: 99.99% uptime
We serve custom inference stacks that have irregular GPU load.
We're looking for people that have done genuinely amazing work in infrastructure and are interested in a challenge, working with both traditional infrastructure such as load balancers, NLB, etc., as well as very different infrastructure around inference engines and GPU loads.
This is a role that will inherently require deep experience with inference engines.
Contributions to vLLM, SGLang, trtllm, or inference frameworks a plus.