Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
River AI Inc. is seeking exceptional inference systems engineers to build engines that serve large models through the River API.
You will own the serving runtime, from request scheduling to distributed model execution, with a focus on latency, throughput, reliability, and cost. You will work with GPU kernel engineers, researchers, and infra engineers to bring models to production, optimize performance, and ensure model-version consistency across deployments.
River AI Inc. is seeking exceptional inference systems engineers to build engines that serve large models through the River API.
You will own the serving runtime, from request scheduling to distributed model execution, with a focus on latency, throughput, reliability, and cost. You will work with GPU kernel engineers, researchers, and infra engineers to bring models to production, optimize performance, and ensure model-version consistency across deployments.