An application made for this job — a tailored resume and cover letter that speak straight to the posting.
OpenAI's robotics team seeks a GPU Inference Engineer in San Francisco to advance model serving efficiency and inference performance. You will drive system optimizations, build serving infrastructure, and partner with researchers to scale multimodal workloads.
Ideal candidates combine deep expertise in inference optimization and low-level data movement with a collaborative approach to extending AI capabilities for real-world robotic applications. This role is hybrid with relocation assistance.
Our Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples' lives.
We're looking for a GPU Inference Engineer to contribute to improvements in model serving efficiency for our Robotics research. This is a high-impact role where you'll drive initiatives to optimize inference performance and scalability. You'll also be engaged in model design, to help assist our researchers in developing inference-friendly models. This role is critical to scaling the team's broader goals - it will directly enable leadership to focus on higher-leverage initiatives by building a stronger technical foundation. In this role you will:
This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees.
OpenAI i