An application made for this job — a tailored resume and cover letter that speak straight to the posting.
OpenAI is seeking a role focused on modeling inference performance across application, model, and fleet layers. You will build cost-to-serve estimates from microbenchmarks and create tools that help teams reason about latency, capacity, utilization, and cost tradeoffs.
The role emphasizes deep expertise in profiling, benchmarking, and optimization, with collaboration across engineering and research to implement concrete improvements in production systems.
OpenAI is seeking a role focused on modeling inference performance across application, model, and fleet layers. You will build cost-to-serve estimates from microbenchmarks and create tools that help teams reason about latency, capacity, utilization, and cost tradeoffs.
The role emphasizes deep expertise in profiling, benchmarking, and optimization, with collaboration across engineering and research to implement concrete improvements in production systems.