Get more replies from employers
Send a job-specific resume in minutes.
OpenAI is seeking a systems-focused engineer to design and implement the LLM inference runtime for frontier models on our custom silicon. You'll bridge model execution with the hardware, shaping how workloads map onto the platform and how production workloads achieve high throughput and low latency.
You will collaborate across model, systems, compiler, kernel, and hardware teams to build a scalable, observable runtime and to deliver reliable, production-grade performance on OpenAI's AI
OpenAI is seeking a systems-focused engineer to design and implement the LLM inference runtime for frontier models on our custom silicon. You'll bridge model execution with the hardware, shaping how workloads map onto the platform and how production workloads achieve high throughput and low latency.
You will collaborate across model, systems, compiler, kernel, and hardware teams to build a scalable, observable runtime and to deliver reliable, production-grade performance on OpenAI's AI