Turn this role into an interview — a resume and cover letter built around what this employer wants.
Designworks Talent LLC is seeking a Staff Inference Engineer to build and operate the model-serving systems powering a next-generation AI inference platform. You will work on high-throughput, low-latency inference for production-scale APIs in a hybrid Bellevue, WA setting.
The role emphasizes optimizing GPU-backed workloads, collaborating with training and platform teams, and contributing to scalable, reliable production services across diverse model architectures.
Designworks Talent LLC is seeking a Staff Inference Engineer to build and operate the model-serving systems powering a next-generation AI inference platform. You will work on high-throughput, low-latency inference for production-scale APIs in a hybrid Bellevue, WA setting.
The role emphasizes optimizing GPU-backed workloads, collaborating with training and platform teams, and contributing to scalable, reliable production services across diverse model architectures.