Get more replies from employers
Send a job-specific resume in minutes.
Designworks Talent is seeking an Inference Engineer to build and operate production-grade model-serving and inference systems in a hybrid Bellevue, WA setting. You’ll work across distributed systems, GPU optimization, and production AI infrastructure to deliver high-throughput, low-latency inference.
Join a fast-moving team shaping next-gen AI platforms with scalable tooling, monitoring, and reliability practices, collaborating with AI training and platform teams.
Designworks Talent is seeking an Inference Engineer to build and operate production-grade model-serving and inference systems in a hybrid Bellevue, WA setting. You’ll work across distributed systems, GPU optimization, and production AI infrastructure to deliver high-throughput, low-latency inference.
Join a fast-moving team shaping next-gen AI platforms with scalable tooling, monitoring, and reliability practices, collaborating with AI training and platform teams.