Get more replies from employers
Send a job-specific resume in minutes.
Amazon in Seattle seeks a Senior Inference Engineer to own end-to-end real-time multimodal inference—from research to production—within strict latency budgets.
You will co-design architectures with scientists, optimize streaming serving, and build offline training/evaluation infra to support RL and deployment at scale.
Collaborate across hardware partners to minimize latency and cost while delivering human-like responsiveness.
Amazon in Seattle seeks a Senior Inference Engineer to own end-to-end real-time multimodal inference—from research to production—within strict latency budgets.
You will co-design architectures with scientists, optimize streaming serving, and build offline training/evaluation infra to support RL and deployment at scale.
Collaborate across hardware partners to minimize latency and cost while delivering human-like responsiveness.