Get more replies from employers
Send a job-specific resume in minutes.
AMD is looking for a performance-obsessed engineer to drive AI inference performance to the absolute limit on AMD GPUs, with SGLang as the primary serving framework. You will lead a small, highly technical team and work end-to-end across the stack: profiling, diagnosing, and optimizing leading models running on SGLang across customer-relevant serving configurations (e.g.
agentic coding, long-context, high-throughput serving).
AMD is looking for a performance-obsessed engineer to drive AI inference performance to the absolute limit on AMD GPUs, with SGLang as the primary serving framework. You will lead a small, highly technical team and work end-to-end across the stack: profiling, diagnosing, and optimizing leading models running on SGLang across customer-relevant serving configurations (e.g.
agentic coding, long-context, high-throughput serving).