Get more replies from employers
Send a job-specific resume in minutes.
SpaceXAI is building a high-performance inference platform used by Grok at scale. As a Member of Technical Staff - Inference, you will design and optimize distributed infrastructure for model serving, including global KV caching, batching, and auto-scaling, with deep GPU kernel work and code generation.
You will own everything from distributed infrastructure to low-level optimizations, shaping how fast and reliably users interact with Grok across millions of requests per day.
SpaceXAI is building a high-performance inference platform used by Grok at scale. As a Member of Technical Staff - Inference, you will design and optimize distributed infrastructure for model serving, including global KV caching, batching, and auto-scaling, with deep GPU kernel work and code generation.
You will own everything from distributed infrastructure to low-level optimizations, shaping how fast and reliably users interact with Grok across millions of requests per day.