Get more replies from employers
Send a job-specific resume in minutes.
SpaceXAI is building a high-performance inference platform used by Grok at scale. As a Member of Technical Staff - Inference, you will design and optimize distributed infrastructure for model serving, including global KV caching, batching, and auto-scaling, with deep GPU kernel work and code generation.
You will own everything from distributed infrastructure to low-level optimizations, shaping how fast and reliably users interact with Grok across millions of requests per day.
SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.