Get more replies from employers
Send a job-specific resume in minutes.
SpaceXAI is building a high-performance inference platform serving Grok at global scale. You will design and optimize distributed model-serving systems, owning from global KV caches to low-level GPU kernel work and code generation.
This role drives latency, throughput, and reliability for billions of users, with opportunities to push innovations in batching, quantization, and speculative decoding on next-gen hardware.
SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.
Responsibilities:
$180,000 - $440,000 USD
Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.
SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.