Get more replies from employers
Send a job-specific resume in minutes.
RadixArk is seeking a Member of Technical Staff to accelerate LLM inference and training on modern GPUs. You will profile, optimize, and extend SGLang and Miles across production workloads, collaborating with partner teams to push the performance envelope and broaden model support.
You will work on kernel development, day-0 model support, and enablement for new silicon, while contributing to the roadmap and ecosystem integrations that power scalable AI inference and training.
RadixArk is seeking a Member of Technical Staff to accelerate LLM inference and training on modern GPUs. You will profile, optimize, and extend SGLang and Miles across production workloads, collaborating with partner teams to push the performance envelope and broaden model support.
You will work on kernel development, day-0 model support, and enablement for new silicon, while contributing to the roadmap and ecosystem integrations that power scalable AI inference and training.