Get more replies from employers
Send a job-specific resume in minutes.
Thinking Machines Lab Inc. in San Francisco seeks an infrastructure research engineer to design, optimize, and maintain the compute foundations powering large-scale language model training.
You will develop high-performance ML kernels (CUDA, CuTe, Triton) and improve the distributed compute stack for scalable AI systems. You’ll collaborate with researchers and systems architects, prototype kernel implementations, and profile performance across hardware generations to shape numerical and
Thinking Machines Lab Inc. in San Francisco seeks an infrastructure research engineer to design, optimize, and maintain the compute foundations powering large-scale language model training.
You will develop high-performance ML kernels (CUDA, CuTe, Triton) and improve the distributed compute stack for scalable AI systems. You’ll collaborate with researchers and systems architects, prototype kernel implementations, and profile performance across hardware generations to shape numerical and