Stand out for this role — generate a tailored resume and cover letter in about a minute.
Acceler8 Talent in San Francisco seeks a Member of Technical Staff focusing on Kernels & GPU Performance to push the limits of production AI inference across diverse accelerators.
You will implement low-level kernels, analyze memory hierarchies, and collaborate with compiler, ML systems, and runtime teams to optimize latency and throughput.
This on-site role offers a full-time path at a fast-growing AI infrastructure company.
Acceler8 Talent in San Francisco seeks a Member of Technical Staff focusing on Kernels & GPU Performance to push the limits of production AI inference across diverse accelerators.
You will implement low-level kernels, analyze memory hierarchies, and collaborate with compiler, ML systems, and runtime teams to optimize latency and throughput.
This on-site role offers a full-time path at a fast-growing AI infrastructure company.