Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Hedra in San Francisco is seeking an Inference Optimization Engineer to join our research-and-systems team. You will push state-of-the-art visual models to run faster at inference time across single and multi-GPU environments.
Role spans algorithmic and systems optimizations, including quantization and advanced memory techniques, with close collaboration with researchers and engineers to deploy scalable solutions.
Hedra in San Francisco is seeking an Inference Optimization Engineer to join our research-and-systems team. You will push state-of-the-art visual models to run faster at inference time across single and multi-GPU environments.
Role spans algorithmic and systems optimizations, including quantization and advanced memory techniques, with close collaboration with researchers and engineers to deploy scalable solutions.