Turn this role into an interview — a resume and cover letter built around what this employer wants.
Hedra is a small, highly technical team in San Francisco building state-of-the-art visual models and the infrastructure to deploy them at scale. We seek an Inference Optimization Engineer to push model performance at inference time, working at the boundary between research and systems.
You’ll collaborate with researchers to run new architectures efficiently on modern hardware, exploring novel optimization techniques and scalable execution across accelerators.
Hedra is a small, highly technical team in San Francisco building state-of-the-art visual models and the infrastructure to deploy them at scale. We seek an Inference Optimization Engineer to push model performance at inference time, working at the boundary between research and systems.
You’ll collaborate with researchers to run new architectures efficiently on modern hardware, exploring novel optimization techniques and scalable execution across accelerators.