Stand out for this role — generate a tailored resume and cover letter in about a minute.
Deepgram is seeking an Embedded AI Engineer to work on the Partner Platform Engineering team, tackling the lowest layer of the edge stack. You will write and optimize custom kernels and operators for diverse hardware, including embedded SoCs, DSPs, and NPUs, enabling Deepgram models to run on non-NVIDIA accelerators.
Your work will involve quantization, operator fusion, and architecture-specific compilation, with collaboration to fit models to constrained devices and delivery of reusable runtime
Deepgram is seeking an Embedded AI Engineer to work on the Partner Platform Engineering team, tackling the lowest layer of the edge stack. You will write and optimize custom kernels and operators for diverse hardware, including embedded SoCs, DSPs, and NPUs, enabling Deepgram models to run on non-NVIDIA accelerators.
Your work will involve quantization, operator fusion, and architecture-specific compilation, with collaboration to fit models to constrained devices and delivery of reusable runtime