Stand out for this role — generate a tailored resume and cover letter in about a minute.
Genesis AI is seeking a senior engineer to build and optimize real-time inference pipelines for on-device and cluster deployment. You will design GPU-accelerated systems, implement CUDA and Triton kernels, and ensure efficient memory and compute scheduling across heterogeneous stacks.
You’ll work on throughput- and latency-optimized workloads, with monitoring and debugging tools to diagnose regressions rapidly.
Genesis AI is seeking a senior engineer to build and optimize real-time inference pipelines for on-device and cluster deployment. You will design GPU-accelerated systems, implement CUDA and Triton kernels, and ensure efficient memory and compute scheduling across heterogeneous stacks.
You’ll work on throughput- and latency-optimized workloads, with monitoring and debugging tools to diagnose regressions rapidly.