A complete application in a minute — tailored resume and cover letter, ready to send.
Genesis seeks a senior ML infrastructure engineer to design and optimize distributed training systems on multi-node GPU clusters. You will push wall-clock time to convergence by profiling, refining data pipelines, and low-level kernels in CUDA/CuDNN/Triton.
You will build scalable PyTorch-based workflows, balance CPU/GPU workloads, and develop monitoring tools to diagnose performance regressions at scale.
Genesis seeks a senior ML infrastructure engineer to design and optimize distributed training systems on multi-node GPU clusters. You will push wall-clock time to convergence by profiling, refining data pipelines, and low-level kernels in CUDA/CuDNN/Triton.
You will build scalable PyTorch-based workflows, balance CPU/GPU workloads, and develop monitoring tools to diagnose performance regressions at scale.