Stand out for this role — generate a tailored resume and cover letter in about a minute.
Sciforium is seeking a distributed training and inference engineer to build and optimize the ML software stack for large-scale AI workloads. You will work across CUDA/ROCm runtimes to high-level frameworks like JAX and PyTorch to ensure fast, scalable training and serving.
This role emphasizes deep systems engineering, debugging hardware–software interactions, and optimizing performance at every layer of the ML stack, enabling training and deployment of next‑gen LLMs and generative AI models.
Sciforium is seeking a distributed training and inference engineer to build and optimize the ML software stack for large-scale AI workloads. You will work across CUDA/ROCm runtimes to high-level frameworks like JAX and PyTorch to ensure fast, scalable training and serving.
This role emphasizes deep systems engineering, debugging hardware–software interactions, and optimizing performance at every layer of the ML stack, enabling training and deployment of next‑gen LLMs and generative AI models.