Get more replies from employers
Send a job-specific resume in minutes.
Veeda AI in Toronto is seeking a Member of Technical Staff - ML Performance to own step time and model throughput for multi-node video world model training, optimize tensor usage, and implement advanced parallelism.
You will write and tune CUDA and Triton kernels, improve precision and numerical stability, and build fault diagnostics and elastic checkpointing to minimize downtime during interruptions.
Veeda AI in Toronto is seeking a Member of Technical Staff - ML Performance to own step time and model throughput for multi-node video world model training, optimize tensor usage, and implement advanced parallelism.
You will write and tune CUDA and Triton kernels, improve precision and numerical stability, and build fault diagnostics and elastic checkpointing to minimize downtime during interruptions.