Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Veeda AI is seeking a highly skilled engineer to advance distributed training and high-performance inference for multimodal world models. You will own multi-node throughput, profile GPU utilization, and optimize kernels and CUDA/Triton code across NVLink fabrics.
Candidates should have deep PyTorch expertise, Python and C++/CUDA fluency, and experience with frameworks like FSDP2, Megatron-Core, DeepSpeed. You will work in a fast-moving research environment to push physical AI capabilities
Veeda AI is seeking a highly skilled engineer to advance distributed training and high-performance inference for multimodal world models. You will own multi-node throughput, profile GPU utilization, and optimize kernels and CUDA/Triton code across NVLink fabrics.
Candidates should have deep PyTorch expertise, Python and C++/CUDA fluency, and experience with frameworks like FSDP2, Megatron-Core, DeepSpeed. You will work in a fast-moving research environment to push physical AI capabilities