Get more replies from employers
Send a job-specific resume in minutes.
NVIDIA AI in Seattle is seeking a Lead performance and scalability analyst to optimize the Kubernetes-based accelerated runtime stack for large-scale AI workloads. You will collaborate with researchers and engineers to design automated workload tests and integrate performance testing into CI/CD pipelines.
The role focuses on performance, reliability, and scalability across GPU-accelerated platforms, requiring deep expertise in distributed systems and modern cloud-native tooling.
Lead performance and scalability analysis across the Kubernetes-based accelerated runtime stack, focusing on NVIDIA components. Collaborate with researchers and developers to design automated workload tests and integrate performance testing into CI/CD workflows.
Bachelor’s or Master’s degree in Engineering or equivalent experience, with at least 8 years in computer architecture and distributed systems. Expertise in Kubernetes and experience with large-scale AI workloads are essential.