Get more replies from employers
Send a job-specific resume in minutes.
NVIDIA is seeking a Sr Site Reliability Engineer in Santa Clara, CA to help operate large-scale GPU platforms and deploy robust infrastructure. You will contribute to deployments and daily operations, and bridge gaps between cluster operations and development.
Required: 8+ years in SRE or software roles, a CS degree or equivalent, and Python fluency. Linux networking is essential; Kubernetes and InfiniBand/Spectrum-X experience are a plus. Equity and benefits are offered.
NVIDIA is seeking a Sr Site Reliability Engineer in Santa Clara, CA to help operate large-scale GPU platforms and deploy robust infrastructure. You will contribute to deployments and daily operations, and bridge gaps between cluster operations and development.
Required: 8+ years in SRE or software roles, a CS degree or equivalent, and Python fluency. Linux networking is essential; Kubernetes and InfiniBand/Spectrum-X experience are a plus. Equity and benefits are offered.