Stand out for this role — generate a tailored resume and cover letter in about a minute.
NVIDIA is seeking an experienced Site Reliability Engineer to lead the reliability roadmap for AI Platform Runtime and enterprise systems. You will architect highly available platforms, drive automation, and mentor senior engineers across multiple teams.
The role requires deep expertise in distributed systems, Kubernetes, and cloud platforms, with a focus on observability and incident-driven improvements. Hybrid work and equity opportunities are offered.
NVIDIA is seeking an experienced Site Reliability Engineer to lead the reliability roadmap for AI Platform Runtime and enterprise systems. You will architect highly available platforms, drive automation, and mentor senior engineers across multiple teams.
The role requires deep expertise in distributed systems, Kubernetes, and cloud platforms, with a focus on observability and incident-driven improvements. Hybrid work and equity opportunities are offered.