Senior Systems Software Engineer, Developer Productivity and Cloud Automation - GeForce NOW

NVIDIA

Santa Clara (CA)

On-site

USD 184,000 - 356,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Comprehensive benefits

Job summary

NVIDIA is hiring for a senior cloud infrastructure engineer to own a Kubernetes-native deployment platform and automation stack. You will build backend services and APIs powering multi-cloud, on-prem and bare-metal deployments, with zero-downtime rollouts and drift detection.

You will drive GitOps pipelines (Flux/Argo CD) across environments and develop CRDs and operators in Go for scale and security. The role demands deep Kubernetes expertise, experience with Terraform/Ansible/Vault, and strong

Qualifications

  • Bachelor's degree in computer science, engineering, or equivalent experience.
  • 10+ years of Cloud Infrastructure and DevOps experience with Kubernetes and GitOps.
  • Expert Kubernetes: CRDs, operators, multi-cluster, security hardening.
  • Proficient with Flux CD or Argo CD, GitLab CI, Jenkins; Go or Python for control plane.
  • Experience with Vault, Terraform, Ansible, and Helm in on-prem and cloud.

Responsibilities

  • Build and develop backend microservices and REST/gRPC/MCP APIs powering the platform.
  • Extend the platform with dynamic delivery, rollback, drift detection, and zone bootstrapping.
  • Create prod-like staging environments to catch regressions early.
  • Develop and sustain GitOps pipelines across on-prem, Nvidia GFN Cloud, AWS, Azure, GCP.
  • Develop Kubernetes CRDs and operators in Go for scheduling and auto-scaling.
  • Integrate CI/CD, observability, and automation into a unified platform.
  • Automate hardware and cloud configurations with Terraform, Ansible, and Vault.
  • Deploy monitoring with Prometheus, Grafana, Datadog, ELK; define SLOs and alerts.
  • Integrate runbooks, StackStorm bots, anomaly remediation, and Slack release bots.

Skills

Kubernetes
GitOps
Cloud infrastructure
Go
Python
CI/CD pipelines

Education

Bachelor's degree in Computer Science or related field

Tools

Terraform
Ansible
Vault
Prometheus
Grafana
ELK

Job description

What You'll Own

This is the Kubernetes-native deployment platform paired with its automation system and the backend services and APIs that support it. We deliver zero-downtime global rollouts with automatic drift detection. Our reliable staging environments replicate production accurately. We provide cloud automation for bare-metal and multi-cloud environments. If a GFN service ships, runs, or self-heals, you build the infrastructure behind it.

What You'll Be Doing:
  • Build and develop backend microservices and REST/gRPC/MCP APIs that power the deployment platform, zone reservation/lease system, and developer self-service tooling

  • Extend the platform with dynamic delivery, automatic rollback, drift detection, and automated zone bootstrapping

  • Build and stabilize prod-like staging environments/zones so teams catch regressions early

  • Develop and sustain GitOps pipelines (Flux CD, Argo CD) across on-premises and Nvidia GFN Cloud/AWS/Azure/GCP

  • Develop Kubernetes CRDs and operators in Go for scheduling, auto-scaling, and compliance across data centers

  • Build backend integrations and control-plane services connecting CI/CD, observability, and automation systems into a unified platform experience

  • Automate dedicated hardware and multiple cloud platform configurations using Terraform, Ansible, and Vault

  • Implement monitoring solutions including Prometheus, Grafana, Datadog, and ELK, paired with SLO/alerting for early detection

  • Integrate automation tools - runbooks, StackStorm bots, anomaly-triggered remediation, and Slack self-service release bots

What We Need to See:
  • Bachelor or higher degree in computer science, engineering, or equivalent experience.

  • 10+ years of experience in Cloud Infrastructure and DevOps, with deep expertise in Kubernetes, GitOps (or equivalent), and production-grade cloud-native CI/CD pipelines

  • Expert Kubernetes: CRDs, operators, multi-cluster management, and security hardening (CIS, PCI/SOC 2)

  • Proficiency in Flux CD or Argo CD, GitLab CI, and Jenkins; Go or Python for control plane development

  • Experience handling Vault, Terraform, Ansible, and Helm in both on-premises and cloud environments

  • Experience in developing and scaling RESTful, gRPC, MCP APIs and backend services.

  • Experience running hybrid multi-cloud and bare-metal at production scale, and owning a platform roadmap end-to-end.

Ways to Stand Out from the crowd:
  • Backstage or similar internal developer portals for self-service tooling

  • Comprehensive expertise covering backend services as well as platform and infrastructure engineering

  • Daily use of AI-assisted tools (Claude Code, GitHub Copilot, Cursor) - we use these daily

  • Hybrid infrastructure spanning on-premises GPU clusters and public cloud

With competitive salaries and a generous benefits package, NVIDIA is widely considered one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 30, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior System Software Engineer for Cloud – GeForce NOW, Senior System Software Engineer for Cl[...]
Senior System Software Engineer for Cloud – GeForce NOW, Senior System Software Engineer for Cl[...]

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Generous benefits package
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA Gruppe • United States

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer, DGX Cloud Production Engineering
Senior Software Engineer, DGX Cloud Production Engineering

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • Town of Texas (WI)

On-site
USD 168,000 - 322,000
Equity
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • Colorado

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • California (MO)

On-site
USD 170,000 - 322,000
Equity
Benefits
Senior Software Engineer - DGX Cloud
Senior Software Engineer - DGX Cloud

Thomas To • Seattle (WA)

On-site
USD 184,000 - 357,000
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • Massachusetts

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer - DGX Cloud
Senior Software Engineer - DGX Cloud

NVIDIA Gruppe • Seattle (WA)

On-site
USD 184,000 - 357,000
Equity
Benefits package
Performance bonuses
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 322,000