Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA

Massachusetts

On-site

USD 168,000 - 322,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Massachusetts is seeking a seasoned infrastructure engineer to help scale our global AI tooling. You will build and operate core services, design secure cloud-native platforms, and advance automation across our production environments.

We value deep expertise in Python and Go, Kubernetes, and distributed systems, plus strong Linux fundamentals, observability, and collaboration skills. This role offers equity and a path to impact at scale within a leading tech company.

Qualifications

  • BS or equivalent experience with 8+ years in the field.
  • Strong proficiency in Python and Go for production-quality software.
  • Experience building cloud-native microservices and APIs on Kubernetes (FastAPI, gRPC, REST).
  • Experience with infrastructure automation (Terraform, Ansible) and orchestration (Temporal).
  • Experience with distributed systems, databases, Redis, and messaging (Kafka, NATS, SQS).
  • Experience designing and operating production infrastructure services (DNS, NTP, RADIUS/OAuth, observability).
  • Strong Linux fundamentals; experience with observability, networking, and security.

Responsibilities

  • Build and operate core infrastructure services powering NVIDIA's global AI infra.
  • Architect secure, scalable cloud-native platform services.
  • Develop software enabling orchestration, workflows, and automation.
  • Own integrations to automate provisioning and lifecycle management.
  • Build observability and security capabilities for reliability.
  • Collaborate with infra and networking teams to scale services.
  • Drive automation, monitoring, incident response, and continuous improvement.

Skills

Python
Go
Kubernetes
Microservices
REST
Terraform
Ansible
Distributed systems
Prometheus
Grafana
OpenTelemetry
gNMI
Networking
Security
Problem solving

Education

BS or equivalent

Tools

Terraform
Ansible
Temporal
Redis
Kafka
NATS
SQS
DNS
NTP
RADIUS
OAuth
Docker
gRPC
FastAPI
REST
OpenTelemetry
NetBox
Nautobot

Job description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

What You'll Be Doing
  • Build and operate core infrastructure services that power NVIDIA's global AI infrastructure.
  • Architect and develop secure, scalable and highly available cloud-native platform services.
  • Develop software that enables infrastructure orchestration, self-service workflows, and platform automation.
  • Own integrations with internal and external platforms to automate infrastructure provisioning and lifecycle management.
  • Build observability and security capabilities that improve the reliability and resilience of our infrastructure.
  • Partner with infrastructure and networking teams to deliver production services at scale.
  • Drive operational excellence through automation, monitoring, incident response, and continuous improvement.
What We Need To See
  • BS or equivalent experience with 8+ years of relevant industry experience.
  • Strong proficiency in Python and Go, with experience building production-quality software.
  • Experience building cloud-native microservices and APIs on Kubernetes using frameworks such as FastAPI, gRPC, or REST.
  • Experience with infrastructure automation (Terraform, Ansible), workflow orchestration (Temporal), and distributed systems using databases, Redis, and messaging platforms (Kafka, NATS, SQS).
  • Experience designing, building, and operating production infrastructure services such as DNS, NTP, AAA (RADIUS/OAuth), and observability platforms.
  • Strong Linux fundamentals with experience in observability (Prometheus, Grafana, OpenTelemetry, gNMI), networking (BGP, switching, routing, load balancing), and security (VPNs, firewalls, iptables/nftables).
  • Excellent problem-solving, communication, and collaboration skills.
Ways To Stand Out From The Crowd
  • Hands-on experience with network infrastructure including switches, routers, and firewalls.
  • Familiarity with InfiniBand, RDMA, and AI/HPC networking.Experience with NetBox, Nautobot, or similar network source of truth platforms.
  • Contributions to open-source software. Experience with public cloud platforms.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 8, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

JR2022552

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • California (MO)

On-site
USD 170,000 - 322,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • Town of Texas (WI)

On-site
USD 168,000 - 322,000
Equity
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • Colorado

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA Gruppe • United States

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 322,000
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA AI • California (MO)

On-site
USD 168,000 - 322,000
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA • United States

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

Socket.dev • Santa Clara (CA)

On-site
USD 168,000 - 322,000
equity (outlined)
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA Corporation • United States

Remote
USD 168,000 - 322,000
Equity
Benefits
Senior Software Engineer, Networking DGX Cloud
Senior Software Engineer, Networking DGX Cloud

NVIDIA AI • New York (NY)

On-site
USD 200,000 - 391,000
Equity
Benefits