Senior SRE - Scalable Infra & Reliability (Equity)

NVIDIA

Durham (NC)

On-site

USD 224,000 - 431,250

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Durham, NC seeks Systems and Software Engineers to build and run reliable large-scale infrastructure platform services atop NVIDIA hardware. You will design, deploy, and operate services, aiming for dependable performance and scalable operations across the stack.

The role emphasizes automation, blameless incident response, and collaboration with peer teams. A BS in CS or related field plus 12+ years of experience is required, with strong Python/Go and Linux expertise.

Qualifications

  • BS degree in Computer Science or a related technical field involving coding or equivalent experience.
  • 12+ years of relevant experience.
  • Proven ability to initiate and drive projects and collaborate across teams.
  • Experience with infrastructure automation and distributed systems for large-scale cloud environments.
  • Proficiency in Python, Go, Perl or Ruby.

Responsibilities

  • Design, build, deploy, and run infrastructure services and manage the software life cycle.
  • Define internal service level objectives and error budgets as part of observability strategy.
  • Automate toil where ROI justifies it.

Skills

Python
Go
Perl
Ruby
Linux
Networking
Storage
Containers
SRE

Education

BS degree in Computer Science or related technical field

Tools

Kubernetes
OpenStack
Docker
Slurm

Job description

NVIDIA in Durham, NC seeks Systems and Software Engineers to build and run reliable large-scale infrastructure platform services atop NVIDIA hardware. You will design, deploy, and operate services, aiming for dependable performance and scalable operations across the stack.

The role emphasizes automation, blameless incident response, and collaboration with peer teams. A BS in CS or related field plus 12+ years of experience is required, with strong Python/Go and Linux expertise.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Infra & SRE Engineer — Scale & Reliability (Equity)
Senior Infra & SRE Engineer — Scale & Reliability (Equity)

NVIDIA • Austin (TX)

On-site
USD 230,000 - 420,000
Equity
Benefits
Senior Infrastructure & Reliability Engineer – Equity
Senior Infrastructure & Reliability Engineer – Equity

NVIDIA • Westford (MA)

On-site
USD 224,000 - 431,000
Equity compensation
Benefits
Senior Systems & Reliability Engineer
Senior Systems & Reliability Engineer

NVIDIA • California (MO)

On-site
USD 224,000 - 431,000
Equity
Benefits
Senior Infrastructure Engineer – SRE & Automation
Senior Infrastructure Engineer – SRE & Automation

NVIDIA AI • Durham (CA)

On-site
USD 180,000 - 260,000
Equity
Benefits
Senior Infrastructure Engineer — Scale, Reliability Equity
Senior Infrastructure Engineer — Scale, Reliability Equity

NVIDIA Gruppe • United States

On-site
USD 230,000 - 430,000
Equity
Benefits
Senior Systems Engineer - Scale Infra & Reliability Equity
Senior Systems Engineer - Scale Infra & Reliability Equity

NVIDIA • Columbia (SC)

On-site
USD 224,000 - 431,250
Equity
Benefits package
Senior SRE, BCM/DGX Cloud - Scale GPU Clusters
Senior SRE, BCM/DGX Cloud - Scale GPU Clusters

NVIDIA • Santa Clara (CA)

On-site
USD 168,000 - 334,000
Equity
Benefits
SRE: Hardware Infrastructure for Reliable, AI‑Driven Ops
SRE: Hardware Infrastructure for Reliable, AI‑Driven Ops

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Senior SRE, BCM/DGX Cloud - Scale GPU Clusters (Equity)
Senior SRE, BCM/DGX Cloud - Scale GPU Clusters (Equity)

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 168,000 - 334,000
Equity
Benefits
Senior Staff SRE — Global Infra, Automation & Observability
Senior Staff SRE — Global Infra, Automation & Observability

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 200,000 - 322,000