Senior Distributed Systems Engineer - EDA Infra

NVIDIA

United States

On-site

USD 152,000 - 242,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is seeking Sr Software Engineer - Distributed Systems to design and scale infrastructure for EDA workloads. You will build automation, health-management, and lifecycle tooling across GPU/CPU compute fleets and coordinate with EDA, networking, storage, and hardware teams.

The role emphasizes automation, reliability, and capacity planning in a multi-team, geographically distributed environment with strong programming in Go or Python and Linux expertise.

Qualifications

  • 5+ years of software or infrastructure engineering experience
  • BS in Computer Science, Engineering, Physics, Mathematics, or related field or equivalent experience
  • Strong programming in Go or Python, with data structures and algorithms knowledge
  • Experience designing automation for distributed systems and large fleets of Linux nodes
  • Understanding of performance, security, reliability, fault tolerance, state management, and data consistency
  • Experience with infrastructure automation, software deployment, observability, and operational recovery
  • Strong communication and collaboration across teams and regions

Responsibilities

  • Design and build platforms that automate provisioning, configuration, operation, and lifecycle management of large-scale GPU and CPU compute infrastructure
  • Develop monitoring, health-management, and remediation systems to improve reliability and utilization
  • Automate hardware deployment, OS configuration, firmware/software updates, enrollment, and recovery workflows
  • Build reliable services and workflows that integrate with schedulers, infrastructure management, and observability platforms
  • Use telemetry to identify failures and recover systems to service
  • Collaborate with EDA, infrastructure, networking, storage, and hardware teams to deliver scalable solutions
  • Participate in incident response, root-cause analysis, capacity planning, and production-service improvements

Skills

Go
Python
Distributed systems
Linux
Data structures
Algorithms
Automation
Observability

Education

BS in Computer Science, Engineering, Physics, Mathematics, or related field

Tools

Slurm
LSF
Kubernetes
Bright Cluster Manager

Job description

NVIDIA is seeking Sr Software Engineer - Distributed Systems to design and scale infrastructure for EDA workloads. You will build automation, health-management, and lifecycle tooling across GPU/CPU compute fleets and coordinate with EDA, networking, storage, and hardware teams.

The role emphasizes automation, reliability, and capacity planning in a multi-team, geographically distributed environment with strong programming in Go or Python and Linux expertise.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Distributed Systems Engineer, EDA Infrastructure
Senior Distributed Systems Engineer, EDA Infrastructure

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Distributed Systems Engineer EDA Infra Equity
Senior Distributed Systems Engineer EDA Infra Equity

Socket.dev • North Carolina

Hybrid
USD 152,000 - 288,000
Distributed Systems Engineer, EDA Infra & Automation
Distributed Systems Engineer, EDA Infra & Automation

NVIDIA Corporation • Northern (KY)

Hybrid
USD 152,000 - 288,000
Senior Distributed Systems Engineer – EDA Infrastructure
Senior Distributed Systems Engineer – EDA Infrastructure

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Systems Engineer: GPU Infra & Distributed Automation
Senior Systems Engineer: GPU Infra & Distributed Automation

NVIDIA AI • Durham (CA)

On-site
USD 150,000 - 210,000
Equity
Benefits
Senior Systems Software Engineer- EDA Infrastructure
Senior Systems Software Engineer- EDA Infrastructure

NVIDIA AI • Durham (CA)

On-site
USD 150,000 - 210,000
Equity
Benefits
Senior Infra Engineer: Scale GPU/CPU Compute & Automation
Senior Infra Engineer: Scale GPU/CPU Compute & Automation

NVIDIA Gruppe • Washington

On-site
USD 180,000 - 288,000
Senior Systems Engineer - EDA Infra with Equity Options
Senior Systems Engineer - EDA Infra with Equity Options

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Software Engineer - Distributed Systems Engineer, EDA Infrastructure
Senior Software Engineer - Distributed Systems Engineer, EDA Infrastructure

NVIDIA • United States

On-site
USD 152,000 - 242,000
Senior Software Engineer - Distributed Systems Engineer, EDA Infrastructure
Senior Software Engineer - Distributed Systems Engineer, EDA Infrastructure

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits