Senior Distributed Systems Engineer — EDA Infra (Equity)

NVIDIA Corporation

Washington

Hybrid

USD 152,000 - 288,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation is seeking a Sr Software Engineer to design and scale distributed infrastructure for EDA workloads in a production environment. You will build automation that provisions, configures, and operates large GPU/CPU compute fleets, and develop health-management and remediation systems to improve reliability across data centers.

The role requires strong Go or Python skills, experience with Linux-based clusters, and collaboration across multiple teams to reduce toil and improve

Qualifications

  • 5+ years of software or infrastructure engineering experience supporting large-scale production systems.
  • A BS in Computer Science, Engineering, Physics, Mathematics, or a related field, or equivalent experience.
  • Strong programming experience in Go or Python, including data structures and algorithms.
  • Experience designing automation for distributed systems and large fleets of Linux-based compute nodes.
  • Understanding of performance, security, reliability, fault tolerance, state management, and data consistency in complex systems.

Responsibilities

  • Design and build platforms that automate provisioning, configuration, operation and lifecycle management of large-scale GPU/CPU compute infrastructure.
  • Develop monitoring, health-management, and remediation systems to improve reliability and utilization of EDA compute environments.
  • Automate hardware deployment, OS configuration, firmware and software updates, cluster enrollment, and recovery workflows.
  • Build reliable services and workflows that integrate with workload schedulers, infrastructure management systems, and observability platforms.
  • Use diagnostics, signals, scheduler data, and telemetry to identify failures and return unhealthy systems to service.
  • Collaborate with EDA, infrastructure, networking, storage, and hardware teams to deliver scalable solutions for chip-design workloads.

Skills

Distributed systems
Go
Python
Automation
Observability
Communication

Education

BS in Computer Science, Engineering, Physics, Mathematics, or related field

Tools

Slurm
LSF
Kubernetes
Bright Cluster Manager

Job description

NVIDIA Corporation is seeking a Sr Software Engineer to design and scale distributed infrastructure for EDA workloads in a production environment. You will build automation that provisions, configures, and operates large GPU/CPU compute fleets, and develop health-management and remediation systems to improve reliability across data centers.

The role requires strong Go or Python skills, experience with Linux-based clusters, and collaboration across multiple teams to reduce toil and improve

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Distributed Systems Engineer — EDA Infra (Equity)
Senior Distributed Systems Engineer — EDA Infra (Equity)

NVIDIA • Austin (TX)

On-site
USD 152,000 - 288,000
Senior Distributed Systems Engineer – EDA Infra (Equity)
Senior Distributed Systems Engineer – EDA Infra (Equity)

NVIDIA • California (MO)

On-site
USD 170,000 - 260,000
Equity
Benefits
Senior Distributed Systems Engineer, EDA Infrastructure
Senior Distributed Systems Engineer, EDA Infrastructure

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Distributed Systems Engineer, EDA Infra - Equity
Senior Distributed Systems Engineer, EDA Infra - Equity

NVIDIA • Westford (MA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Distributed Systems Engineer, EDA Infra & Automation
Distributed Systems Engineer, EDA Infra & Automation

NVIDIA Corporation • Northern (KY)

Hybrid
USD 152,000 - 288,000
Senior Software Engineer, Distributed Systems for EDA Infra
Senior Software Engineer, Distributed Systems for EDA Infra

NVIDIA • Durham (NC)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Linux Systems Engineer – EDA Infra (Equity)
Senior Linux Systems Engineer – EDA Infra (Equity)

NVIDIA Corporation • Northern (KY)

Hybrid
USD 184,000 - 357,000
Senior Systems Engineer: GPU Infra & Distributed Automation
Senior Systems Engineer: GPU Infra & Distributed Automation

NVIDIA AI • Durham (CA)

On-site
USD 150,000 - 210,000
Equity
Benefits
Senior Systems & Infra Automation Engineer (Equity Eligible)
Senior Systems & Infra Automation Engineer (Equity Eligible)

NVIDIA • Austin (TX)

On-site
USD 184,000 - 357,000
Equity
Senior Software Engineer - EDA Infra & System Validation
Senior Software Engineer - EDA Infra & System Validation

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity
Benefits