Principal Firmware Engineer - Server Manageability and Observability

NVIDIA

United States

On-site

USD 272,000 - 431,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA is seeking a strong technical architect to own end-to-end system software architecture for its data center products, including firmware, kernel drivers, operating systems, and user mode drivers. You will collaborate with component leads and major cloud providers to drive innovation, architecture decisions, and roadmap alignment across hardware and software interfaces.

The role demands deep expertise in scalable server architectures, accelerators, and cross-functional leadership to deliver

Qualifications

  • Deep expertise in scalable and performant server system architecture, focusing on SW/HW interfaces.
  • Extensive experience with complex system software for accelerators (GPUs, DPUs, FPGAs).
  • Mastery of system firmware (SBIOS, OpenBMC), embedded systems, and Linux kernel internals.
  • Proficiency in Out-of-Band and In-Band management architectures and system management protocols (Redfish, IPMI).

Responsibilities

  • Serve as the primary technical contact for major customers, defining KPIs and gathering requirements.
  • Lead technical innovation and collaborations with hyperscalers to architect next-generation data center products.
  • Align NVIDIA roadmaps with customer requirements through direct engagement.
  • Develop and drive adoption of new technologies and protocols.
  • Make critical technical decisions in ambiguous situations to de-risk programs.

Skills

System architecture
GPU software
Firmware
Linux kernel
Networking
Security tradeoffs
Cross-functional collaboration

Education

BS or MS in Computer Science or Electrical Engineering

Tools

OpenBMC
MCTP/PLDM/SPDM

Job description

About NVIDIA

NVIDIA data center systems, such as DGX and HGX, have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. These platforms bring together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We're looking for a strong technical architect to own the end-to-end architecture of these products, at the system software level, including firmware, kernel drivers, operating systems, and user mode drivers. You will work with component leads internally and engage with industry leading cloud service providers on taking these products to market.

What you’ll be doing
  • Serve as the primary technical point of contact for major customers, leading technological discussions, defining KPIs, gathering requirements, and addressing complex technical queries.
  • As a system software architect, lead technical innovation and strategic collaborations with major hyperscalers to architect next-generation data center products.
  • Align NVIDIA's roadmap with major customers' requirements through direct engagement.
  • Develop and drive adoption of new technologies and protocols.
  • Make critical technical decisions in ambiguous situations, mitigating risks through left-shift strategies.
What we need to see
  • Deep expertise in scalable and performant server system architecture, focusing on SW/HW interfaces.
  • Extensive experience with complex system software for accelerators (GPUs, DPUs, FPGAs).
  • Mastery of system firmware (SBIOS, OpenBMC), embedded systems, and Linux kernel internals.
  • Proficiency in Out-of-Band and In-Band management architectures, device management protocols (e.g., MCTP, PLDM, SPDM, RDE) and system management protocols (Redfish, IPMI).
  • Extensive knowledge of networking technologies and protocols, including TCP/IP, Ethernet, InfiniBand, as well as advanced switching and routing concepts.
  • Experience collaborating with platform security experts to define tradeoffs between security and ease of use.
  • Demonstrated success in leading complex, cross-functional projects to completion, showcasing the ability to influence and achieve results without direct authority in large-scale, collaborative environments. Demonstrable experience in implementing left shift strategy to de-risk program execution.
  • BS or MS degree in Computer Science, Electrical Engineering or related field (or equivalent experience).
  • 15+ years in the area of System architecture and design.
Ways to stand out from the crowd
  • Knowledge of cloud and cluster level deployment and management systems. Participation and contributions in standards bodies such as OCP and DMTF.
  • Familiarity with NVIDIA HPC programming models and libraries (CUDA, cuDNN, DOCA)
  • Knowledge of enterprise storage architectures and distributed parallel processing paradigms
Compensation

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD. You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 21, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Firmware Engineer – Server Manageability and Observability
Principal Firmware Engineer – Server Manageability and Observability

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Equity options
Comprehensive benefits
Principal Firmware Engineer – Server Manageability and Observability
Principal Firmware Engineer – Server Manageability and Observability

Segment (Twilio) • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Equity
Benefits package
Distinguished Engineer – Data Center System Software Architect
Distinguished Engineer – Data Center System Software Architect

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Equity options
Comprehensive benefits package
Distinguished Engineer – Data Center System Software Architect
Distinguished Engineer – Data Center System Software Architect

Segment (Twilio) • Santa Clara (CA)

On-site
USD 320,000 - 488,750
Equity
Benefits
Principal Firmware Engineer - Data Center Server Management
Principal Firmware Engineer - Data Center Server Management

NVIDIA • United States

On-site
USD 320,000 - 489,000
Equity
Benefits
Senior Firmware Engineer - CSP Engagements
Senior Firmware Engineer - CSP Engagements

NVIDIA • United States

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Linux Kernel Systems Software Engineer - CSP Engagements
Senior Linux Kernel Systems Software Engineer - CSP Engagements

NVIDIA • United States

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Software Architect - Data Center Systems
Senior Software Architect - Data Center Systems

NVIDIA • United States

On-site
USD 224,000 - 431,000
System Architect
System Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Principal Software Engineer – CSP Engagements
Principal Software Engineer – CSP Engagements

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package