Performance/ Benchmark Engineer - NVIDIA GPU Systems

Yoh,-A-Day-

Santa Clara (CA)

On-site

USD 250,000 - 300,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, Prescription, Dental & Vision
401K Retirement Savings Plan
Direct Deposit & weekly epayroll
Certification and training

Job summary

Yoh is seeking a hands-on Performance and Benchmark Engineer to characterize AI workloads on NVIDIA GPU systems in Santa Clara. You will scope, run, and optimize benchmarks across DGX and related platforms, focusing on latency, throughput, and scaling.

Collaborate with architecture, compute, and networking teams to drive end-to-end performance improvements while delivering repeatable results and clear reporting.

Qualifications

  • Deep hands-on experience with NVIDIA GPU compute platforms and AI/ML performance benchmarking.
  • Strong understanding of AI inference, model performance, workload characterization, and GPU architecture.
  • Experience with NVIDIA DGX, B200/B300, H100/H200, Blackwell, Hopper, or comparable GPU systems.
  • Experience analyzing performance metrics including latency, throughput, GPU utilization, memory bandwidth, and multi-GPU scaling.
  • Strong scripting and automation skills using Python or similar languages.

Responsibilities

  • Develop and execute performance benchmarks for AI inference and machine learning workloads across NVIDIA GPU systems.
  • Characterize performance on platforms including NVIDIA DGX and B200/B300-based systems, analyzing throughput, latency, utilization, memory behavior, and scaling efficiency.
  • Evaluate AI models and workload configurations to identify performance bottlenecks and recommend system or architecture improvements.
  • Build benchmarking methodologies, automation, and reporting frameworks to produce repeatable performance results.
  • Collaborate with architecture, compute, networking, and software teams to optimize end-to-end AI cluster performance.

Skills

NVIDIA GPU compute platforms
AI/ML performance benchmarking
Python scripting
Performance metrics

Tools

NVIDIA DGX
B200/B300 systems
H100/H200
Nsight

Job description

Performance/ Benchmark Engineer - NVIDIA GPU Systems (BH-399465)

Location Santa Clara, United States Sector Engineering Salary $250,000.00 to $300,000.00 per hour Benefits $250k-$300k Base + Equity Options

Performance/ Benchmark Engineer - NVIDIA GPU Systems
NVIDIA GPU Systems / AI Inference / Performance Engineering
Overview

Seeking a hands-on Performance and Benchmarking Engineer to characterize and optimize AI workloads running on large-scale NVIDIA GPU infrastructure. This role sits within an architecture team and focuses primarily on GPU compute performance, AI inference, and system-level benchmarking, with networking performance as a secondary consideration.

Key Responsibilities
  • Develop and execute performance benchmarks for AI inference and machine learning workloads across NVIDIA GPU systems.
  • Characterize performance on platforms including NVIDIA DGX and B200/B300-based systems, analyzing throughput, latency, utilization, memory behavior, and scaling efficiency.
  • Evaluate AI models and workload configurations to identify performance bottlenecks and recommend system or architecture improvements.
  • Build benchmarking methodologies, automation, and reporting frameworks to produce repeatable performance results.
  • Collaborate with architecture, compute, networking, and software teams to optimize end-to-end AI cluster performance.
Required Qualifications
  • Deep hands-on experience with NVIDIA GPU compute platforms and AI/ML performance benchmarking.
  • Strong understanding of AI inference, model performance, workload characterization, and GPU architecture.
  • Experience with NVIDIA DGX, B200/B300, H100/H200, Blackwell, Hopper, or comparable GPU systems.
  • Experience analyzing performance metrics including latency, throughput, GPU utilization, memory bandwidth, and multi-GPU scaling.
  • Strong scripting and automation skills using Python or similar languages.
Preferred Qualifications
  • Experience with MLPerf, CUDA, NCCL, TensorRT, Triton Inference Server, PyTorch, Nsight, or similar AI performance and profiling technologies.
  • Experience benchmarking LLMs, inference workloads, distributed training, or large-scale GPU clusters.
  • Familiarity with RDMA, RoCE, InfiniBand, Ethernet, GPUDirect RDMA, or networking considerations affecting GPU cluster performance.

Estimated Min Rate: $250,000.00/Annually

Estimated Max Rate: $300,000.00/Annually

What's In It for You?

We welcome you to be a part of the largest and legendary global staffing companies to meet your career aspirations. Yoh's network of client companies has been employing professionals like you for over 65 years in the U.S., UK and Canada. Join Yoh's extensive talent community that will provide you with access to Yoh's vast network of opportunities and gain access to this exclusive opportunity available to you. Benefit eligibility is in accordance with applicable laws and client requirements. Benefits include:

  • Medical, Prescription, Dental & Vision Benefits (for employees working 20+ hours per week)
  • Health Savings Account (HSA) (for employees working 20+ hours per week)
  • Life & Disability Insurance (for employees working 20+ hours per week)
  • MetLife Voluntary Benefits
  • Employee Assistance Program (EAP)
  • 401K Retirement Savings Plan
  • Direct Deposit & weekly epayroll
  • Referral Bonus Programs
  • Certification and training opportunities

Note: Any pay ranges displayed are estimations. Actual pay is determined by an applicant's experience, technical expertise, and other qualifications as listed in the job description.

Yoh, a Day & Zimmermann company, is an Equal Opportunity Employer.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Visit https://www.yoh.com/applicants-with-disabilities to contact us if you are an individual with a disability and require accommodation in the application process.

For California applicants, qualified applicants with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. All of the material job duties described in this posting are job duties for which a criminal history may have a direct, adverse, and negative relationship potentially resulting in the withdrawal of a conditional offer of employment.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Stay Safe During Your Job Search

Fraudulent recruiting communications have become increasingly common. Emails from Yoh recruiters will only come from an @yoh.com email address. Yoh will never ask candidates to pay fees, purchase equipment, send gift cards, or transfer funds as part of the recruiting or hiring process. If you receive a communication that appears suspicious or requests payment, contact Yoh directly before responding or sharing personal information.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance/ Benchmark Engineer - NVIDIA GPU Systems
Performance/ Benchmark Engineer - NVIDIA GPU Systems

Yoh Services LLC • California (MO)

On-site
USD 250,000 - 300,000
Medical benefits
Dental & Vision
401K Retirement
Performance/ Benchmark Engineer - NVIDIA GPU Systems
Performance/ Benchmark Engineer - NVIDIA GPU Systems

YOH Services LLC • Santa Clara (CA)

On-site
USD 250,000 - 300,000
Medical benefits
Health Savings Account (HSA)
401K Retirement Savings Plan
+2
Performance/ Benchmark Engineer - NVIDIA GPU Systems
Performance/ Benchmark Engineer - NVIDIA GPU Systems

Yoh, A Day & Zimmermann Company • California (MO)

On-site
USD 250,000 - 300,000
Medical+Vision
HSA
Life & Disability
+5
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 124,000 - 196,000
Equity
Benefits package
Hybrid work model
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

NVIDIA • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Equity
Benefits package
Inference Performance Engineer, Agent Driven Inference Optimization
Inference Performance Engineer, Agent Driven Inference Optimization

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity eligibility
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Senior Data Center Performance Engineer - Benchmarking and Optimization
Senior Data Center Performance Engineer - Benchmarking and Optimization

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Benefits
Inference Performance Engineer, Agent Driven Inference Optimization
Inference Performance Engineer, Agent Driven Inference Optimization

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Senior Performance Engineer - DGX Cloud
Senior Performance Engineer - DGX Cloud

NVIDIA AI • Eugene (OR)

On-site
USD 224,000 - 432,000
Equity
Benefits