Performance/ Benchmark Engineer - NVIDIA GPU Systems

Yoh Services LLC

California (MO)

On-site

USD 250,000 - 300,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical benefits
Dental & Vision
401K Retirement

Job summary

Yoh Services LLC is seeking a hands-on Performance/ Benchmark Engineer to characterize and optimize AI workloads on large-scale NVIDIA GPU infrastructure, including DGX and B200/B300 systems. You will focus on GPU compute performance, AI inference, and system benchmarking while collaborating across architecture, compute, and networking teams.

The role requires deep hands-on GPU benchmarking experience, strong Python scripting, and proficiency with CUDA/NCCL.

Qualifications

  • Deep hands-on experience with NVIDIA GPU compute platforms and AI/ML performance benchmarking.
  • Strong understanding of AI inference, model performance, workload characterization, and GPU architecture.
  • Experience with NVIDIA DGX, B200/B300, H100/H200, Blackwell, Hopper, or comparable GPU systems.
  • Experience analyzing performance metrics including latency, throughput, GPU utilization, memory bandwidth, and multi-GPU scaling.
  • Strong scripting and automation skills using Python or similar languages.

Responsibilities

  • Develop and execute performance benchmarks for AI inference and machine learning workloads across NVIDIA GPU systems.
  • Characterize performance on platforms including NVIDIA DGX and B200/B300-based systems, analyzing throughput, latency, utilization, memory behavior, and scaling efficiency.
  • Evaluate AI models and workload configurations to identify performance bottlenecks and recommend system or architecture improvements.
  • Build benchmarking methodologies, automation, and reporting frameworks to produce repeatable performance results.
  • Collaborate with architecture, compute, networking, and software teams to optimize end-to-end AI cluster performance.

Skills

NVIDIA GPU
AI benchmarking
Python scripting
Performance analysis

Education

BS in CS/EE

Tools

CUDA
NCCL
TensorRT
PyTorch
Nsight

Job description

Performance/ Benchmark Engineer - NVIDIA GPU Systems
NVIDIA GPU Systems / AI Inference / Performance Engineering
Overview

Seeking a hands-on Performance and Benchmarking Engineer to characterize and optimize AI workloads running on large-scale NVIDIA GPU infrastructure. This role sits within an architecture team and focuses primarily on GPU compute performance, AI inference, and system-level benchmarking, with networking performance as a secondary consideration.

Key Responsibilities
  • Develop and execute performance benchmarks for AI inference and machine learning workloads across NVIDIA GPU systems.
  • Characterize performance on platforms including NVIDIA DGX and B200/B300-based systems, analyzing throughput, latency, utilization, memory behavior, and scaling efficiency.
  • Evaluate AI models and workload configurations to identify performance bottlenecks and recommend system or architecture improvements.
  • Build benchmarking methodologies, automation, and reporting frameworks to produce repeatable performance results.
  • Collaborate with architecture, compute, networking, and software teams to optimize end-to-end AI cluster performance.
Required Qualifications
  • Deep hands-on experience with NVIDIA GPU compute platforms and AI/ML performance benchmarking.
  • Strong understanding of AI inference, model performance, workload characterization, and GPU architecture.
  • Experience with NVIDIA DGX, B200/B300, H100/H200, Blackwell, Hopper, or comparable GPU systems.
  • Experience analyzing performance metrics including latency, throughput, GPU utilization, memory bandwidth, and multi-GPU scaling.
  • Strong scripting and automation skills using Python or similar languages.
Preferred Qualifications
  • Experience with MLPerf, CUDA, NCCL, TensorRT, Triton Inference Server, PyTorch, Nsight, or similar AI performance and profiling technologies.
  • Experience benchmarking LLMs, inference workloads, distributed training, or large-scale GPU clusters.
  • Familiarity with RDMA, RoCE, InfiniBand, Ethernet, GPUDirect RDMA, or networking considerations affecting GPU cluster performance.

Estimated Min Rate: $250,000.00/Annually

Estimated Max Rate: $300,000.00/Annually

What’s In It for You?

We welcome you to be a part of the largest and legendary global staffing companies to meet your career aspirations. Yoh’s network of client companies has been employing professionals like you for over 65 years in the U.S., UK and Canada. Join Yoh’s extensive talent community that will provide you with access to Yoh’s vast network of opportunities and gain access to this exclusive opportunity available to you. Benefit eligibility is in accordance with applicable laws and client requirements.

  • Medical, Prescription, Dental & Vision Benefits (for employees working 20+ hours per week)
  • Health Savings Account (HSA) (for employees working 20+ hours per week)
  • Life & Disability Insurance (for employees working 20+ hours per week)
  • MetLife Voluntary Benefits
  • Employee Assistance Program (EAP)
  • 401K Retirement Savings Plan
  • Direct Deposit & weekly epayroll
  • Referral Bonus Programs
  • Certification and training opportunities

Note: Any pay ranges displayed are estimations. Actual pay is determined by an applicant's experience, technical expertise, and other qualifications as listed in the job description.

Yoh, a Day & Zimmermann company, is an Equal Opportunity Employer.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Visit https://www.yoh.com/applicants-with-disabilities to contact us if you are an individual with a disability and require accommodation in the application process.

For California applicants, qualified applicants with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. All of the material job duties described in this posting are job duties for which a criminal history may have a direct, adverse, and negative relationship potentially resulting in the withdrawal of a conditional offer of employment.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Stay Safe During Your Job Search

Fraudulent recruiting communications have become increasingly common. Emails from Yoh recruiters will only come from an @yoh.com email address. Yoh will never ask candidates to pay fees, purchase equipment, send gift cards, or transfer funds as part of the recruiting or hiring process. If you receive a communication that appears suspicious or requests payment, contact Yoh directly before responding or sharing personal information.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance/ Benchmark Engineer - NVIDIA GPU Systems
Performance/ Benchmark Engineer - NVIDIA GPU Systems

Yoh,-A-Day- • Santa Clara (CA)

On-site
USD 250,000 - 300,000
Medical, Prescription, Dental & Vision
401K Retirement Savings Plan
Direct Deposit & weekly epayroll
+1
Performance/ Benchmark Engineer - NVIDIA GPU Systems
Performance/ Benchmark Engineer - NVIDIA GPU Systems

YOH Services LLC • Santa Clara (CA)

On-site
USD 250,000 - 300,000
Medical benefits
Health Savings Account (HSA)
401K Retirement Savings Plan
+2
Performance/ Benchmark Engineer - NVIDIA GPU Systems
Performance/ Benchmark Engineer - NVIDIA GPU Systems

Yoh, A Day & Zimmermann Company • California (MO)

On-site
USD 250,000 - 300,000
Medical+Vision
HSA
Life & Disability
+5
Principal Engineer Technical Support
Principal Engineer Technical Support

Yoh Services LLC • Santa Clara (CA)

On-site
USD 248,000 - 269,000
Medical, Prescription, Dental & Vision
Health Savings Account (HSA)
Life & Disability Insurance
+4
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

NVIDIA • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Equity
Benefits package
Distinguished Engineer - Cloud Services
Distinguished Engineer - Cloud Services

Yoh Services LLC • Santa Clara (CA)

On-site
USD 313,000 - 353,000
Medical Benefits
Dental & Vision Benefits
Health Savings Account
+3
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 124,000 - 196,000
Equity
Benefits package
Hybrid work model
Inference Performance Engineer, Agent Driven Inference Optimization
Inference Performance Engineer, Agent Driven Inference Optimization

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity eligibility
Senior Data Center Performance Engineer - Benchmarking and Optimization
Senior Data Center Performance Engineer - Benchmarking and Optimization

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Benefits
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000