GPU Kernel Evaluation Expert

OpenTrain AI

Northern (KY)

Hybrid

USD 96,000 - 124,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OpenTrain AI is seeking a GPU Kernel Evaluation Expert to assess the quality, correctness, and completeness of GPU and accelerator kernel development tasks used to train and evaluate frontier AI models.

You will evaluate numerical correctness, benchmarking fairness, task scope, compilation validity, and runtime behavior across a range of kernel task types. This is a fully remote contract role for eligible candidates in the United States, with a 40-hour work week and a 20+ hour minimum.

Qualifications

  • 3+ years developing, optimizing, or verifying GPU kernels and accelerators.
  • Experience with at least two of CUDA, Triton, NKI, or Pallas for JAX.
  • Strong understanding of numerical-correctness criteria for kernels.
  • Experience with Nsight, NCU, roofline analysis, or profiling tools.
  • Experience with three task types: generation, translation/lowering, migration, debugging, optimization, or operator fusion.

Responsibilities

  • Review GPU/kernel task submissions for quality, correctness, and completeness.
  • Evaluate numerical accuracy using absolute, relative, and ULP tolerances.
  • Assess selection and suitability of reference implementations.
  • Analyze performance-benchmarking fairness and profiling results.
  • Ensure compilation and runtime validity across environments.
  • Provide rubric-based written feedback for every evaluated task.

Skills

GPU kernels
CUDA
Triton
NKI
Pallas for JAX

Tools

Nsight
NCU
Roofline analysis
Profiling tools

Job description

About OpenTrain

OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping contributors discover projects, build a professional profile, and apply in minutes.

Creating an OpenTrain account is free, and this opportunity offers a way to apply specialized GPU programming expertise to the development and evaluation of advanced AI systems.

About AI Training Work

AI training is the human side of building artificial intelligence. Specialists review code, assess model outputs, and provide structured feedback that helps AI systems become more capable, reliable, and useful.

In this role, your technical evaluations will support training and evaluation workflows involving GPU and accelerator kernel development. The work is fully remote and combines deep engineering expertise with cutting‑edge AI development.

The GPU Kernel Evaluation Expert Role

OpenTrain AI is seeking a GPU Kernel Evaluation Expert to assess the quality, correctness, and completeness of GPU and accelerator kernel development tasks used to train and evaluate frontier AI models.

You will evaluate numerical correctness, benchmarking fairness, task scope, compilation validity, and runtime behavior across a range of kernel task types. Each submission requires clear, rubric‑based written feedback.

  • Fully remote contract role for eligible candidates in the United States
  • Pay range: $70.00 to $90.00 per hour
  • The role description specifies a 40-hour-per-week commitment
  • The structured time requirement is listed as 20+ hours per week
  • Employment types: contractor and part‑time
  • Primary language: English
What You'll Evaluate

You will review technical submissions and task designs across multiple aspects of GPU and accelerator kernel development. Your assessments should be accurate, consistent, and grounded in the applicable evaluation rubric.

  • GPU and accelerator kernel tasks for quality, correctness, and completeness
  • Numerical correctness using absolute, relative, and ULP tolerances
  • Selection and suitability of reference implementations
  • Performance‑benchmarking fairness and profiling results
  • Compilation and runtime validity across different environments
  • Generation from specification, translation or lowering, migration, debugging, optimization, and operator‑fusion tasks
  • Written, rubric‑based feedback for every evaluated task
Required Qualifications

The listing is marked entry level, but the role specifically requires at least three years of hands‑on experience developing, optimizing, or verifying GPU or accelerator kernels. You must have experience in at least two of CUDA, Triton, NKI, or Pallas for JAX.

  • 3+ years developing, optimizing, or verifying GPU or accelerator kernels
  • Hands‑on experience with at least two of CUDA, Triton, NKI, or Pallas for JAX
  • Strong understanding of numerical‑correctness criteria for kernels
  • Experience with Nsight, NCU, roofline analysis, or framework‑native profiling tools
  • Familiarity with common compilation and runtime failure modes
  • Experience with at least three task types: generation, translation or lowering, migration, debugging, optimization, or operator fusion
Helpful Technical Background

The following experience is helpful for evaluating a broad range of kernel tasks and accelerator environments. These qualifications are preferred background rather than listed minimum requirements.

  • Experience across NVIDIA GPU ecosystems such as CUDA and Triton
  • Experience with custom‑accelerator ecosystems such as NKI, Pallas, or TPU
  • Compiler engineering, MLIR, or intermediate‑representation lowering
  • Memory‑hierarchy optimization, including shared‑memory tiling, register pressure, bank conflicts, and coalescing patterns
  • Contributions to cuBLAS, cuDNN, Triton community kernels, or JAX/XLA custom calls
How This Work Supports AI Development

Modern AI systems depend on people who can inspect technical outputs, identify failure modes, and distinguish reliable results from misleading ones. By evaluating kernel implementations and their benchmarks, you help improve the quality of the data and feedback used in AI development.

This is a specialized path within the broader AI‑training industry, where technical contributors can use software and systems expertise to shape how advanced models are built and evaluated.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote GPU Kernel Evaluation Specialist - Contract
Remote GPU Kernel Evaluation Specialist - Contract

OpenTrain AI • Northern (KY)

Hybrid
USD 96,000 - 124,000
Trainium NKI Kernel Expert
Trainium NKI Kernel Expert

OpenTrain AI • Northern (KY)

Hybrid
USD 96,000 - 124,000
Contractor role
Part-time opportunity
US-based role
+3
GPU Kernel Expert - AI Specialist
GPU Kernel Expert - AI Specialist

Mercor • San Francisco (CA)

On-site
USD 150,000 - 210,000
GPU Kernel Developer - AI Trainer
GPU Kernel Developer - AI Trainer

Mercor • Chicago (IL)

On-site
USD 120,000 - 180,000
GPU Kernel Expert - AI Specialist
GPU Kernel Expert - AI Specialist

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 180,000
GPU Kernel Developer - AI Trainer
GPU Kernel Developer - AI Trainer

Obsidian • Chicago (IL)

On-site
USD 120,000 - 180,000
GPU Kernel Expert Mercor · Remote — United States $70-90/hr →
GPU Kernel Expert Mercor · Remote — United States $70-90/hr →

Dorado • Northern (KY)

Hybrid
USD 100,000 - 180,000
GPU Kernel Architect for AI Training & Evaluation
GPU Kernel Architect for AI Training & Evaluation

Mercor • Chicago (IL)

On-site
USD 120,000 - 180,000
AI Kernel Quality Engineer — GPU/Accelerator Expert
AI Kernel Quality Engineer — GPU/Accelerator Expert

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 180,000
GPU Kernel Quality Engineer
GPU Kernel Quality Engineer

Obsidian • New York (NY)

Remote
USD 140,000 - 200,000