37303 HD - CUDA Engineering Expert

Cephas Consultancy Services Private Limited

California (MO)

On-site

USD 120,000 - 180,000

Part time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cephas Consultancy Services Private Limited seeks a CUDA Engineering Expert for a remote contractor role on a cutting-edge AI project. You will analyze, profile, and optimize GPU kernels with CUDA to maximize throughput on modern hardware and contribute to training next-generation AI systems.

With 5–30 years in CUDA, advanced C++ HPC development, GLSL and WebGPU, you will refactor code for maintainability, implement shader logic, and document performance improvements.

Qualifications

  • CUDA programming expertise with kernel performance tuning (≤140 chars).
  • Advanced C++ development in HPC environments (≤140 chars).
  • Experience with GLSL and WebGPU for graphics/compute shaders (≤140 chars).
  • Proficiency using GPU profilers (Nsight, Visual Profiler, etc.) for guided optimization (≤140 chars).
  • Strong analytical abilities to evaluate kernel performance across hardware generations (≤140 chars).
  • Excellent written and verbal communication skills; remote collaboration experience (≤140 chars).

Responsibilities

  • Analyze, profile, and optimize GPU kernels using CUDA.
  • Identify bottlenecks and propose optimization strategies.
  • Refactor C++/CUDA for cross-GPU architecture and maintainability.
  • Implement shader logic using GLSL and WebGPU.
  • Document findings and performance improvements with actionable reports.
  • Collaborate with AI lab and remote, cross-disciplinary teams.

Skills

CUDA
C++
GLSL
WebGPU

Job description

37303 HD - CUDA Engineering Expert

Cephas Consultancy Services Private Limited Permanent Remote Work, California, United States

About this position

Positions:50 Contract
Experience
5 - 30 Years
Role Type:Contractor
Location:Remote

We are engaging CUDA Engineering Experts to contribute to a cutting-edge customer project focused on GPU kernel optimization in collaboration with a leading AI lab. In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input. No prior experience in AI is required - your domain knowledge is what matters.

Scope of Work

Analyze, profile, and optimize GPU kernels using CUDA and relevant profiling tools to maximize computational throughput on modern hardware. Collaborate with project stakeholders to assess and identify kernel bottlenecks, proposing targeted optimization strategies. Refactor C++ and CUDA codebases for improved maintainability, efficiency, and adaptability across diverse GPU architectures. Implement shader logic and graphics workflows using GLSL and WebGPU, ensuring seamless integration with existing pipelines. Document key findings, optimization steps, and performance improvements with clear, actionable reports and technical communication. Contribute expertise to design discussions, supporting the evaluation of new GPU-based approaches and performance metrics. Stay informed on advancements in GPU programming and share relevant insights to enhance project outcomes.

Preferred Qualifications

Demonstrated expertise in CUDA programming, with a strong track record of performance-tuning GPU kernels. Advanced C++ development skills, particularly in high-performance computing environments. Hands-on experience with GLSL and WebGPU for graphics and compute shader development. Proficiency using GPU profilers (such as Nsight, Visual Profiler, or similar tools) for guided optimization. Strong analytical abilities to evaluate and reason about kernel performance across hardware generations. Excellent written and verbal communication skills—clear documentation and technical reporting are essential. Experience collaborating in remote, cross-disciplinary project settings is a plus.

Skills

Skills: CUDA, C++, GLSL, WebGPU

I agree to receive messaging notifications regarding my applications at the phone number provided. MSG & Data rates may apply. MSG frequency varies. Reply help for help and stop to end. For more information view our terms of service and privacy policy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote CUDA Kernel Optimization Engineer
Remote CUDA Kernel Optimization Engineer

Cephas Consultancy Services Private Limited • California (MO)

Hybrid
USD 120,000 - 180,000
Remote CUDA Engineer: GPU Kernel Optimizer & HPC Expert
Remote CUDA Engineer: GPU Kernel Optimizer & HPC Expert

YO HR Consultancy • United States

Remote
USD 120,000 - 180,000
Remote CUDA Kernel Engineer (Contract)
Remote CUDA Kernel Engineer (Contract)

Congruex • Memphis (TN)

On-site
USD 83,000 - 138,000
CUDA Engineer - Kernel Optimization
CUDA Engineer - Kernel Optimization

Mercor • San Francisco (CA)

On-site
USD 96,000 - 165,000
CUDA Engineer - Kernel Optimization - AI Trainer
CUDA Engineer - Kernel Optimization - AI Trainer

Mercor • Chicago (IL)

On-site
USD 83,000 - 165,000
Remote CUDA Kernel Performance Engineer
Remote CUDA Kernel Performance Engineer

Mercor • United States

On-site
USD 110,000 - 165,000
Remote CUDA Kernel Engineer – Optimize GPU Performance
Remote CUDA Kernel Engineer – Optimize GPU Performance

Pragmatike • Cambridge (MA)

On-site
USD 150,000 - 230,000
Salary + equity
Sign-on bonus
Health/Dental/Vision
+1
CUDA GPU Kernel Performance Engineer (Contract)
CUDA GPU Kernel Performance Engineer (Contract)

Mercor • United States

Remote
Software Engineer, CUDA Deep Learning Systems
Software Engineer, CUDA Deep Learning Systems

NVIDIA AI • Santa Clara (TX)

On-site
USD 124,000 - 196,000
Equity
Senior CUDA HPC Engineer — Remote
Senior CUDA HPC Engineer — Remote

Bright Vision Technologies • Raleigh (NC), Concord (NC)

On-site
USD 130,000 - 180,000