Developer Technology Engineer, Energy

NVIDIA

United Arab Emirates

On-site

AED 420,000 - 660,000

Full time

13 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

NVIDIA is seeking a Compute Developer Technology (DevTech) expert to accelerate energy simulations and AI workflows. You will optimize CUDA performance on multi-GPU and multi-node systems and work with customer teams to deliver measurable speedups.

The role requires strong C/C++ and Python, CUDA programming, and experience with Nsight tools for profiling. You will engage with internal teams to influence roadmaps and build reusable libraries for future NVIDIA products.

Qualifications

  • BS/MS or equivalent in CS/CE/EE/Physics/Applied Math or related field.
  • Strong programming skills in C/C++ and Python on Linux.
  • Hands-on CUDA programming and GPU performance optimization experience.
  • Experience with profiling using NVIDIA Nsight Systems / Nsight Compute.

Responsibilities

  • Profile, analyze, and optimize GPU-accelerated applications focusing on CUDA kernels and memory movement.
  • Drive performance improvements across the stack, including MPI/NCCL multi-GPU/multi-node scaling.
  • Build reproducible benchmarks, performance reports, and tuning recommendations.
  • Develop reference implementations and patches to enable performance and portability.
  • Support customer engagements from POC to production and advise on deployment best practices.

Skills

C/C++
Python
CUDA programming
Profiling
MPI/NCCL
Linux
English
HPC optimization
Nsight tools

Education

BS/MS in CS/CE/EE/Physics/Applied Math

Tools

Nsight Systems
Nsight Compute

Job description

Our work at NVIDIA is dedicated towards a computing model focused on visual and AI computing. For two decades, NVIDIA has pioneered visual computing, the art and science of computer graphics, with our invention of the GPU. The GPU has also shown to be spectacularly effective at solving some of the most complex problems in computer science. Today, NVIDIA's GPU simulates human intelligence, running deep learning algorithms and acting as the brain of computers, robots and self-driving cars that can perceive and understand the world. We are looking to grow our company and teams with the smartest people in the world and there has never been a more exciting time to join our team!

NVIDIA is looking for a passionate, world-class computer scientists and engineers (Compute Developer Technology - DevTech) to accelerate Energy simulation and AI workflows on NVIDIA platforms. You will focus on CUDA performance optimization for workloads such as seismic processing (e.g., imaging/inversion pipelines), reservoir simulation, power grid simulators, and related HPC/AI production workflows. You will work hands-on with customer and partner engineering teams as well as NVIDIA product and engineering groups to deliver measurable speedups and scalable performance on multi-GPU and multi-node systems.

What You Will Be Doing
  • Profile, analyze, and optimize GPU-accelerated applications with emphasis on CUDA kernels, memory movement, concurrency, and end-to-end throughput.
  • Drive performance improvements across the stack:
    • CUDA C++ kernel optimization, launch configuration, memory hierarchy, streams/events
    • GPU libraries (as applicable): cuBLAS, cuFFT, cuSPARSE, cuSOLVER, NCCL
    • Multi-GPU and multi-node scaling using MPI + NCCL, CPU/GPU overlap, communication patterns
  • Build reproducible benchmarks, performance reports, and tuning recommendations (before/after, methodology, scaling curves).
  • Develop and maintain reference implementations, examples, and/or patches to customer code to enable performance and portability.
  • Support customer engagements (POCs to production), including debugging correctness/performance issues and advising on best practices for deployment (containers, schedulers, clusters).
  • Collaborate with internal teams to file actionable issues, validate fixes, and influence roadmap based on real customer requirements in Energy.
  • Build internal libraries and resusable code that would lead to future NVIDIA products.
What We Need To See
  • BS/MS (or equivalent experience) in CS/CE/EE/Physics/Applied Math or related field.
  • Strong programming skills in C/C++ and Python on Linux.
  • Hands-on experience with CUDA programming and GPU performance optimization concepts.
  • Experience profiling and debugging performance using tools such as NVIDIA Nsight Systems / Nsight Compute (or equivalent).
  • Understanding of parallel computing and performance fundamentals (vectorization, threading, NUMA, memory bandwidth/latency).
  • Ability to communicate technical findings clearly to both engineers and non-engineers.
  • 5+ years relevant experience in GPU/HPC optimization; strong track record of delivered speedups and scaling improvements.
Ways To Stand Out From The Crowd
  • Leads performance reviews with customer stakeholders; creates reusable playbooks/reference designs. Experience/Skills (typical)
  • HPC experience with MPI, distributed systems, and multi-node performance tuning.
  • Energy/HPC domain exposure:
    • Seismic processing pipelines, RTM/FWI-style patterns, FFT/stencil/linear algebra heavy codes
    • Reservoir simulation (sparse/iterative solvers), preconditioning, domain decomposition
    • Power grid simulation / transient stability / optimization workflows
  • Experience with CI/perf regression testing, containerized workflows (Docker/Apptainer), and schedulers (Slurm).
  • Familiarity with AI workflows used alongside simulation (data prep, training/inference integration, pipeline performance).

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!, , JR2018521

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Energy HPC CUDA Performance Engineer
Energy HPC CUDA Performance Engineer

NVIDIA • United Arab Emirates

On-site
AED 420,000 - 660,000
Solutions Architect - AI Factory, South Africa
Solutions Architect - AI Factory, South Africa

NVIDIA Corporation • Dubai

On-site
AED 300,000 - 540,000
Solutions Architect - AI Factory, North Africa
Solutions Architect - AI Factory, North Africa

NVIDIA Corporation • Dubai

On-site
AED 661,000 - 1,212,000
Solutions Architect - AI Factory, North Africa
Solutions Architect - AI Factory, North Africa

NVIDIA • Dubai

On-site
AED 500,000 - 900,000
Technical Lead - GPU Infrastructure (100% Remote - Worldwide) at Tether Operations Limited
Technical Lead - GPU Infrastructure (100% Remote - Worldwide) at Tether Operations Limited

Tether Operations Limited • Abu Dhabi

Remote
AED 661,000 - 881,000
Senior HPC Engineer – IFM
Senior HPC Engineer – IFM

The Chronicle Of Higher Education, Inc. • United Arab Emirates

On-site
AED 223,200 - 334,800
Solutions Architect - AI Factory, North Africa
Solutions Architect - AI Factory, North Africa

NVIDIA • United Arab Emirates

On-site
AED 240,000 - 420,000
Global travel up to 20%
Competitive benefits
High Performance Computing Software Engineer - Supercomputing
High Performance Computing Software Engineer - Supercomputing

Institute of Foundation Models • Abu Dhabi

On-site
AED 350,000 - 700,000
Site Reliability Engineer – GPU/HPC Infrastructure (Remote – MENA)
Site Reliability Engineer – GPU/HPC Infrastructure (Remote – MENA)

Saturn Cloud • United Arab Emirates

Remote
AED 360,000 - 600,000
AI Research Engineer (Kernel & Inference Optimization)
AI Research Engineer (Kernel & Inference Optimization)

Lever, Inc. • United Arab Emirates

Remote
AED 350,000 - 700,000
Remote-first
International team
Cutting-edge research
+2