Developer Technology Engineer, Energy

NVIDIA

United Arab Emirates

On-site

AED 550,000 - 1,000,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

NVIDIA in United Arab Emirates is seeking a world-class Compute Developer Technology (DevTech) engineer to accelerate energy simulation and AI workflows on NVIDIA platforms. You will focus on CUDA performance optimization for workloads such as seismic processing, reservoir simulation, and power grid simulations.

You will collaborate with customer and partner teams as well as NVIDIA product and engineering groups to deliver measurable speedups on multi-GPU and multi-node systems, build

Qualifications

  • BS/MS in CS/CE/EE/Physics/Applied Math or equivalent experience.
  • Strong programming skills in C/C++ and Python on Linux.
  • Hands-on experience with CUDA programming and GPU performance optimization concepts.
  • Experience profiling and debugging performance using tools such as NVIDIA Nsight Systems / Nsight Compute (or equivalent).
  • Understanding of parallel computing and performance fundamentals (vectorization, threading, NUMA, memory bandwidth/latency).
  • Ability to communicate technical findings clearly to both engineers and non-engineers.
  • 5+ years relevant experience in GPU/HPC optimization; strong track record of delivered speedups and scaling improvements.

Responsibilities

  • Profile, analyze, and optimize GPU-accelerated applications with emphasis on CUDA kernels, memory movement, concurrency, and end-to-end throughput.
  • Drive performance improvements across the stack: CUDA C++ kernel optimization, launch configuration, memory hierarchy, streams/events.
  • Build reproducible benchmarks, performance reports, and tuning recommendations (before/after, methodology, scaling curves).
  • Develop and maintain reference implementations, examples, and/or patches to customer code to enable performance and portability.
  • Support customer engagements (POCs to production), including debugging correctness/performance issues and advising on best practices for deployment (containers, schedulers, clusters).
  • Collaborate with internal teams to file actionable issues, validate fixes, and influence roadmap based on real customer requirements in Energy.
  • Build internal libraries and reusable code that would lead to future NVIDIA products.

Skills

CUDA programming
C/C++
Python
Linux
GPU performance
Nsight tools
MPI/NCCL

Education

BS/MS in CS/CE/EE/Physics/Applied Math

Tools

NVIDIA Nsight Systems
NVIDIA Nsight Compute
MPI
NCCL
Docker/Apptainer

Job description

Our work at NVIDIA is dedicated towards a computing model focused on visual and AI computing. For two decades, NVIDIA has pioneered visual computing, the art and science of computer graphics, with our invention of the GPU. The GPU has also shown to be spectacularly effective at solving some of the most complex problems in computer science. Today, NVIDIA’s GPU simulates human intelligence, running deep learning algorithms and acting as the brain of computers, robots and self-driving cars that can perceive and understand the world. We are looking to grow our company and teams with the smartest people in the world and there has never been a more exciting time to join our team!

NVIDIA is looking for a passionate, world-class computer scientists and engineers (Compute Developer Technology - DevTech) to accelerate Energy simulation and AI workflows on NVIDIA platforms. You will focus on CUDA performance optimization for workloads such as seismic processing (e.g., imaging/inversion pipelines), reservoir simulation, power grid simulators, and related HPC/AI production workflows. You will work hands-on with customer and partner engineering teams as well as NVIDIA product and engineering groups to deliver measurable speedups and scalable performance on multi-GPU and multi-node systems.

What You Will Be Doing
  • Profile, analyze, and optimize GPU-accelerated applications with emphasis on CUDA kernels, memory movement, concurrency, and end-to-end throughput.
  • Drive performance improvements across the stack:
    • CUDA C++ kernel optimization, launch configuration, memory hierarchy, streams/events
    • GPU libraries (as applicable): cuBLAS, cuFFT, cuSPARSE, cuSOLVER, NCCL
    • Multi-GPU and multi-node scaling using MPI + NCCL, CPU/GPU overlap, communication patterns
  • Build reproducible benchmarks, performance reports, and tuning recommendations (before/after, methodology, scaling curves).
  • Develop and maintain reference implementations, examples, and/or patches to customer code to enable performance and portability.
  • Support customer engagements (POCs to production), including debugging correctness/performance issues and advising on best practices for deployment (containers, schedulers, clusters).
  • Collaborate with internal teams to file actionable issues, validate fixes, and influence roadmap based on real customer requirements in Energy.
  • Build internal libraries and resuable code that would lead to future NVIDIA products.
What We Need To See
  • BS/MS (or equivalent experience) in CS/CE/EE/Physics/Applied Math or related field.
  • Strong programming skills in C/C++ and Python on Linux.
  • Hands-on experience with CUDA programming and GPU performance optimization concepts.
  • Experience profiling and debugging performance using tools such as NVIDIA Nsight Systems / Nsight Compute (or equivalent).
  • Understanding of parallel computing and performance fundamentals (vectorization, threading, NUMA, memory bandwidth/latency).
  • Ability to communicate technical findings clearly to both engineers and non-engineers.
  • 5+ years relevant experience in GPU/HPC optimization; strong track record of delivered speedups and scaling improvements.
Ways To Stand Out From The Crowd
  • Leads performance reviews with customer stakeholders; creates reusable playbooks/reference designs. Experience/Skills (typical)
  • HPC experience with MPI, distributed systems, and multi-node performance tuning.
  • Energy/HPC domain exposure:
    • Seismic processing pipelines, RTM/FWI-style patterns, FFT/stencil/linear algebra heavy codes
    • Reservoir simulation (sparse/iterative solvers), preconditioning, domain decomposition
    • Power grid simulation / transient stability / optimization workflows
  • Experience with CI/perf regression testing, containerized workflows (Docker/Apptainer), and schedulers (Slurm).
  • Familiarity with AI workflows used alongside simulation (data prep, training/inference integration, pipeline performance).

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Energy HPC CUDA Performance Engineer
Energy HPC CUDA Performance Engineer

NVIDIA • United Arab Emirates

On-site
AED 550,000 - 1,000,000
CUDA Performance Engineer - GPU Kernels
CUDA Performance Engineer - GPU Kernels

YO IT Consulting • Al Ruways Industrial City

On-site
AED 279,000 - 502,000
GPU Software Engineer - AI Acceleration
GPU Software Engineer - AI Acceleration

YO IT Consulting • Al Ruways Industrial City

On-site
AED 279,000 - 502,000
Remote contractor
Flexible hours
CUDA Engineering Expert - GPU Optimization
CUDA Engineering Expert - GPU Optimization

YO IT Consulting • Al Ruways Industrial City

On-site
AED 25,000 - 45,000
CUDA Engineering Expert - GPU Optimization
CUDA Engineering Expert - GPU Optimization

YO IT Consulting • Dubai

On-site
AED 246,000 - 391,000
Senior System Software Engineer
Senior System Software Engineer

NVIDIA • Dubai

On-site
AED 350,000 - 550,000
Competitive salary
Benefits package
Remote GPU Software Engineer - AI Acceleration
Remote GPU Software Engineer - AI Acceleration

YO IT Consulting • Dubai

Remote
AED 246,000 - 402,000
Senior HPC Engineer – IFM
Senior HPC Engineer – IFM

The Chronicle Of Higher Education, Inc. • United Arab Emirates

On-site
DevOps / IT Infra Engineer
DevOps / IT Infra Engineer

Evollabs Tech • Dubai Emirate

On-site
AED 300,000 - 520,000
High Performance Computing Software Engineer - Supercomputing
High Performance Computing Software Engineer - Supercomputing

Institute of Foundation Models • Abu Dhabi

On-site
AED 350,000 - 700,000