Member of Technical Staff (GPU Performance Engineer)

Reka

United States

Remote

USD 120,000 - 150,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Five weeks of paid leave
Comprehensive healthcare benefits
Visa support for H1B and OPT transfers

Job summary

A global AI foundation model startup is seeking an experienced GPU Performance Engineer to enhance training infrastructure and optimize model performance. The ideal candidate will have strong skills in Python and experience with large-scale model training, including GPU code optimization. This role offers a collaborative environment featuring top-tier talent and generous benefits, including extensive paid leave and visa support. Join a team committed to AI innovation and excellence.

Qualifications

  • Strong engineering skills with fluency in Python and PyTorch (or other frameworks).
  • Proven experience implementing and training large deep learning models.
  • Experience writing and debugging low-level GPU code (CUDA, C++).
  • Experience scaling GPU jobs using large-scale compute clusters.
  • Ability to analyze and optimize GPU-accelerated workloads.

Responsibilities

  • Design and implement improvements to training infrastructure.
  • Contribute to technical decisions optimizing model performance.
  • Work on post-training processes including reinforcement learning and fine-tuning.
  • Improve efficiency and scalability of model serving infrastructure.

Skills

Fluent in Python
Proficient in PyTorch
Experience with CUDA
Large scale model training
Performance optimization

Job description

We are seeking an experienced GPU Performance Engineer with a strong background in Python and large-scale model training. In this role, you will design and implement improvements to our training infrastructure and directly contribute to technical decisions that optimize performance of our models. You will also work on post-training processes, including reinforcement learning and fine-tuning. Furthermore, you will contribute to improving the efficiency and scalability of our model serving infrastructure.

Ideal Experience
  • Strong engineering skills with fluency in Python and PyTorch (or other frameworks).
  • Proven experience implementing and training large deep learning models.
  • Experience writing and debugging low-level GPU code (CUDA, C++).
  • Experience scaling up GPU jobs using large-scale compute clusters (e.g., Slurm or Kubernetes).
  • Demonstrated ability to analyze and optimize the performance of GPU-accelerated workloads, including profiling, identifying bottlenecks, and implementing performance tuning techniques.
Reka's Mission

Reka's mission is to build useful multimodal artificial intelligence and use it to empower organizations and businesses. We are a globally distributed foundation model startup, headquartered in the San Francisco Bay Area, California. Embracing a remote-first approach, our team brings together top talent from around the world. Our founding team, along with many of our team members, has contributed to numerous breakthroughs in AI over the past decade.

Why Reka?
  • An Elite Team: Collaborate with top-tier engineers, researchers, and operators from renowned organizations like Google DeepMind, Facebook AI Research (FAIR), and successful startups, driving innovation in AI technology.
  • Cutting-edge Infrastructure: Train state-of-the-art models leveraging the latest software and hardware, expanding the frontier of innovation in AI infrastructure development.
  • Inclusive and Open Culture: Thrive in an open and inclusive work environment that values diverse perspectives and fosters creativity.
  • Generous Benefits: Enjoy five weeks of paid leave to recharge, comprehensive healthcare benefits (including vision and dental), and additional perks that support your well-being.
  • Visa Support: We provide visa assistance, including H1B and OPT transfers, for US employees to ensure a smooth transition and support your career with us.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff (GPU Performance Engineer) Mon, 07 Sep 2026 19:38:55 +0000
Member of Technical Staff (GPU Performance Engineer) Mon, 07 Sep 2026 19:38:55 +0000

Megalojobs • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Five weeks of paid leave
Healthcare (vision and dental)
Visa assistance (H1B/OPT)
Member of Technical Staff (Machine Learning Engineer)
Member of Technical Staff (Machine Learning Engineer)

Reka AI • United States

On-site
USD 120,000 - 160,000
5 weeks of paid leave
Comprehensive healthcare benefits
Visa assistance for US employees
GPU Performance Engineer - Remote, PTO & Visa Support
GPU Performance Engineer - Remote, PTO & Visa Support

Megalojobs • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Five weeks of paid leave
Healthcare (vision and dental)
Visa assistance (H1B/OPT)
Member of Sales Staff
Member of Sales Staff

Neura Market • United States

On-site
USD 90,000 - 150,000
5 weeks paid leave
Healthcare (vision & dental)
Remote-friendly culture
Member of Technical Staff - Mid-Training Infra
Member of Technical Staff - Mid-Training Infra

Reflection • New York (NY)

On-site
USD 150,000 - 200,000
Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
Daily meals provided
+2
Program Lead
Program Lead

Reka AI • San Francisco (CA)

On-site
USD 140,000 - 200,000
5 weeks paid leave
Comprehensive healthcare (vision &
Dental
Member of Technical Staff, Post-Training & Applied Research
Member of Technical Staff, Post-Training & Applied Research

SF Tensor • San Francisco (CA)

On-site
USD 275,000 - 315,000
Relocation assistance
Equity
Software Engineer, AI Infra
Software Engineer, AI Infra

Makers Fund • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Monthly stipends
+1
Supply Lead
Supply Lead

Neura Market • Northern (KY)

Hybrid
USD 90,000 - 150,000
5 weeks paid leave
Health, vision and dental benefits
Remote-first culture
ML Researcher - Image / Video Diffusion
ML Researcher - Image / Video Diffusion

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Health insurance
401k with 4% company match
Meals in the office
+2