Machine Learning Engineer - Kernels

7Seventy

Northern (KY)

Presencial

USD 150.000 - 190.000

Jornada completa

hace 27 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Consigue una respuesta de este empleador — un currículum y una carta de presentación adaptados exactamente a lo que busca para contratar.

Supera los filtros ATS

Descripción de la vacante

7Seventy is seeking a Machine Learning Engineer - Kernels focused on custom GPU and accelerator kernel development for high-performance AI workloads. You will design low-level implementations, optimize performance, and translate advances in ML into production-ready code.

The role blends GPU programming, parallel computing, hardware-aware optimization, and ML infrastructure, with work spanning C++, CUDA, performance profiling, benchmarking, and optimization in distributed and heterogeneous

Formación

  • 2+ years of GPU programming or parallel computing.
  • Experience optimizing workloads for distributed/heterogeneous computing.
  • Strong C++ and CUDA programming skills.
  • Familiarity with ML backends and profiling tools.

Responsabilidades

  • Design, develop, and implement custom GPU/accelerator kernels for ML workloads.
  • Profile and benchmark ML workloads; identify bottlenecks; implement optimizations.
  • Collaborate with ML researchers to translate advances into production-ready code.
  • Evaluate accelerator hardware and programming tech (CUDA, ROCm, TPUs) to guide kernel design.
  • Document optimization techniques and share best practices across teams.

Conocimientos

GPU kernel development
Parallel computing
Performance profiling
Distributed computing
Hardware acceleration
C++ programming
Low-level optimization
Collaborating with researchers

Educación

Bachelor's degree in Computer Science or Electrical Engineering
Master's degree or PhD preferred

Herramientas

CUDA
ROCm
TPUs

Descripción del empleo

Experience: 2+ years in GPU programming, parallel computing, or systems-level optimization

Core Areas: GPU Kernel Development, Machine Learning Workload Optimization, CUDA, C++, Parallel Computing, Performance Profiling, Hardware Acceleration, Distributed Computing

Compensation: $150,000 – $190,000 per year

About the Role

This opportunity is for a Machine Learning Engineer - Kernels specializing in custom GPU and accelerator kernel development for high-performance AI and machine learning workloads. The role focuses on designing efficient low-level implementations, optimizing computational performance, and translating advances in machine learning algorithms into reliable, production-ready code.

The position combines GPU programming, parallel computing, hardware-aware optimization, and machine learning infrastructure. Work involves C++, CUDA, performance profiling, benchmarking, and optimization across distributed and heterogeneous computing environments. The engineer will collaborate with researchers, evaluate hardware developments involving CUDA, ROCm, and TPUs, and contribute to engineering practices that improve computational efficiency for advanced AI workloads.

What You'll Do
  • Design, develop, and implement custom GPU and accelerator kernels to maximize computational performance for machine learning workloads.
  • Profile and benchmark performance-critical ML workloads, identify computational bottlenecks, and implement low-level optimizations to improve execution efficiency.
  • Collaborate with machine learning researchers to translate algorithmic advances into efficient, production-ready software implementations.
  • Evaluate developments in accelerator hardware and programming technologies, including CUDA, ROCm, and TPUs, to guide kernel design and optimization decisions.
  • Document low-level optimization techniques and share performance engineering best practices with technical teams.
Qualifications

Required Experience

  • At least 2 years of professional experience in GPU programming, parallel computing, or systems-level performance optimization.
  • Experience optimizing computational workloads for distributed and heterogeneous computing environments.

Required Skills

  • Strong programming skills in C++, CUDA, or comparable low-level programming languages used for performance-intensive computing.
  • Familiarity with machine learning frameworks and their low-level computational backends.
  • Ability to use profiling tools and performance diagnostics to investigate execution behavior and identify optimization opportunities.
  • Strong understanding of performance-oriented programming and the interaction between algorithms and hardware.
  • Detail-oriented approach to engineering, with a strong focus on computational efficiency and performance optimization.
  • Ability to collaborate effectively on research-driven engineering challenges and contribute to technically ambitious development work.

Education

  • Bachelor's, Master's, or Ph.D. degree in Computer Science, Electrical Engineering, or a related field, or equivalent practical experience.
Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Kernel Engineer
Kernel Engineer

Acceler8 Talent • San Francisco (CA)

Presencial
USD 180.000 - 240.000
GPU Kernel Engineer — High-Performance ML & Acceleration
GPU Kernel Engineer — High-Performance ML & Acceleration

7Seventy • Northern (KY)

Presencial
USD 150.000 - 190.000
GPU Kernel Engineer — High-Performance ML at Scale
GPU Kernel Engineer — High-Performance ML at Scale

The Consensus • San Francisco (CA)

Presencial
USD 120.000 - 160.000
Competitive compensation with equity
100% medical, dental, and vision insurance coverage
Flexible PTO policy including a Winter Break
+2
CUDA Engineer - Kernel Optimization - AI Trainer
CUDA Engineer - Kernel Optimization - AI Trainer

Mercor • Chicago (IL)

Presencial
USD 83.000 - 165.000
CUDA Engineer - Kernel Optimization
CUDA Engineer - Kernel Optimization

Mercor • San Francisco (CA)

Presencial
USD 165.000 - 276.000
GPU Kernel Developer - Fully Remote | Upto $90/hr
GPU Kernel Developer - Fully Remote | Upto $90/hr

mercor • EE. UU.

A distancia
USD 96.000 - 124.000
Software Engineer - GPU Kernel
Software Engineer - GPU Kernel

FriendliAI • San Francisco (CA)

Presencial
USD 120.000 - 150.000
Flexible working hours
Daily lunch and dinner
Health check-up support
+3
Member of Technical Staff - Kernels & GPU Performance
Member of Technical Staff - Kernels & GPU Performance

Gimlet Labs • San Francisco (CA)

Presencial
USD 120.000 - 160.000
Senior Kernel Engineer: 26-02794
Senior Kernel Engineer: 26-02794

Akraya, Inc. • Bellevue (WA)

Presencial
USD 117.000 - 124.000
Machine Learning Engineer — GPU Kernel
Machine Learning Engineer — GPU Kernel

Institute of Foundation Models • Sunnyvale (CA)

Presencial
USD 150.000 - 450.000
Comprehensive medical, dental, and视觉 V
Bonus
401K Plan
+4