Member of Technical Staff, ML Kernels

Netpreme

Boston (MA)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity awards
Health insurance
401(k) with company matching
Charity donation matching
Visa sponsorship
Flexible PTO

Job summary

Netpreme is seeking a skilled professional to build highly optimized ML kernels for GPUs, FPGAs, and custom silicon. Responsibilities include designing and implementing compute and data movement kernels, targeting multiple hardware platforms, and collaborating with hardware architects and ML researchers. The ideal candidate should have solid experience in programming accelerators and knowledge of performance optimizations. The role offers equity awards, comprehensive insurance, 401(k) matching, and flexible PTO.

Qualifications

  • Solid experience in programming accelerators (GPUs, FPGAs).
  • Deep knowledge of at least one accelerator programming model (e.g. CUDA).
  • Interest in performance optimizations and enthusiasm about extracting maximum efficiency.
  • Knowledge of high-level synthesis (HLS) is a strong plus.

Responsibilities

  • Design and implement highly-optimized compute and data movement kernels for ML workloads.
  • Target multiple hardware platforms: GPUs, FPGAs, custom ASICs.
  • Work closely with hardware architects and ML researchers.

Skills

Programming accelerators (GPUs, FPGAs)
CUDA
Performance optimizations
High-level synthesis (HLS)

Job description

Build highly optimized ML kernels for GPUs, FPGAs, and custom silicon.

Location: Boston / Santa Clara

In this role, you will
  • design and implement highly-optimized compute and data movement kernels for ML workloads.
  • target multiple hardware platforms: GPUs, FPGAs, custom ASICs.
  • work closely with hardware architects and ML researchers.
Role requirements
  • solid experience in programming accelerators (GPUs, FPGAs).
  • deep knowledge of at least one accelerator programming model (e.g. CUDA).
  • interest in performance optimizations and enthusiasm about extracting maximum efficiency from the underlying hardware.
  • knowledge of high-level synthesis (HLS) is a strong plus.
Perks
  • Equity awards granting you ownership of the company, in addition to salary.
  • Full health, dental, vision, disability, and life insurance.
  • 401(k) with company matching.
  • 100% charity donation matching.
  • Visa sponsorship and relocation support to assist you in moving to our offices in Boston or Santa Clara.
  • Flexible PTO and remote work options.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Kernel Architect: GPUs/FPGA/ASICs, Equity & Remote
ML Kernel Architect: GPUs/FPGA/ASICs, Equity & Remote

Netpreme • Boston (MA)

Hybrid
USD 100,000 - 130,000
Equity awards
Health insurance
401(k) with company matching
+3
Principal Software Engineer - Kernels
Principal Software Engineer - Kernels

d-Matrix • Santa Clara (CA)

Hybrid
USD 120,000 - 180,000
Member of Technical Staff, ML Systems
Member of Technical Staff, ML Systems

Netpreme • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Relocation assistance
Visa sponsorship
Lunch stipend
+1
Founding GPU Kernel Engineer
Founding GPU Kernel Engineer

SF Tensor • San Francisco (CA)

On-site
USD 285,000 - 315,000
Member of Technical Staff, ML Systems
Member of Technical Staff, ML Systems

Netpreme • Cambridge (MA)

On-site
USD 190,000 - 230,000
Relocation assistance
Visa sponsorship
Daily lunch stipend
+2
Principal Software Engineer – Kernels
Principal Software Engineer – Kernels

MixMode • Santa Clara (CA)

Hybrid
USD 150,000 - 200,000
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Kernel Engineer (Compute / Accelerator)
Kernel Engineer (Compute / Accelerator)

DensityAI • Mountain View (WY)

On-site
USD 260,000 - 320,000
Equity grant
Medical/dental/vision
401(k)
+1
Member of Technical Staff - Kernels & GPU Performance
Member of Technical Staff - Kernels & GPU Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Software Engineer, ML Performance & Systems
Staff Software Engineer, ML Performance & Systems

fal • San Francisco (CA)

On-site
USD 180,000 - 250,000
Competitive salary and equity
Visa sponsorship
Health, dental, and vision insurance
+1