High-Performance ML Runtime & Kernel Engineer

Cerebras Systems

Sunnyvale (CA)

On-site

USD 150,000 - 210,000

Full time

8 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Cerebras Systems is seeking an engineer for the Core ML team to bridge cutting-edge research with efficient execution on the Cerebras Wafer-Scale Engine.

You will work across ML frameworks, compilers, runtimes, and low-level kernels to turn research prototypes into robust, high-performance demonstrations. Depending on background, emphasis may be on runtime capabilities, kernel development, or both, contributing to scalable, real-time AI workloads.

Qualifications

  • Bachelor’s, Master’s, or PhD in CS/CE/EE or related field.
  • Experience developing high-performance systems software.
  • Strong C++ and Python programming skills.
  • Familiarity with PyTorch or JAX.
  • Ability to translate research into production software.

Responsibilities

  • Design and implement runtime components and high-performance kernels for Core ML.
  • Translate research prototypes into Cerebras platform implementations.
  • Profile and debug performance across ML stack layers.
  • Optimize compute, memory, and communication for large-scale training.
  • Develop benchmarks, instrumentation, and tests for correctness.

Skills

C++
Python
Performance profiling
Parallel programming
Debugging
PyTorch
JAX
Production deployment

Education

Bachelor's/Master's/PhD in CS/CE/EE

Tools

CUDA
Triton
LLVM

Job description

Cerebras Systems is seeking an engineer for the Core ML team to bridge cutting-edge research with efficient execution on the Cerebras Wafer-Scale Engine.

You will work across ML frameworks, compilers, runtimes, and low-level kernels to turn research prototypes into robust, high-performance demonstrations. Depending on background, emphasis may be on runtime capabilities, kernel development, or both, contributing to scalable, real-time AI workloads.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Runtime and Kernel Engineer - Core ML
ML Runtime and Kernel Engineer - Core ML

Cerebras Systems • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
New Grad Kernel Engineer for High-Performance AI
New Grad Kernel Engineer for High-Performance AI

Cerebras • Sunnyvale (CA)

On-site
USD 150,000 - 230,000
ML Algorithm Mapping and Performance Engineer, Core ML
ML Algorithm Mapping and Performance Engineer, Core ML

Cerebras • United States

Remote
USD 140,000 - 220,000
New Grad Kernel Engineer - ML & HPC Systems
New Grad Kernel Engineer - ML & HPC Systems

Foundation Capital • Sunnyvale (CA)

On-site
USD 180,000 - 260,000
ML Systems Performance Engineer — Hardware Co-Design
ML Systems Performance Engineer — Hardware Co-Design

Cerebras Systems • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Kernel Engineer - High-Performance ML/HPC on Custom AI Chip
Kernel Engineer - High-Performance ML/HPC on Custom AI Chip

Foundation Capital • United States

On-site
USD 100,000 - 130,000
Opportunity to publish open-source AI research
Work with one of the fastest AI supercomputers
Non-corporate work culture
Kernel Engineer
Kernel Engineer

Foundation Capital • United States

On-site
USD 100,000 - 130,000
Opportunity to publish open-source AI research
Work with one of the fastest AI supercomputers
Non-corporate work culture
ML Algorithm Mapping and Performance Engineer, Core ML
ML Algorithm Mapping and Performance Engineer, Core ML

Cerebras Systems • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
ML Systems Performance Engineer
ML Systems Performance Engineer

Foundation Capital • United States

On-site
USD 100,000 - 130,000
ML Performance Engineer: Algorithm Mapping for AI Accelerators
ML Performance Engineer: Algorithm Mapping for AI Accelerators

Cerebras • United States

Remote
USD 140,000 - 220,000