ML Inference Performance Engineer - Kernel Optimizer

Cerebras Systems

United States

On-site

USD 110,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cerebras Systems is seeking engineers for the inference performance team, focusing on optimizing AI model inference speed on the world’s largest AI chip. This role involves performance modeling, debugging, and tooling development to enhance machine learning capabilities.

Ideal candidates should have a relevant degree, strong experience in computer architecture, and proficiency in C++ and Python. Join us at Cerebras to help shape the future of AI technology.

Qualifications

  • Strong background in computer architecture.
  • Exposure to and understanding of low-level deep learning / LLM math.
  • 3+ years of experience in a relevant domain.

Responsibilities

  • Build performance models to estimate ML model performance.
  • Optimize and debug kernel micro code for higher inference speed.
  • Develop tools to visualize performance data.

Skills

Kernel optimization
C++
Python
Performance profiling
Problem-solving

Education

Bachelors / Masters / PhD in Electrical Engineering or Computer Science

Tools

CPU/GPU simulators

Job description

Cerebras Systems is seeking engineers for the inference performance team, focusing on optimizing AI model inference speed on the world’s largest AI chip. This role involves performance modeling, debugging, and tooling development to enhance machine learning capabilities.

Ideal candidates should have a relevant degree, strong experience in computer architecture, and proficiency in C++ and Python. Join us at Cerebras to help shape the future of AI technology.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Inference Performance Architect
ML Inference Performance Architect

Cerebras • United States

On-site
USD 100,000 - 130,000
ML Systems Performance Engineer
ML Systems Performance Engineer

Cerebras Systems • United States

On-site
USD 110,000 - 150,000
Senior AI Inference Performance Engineer
Senior AI Inference Performance Engineer

Cerebras Systems, Inc. • Sunnyvale (CA)

On-site
USD 130,000 - 160,000
Opportunities for contributing to open-source projects
Stable work environment with startup vitality
ML Systems Performance Engineer
ML Systems Performance Engineer

Cerebras • United States

On-site
USD 100,000 - 130,000
Kernel Engineer - High-Performance ML/HPC on Custom AI Chip
Kernel Engineer - High-Performance ML/HPC on Custom AI Chip

Cerebras Systems • United States

On-site
USD 110,000 - 140,000
Non-corporate work culture
Equal opportunity employer
Continuous learning and growth opportunities
AI Hardware Kernel Performance Engineer
AI Hardware Kernel Performance Engineer

Cerebras • Sterling (VA)

On-site
USD 100,000 - 150,000
Kernel Engineer - High-Performance ML/HPC on Custom AI Chip
Kernel Engineer - High-Performance ML/HPC on Custom AI Chip

Cerebras • United States

On-site
USD 100,000 - 130,000
Opportunity to publish open-source AI research
Work with one of the fastest AI supercomputers
Non-corporate work culture
CoDesign & NextGen Performance Engineer
CoDesign & NextGen Performance Engineer

Cerebras • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
High-Performance Kernel Engineer for AI & HPC
High-Performance Kernel Engineer for AI & HPC

Cerebras • Raleigh (NC)

On-site
USD 100,000 - 140,000
Equal opportunity work environment
Continuous learning and support
Diverse team culture
Kernel Engineer
Kernel Engineer

Cerebras • Raleigh (NC)

On-site
USD 100,000 - 140,000
Equal opportunity work environment
Continuous learning and support
Diverse team culture