Kernel Engineer - New Grad

Cerebras

United States

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cerebras Systems is seeking a Kernel Engineer to develop high-performance software at the intersection of hardware and software for AI and HPC workloads. You will map ML and linear algebra operations to the Cerebras Wafer-Scale Engine, optimizing compute utilization and system performance.

You will collaborate with experienced kernel, compiler, performance, and hardware engineers, learn architecture details, and contribute to kernel libraries and validation methodologies.

Qualifications

  • Bachelor’s, Master’s, or PhD in CS, CE, EE, math or related field.
  • Strong C++ fundamentals and Python familiarity.
  • Understanding of computer architecture concepts.
  • Knowledge of data structures, algorithms, and software fundamentals.
  • Experience debugging software through coursework, internships, research, co-op placements, or technical projects.
  • Strong analytical and problem-solving skills.
  • Interest in low-level software, parallel computing, performance optimization, or hardware/software co-design.
  • Ability to learn unfamiliar systems and collaborate effectively within a technical team.

Responsibilities

  • Design and implement ML and linear algebra kernels for Cerebras Wafer-Scale Engine.
  • Develop and debug high-performance kernel routines using the Cerebras Software Language.
  • Apply parallel programming to map workloads onto Cerebras architecture.
  • Use analysis, data, and profiling tools to evaluate kernel behavior and guide design.
  • Investigate correctness, performance, and hardware utilization issues.
  • Develop unit tests and system validations for kernel libraries.
  • Collaborate with kernel, compiler, performance, and hardware engineers to improve software.
  • Study ML workloads and evolve the kernel library.
  • Participate in code reviews and software development processes.
  • Build understanding of Cerebras architecture and memory system.

Skills

C++
Python
Arch concepts
Data structures
Debugging
Parallel computing
Co-design
Team collaboration

Education

Bachelor’s/Master’s/PhD in CS/CE/EE/Math

Tools

Cerebras Software Language
CUDA
OpenCL
Assembly

Job description

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.

Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

About the Role

As a Kernel Engineer at Cerebras, you will develop high-performance software at the intersection of hardware and software for cutting-edge artificial intelligence and high-performance computing workloads.

You will help implement, optimize, and validate machine learning and linear algebra operations for the Cerebras Wafer-Scale Engine, our custom massively parallel processor architecture. Working alongside experienced kernel, compiler, performance, and hardware engineers, you will learn how algorithms are mapped to specialized hardware and contribute to software that maximizes compute utilization and system performance.

You will be part of a team responsible for the design, development, performance tuning, and validation of foundational ML and HPC kernels. This is an excellent opportunity for a new graduate who is interested in computer architecture, parallel programming, low-level software, and machine learning systems.

Responsibilities
  • Help design and implement machine learning and linear algebra kernels for the Cerebras Wafer-Scale Engine.

  • Develop and debug high-performance kernel routines using low-level programming techniques and the Cerebras Software Language, a custom C-like language.

  • Apply parallel programming algorithms to map computational workloads efficiently onto the Cerebras architecture.

  • Use mathematical analysis, performance data, and profiling tools to evaluate kernel behavior and inform design decisions.

  • Identify and investigate correctness, performance, and hardware utilization issues.

  • Develop unit tests and system-level validation methodologies to verify the functionality and performance of kernel libraries.

  • Collaborate with kernel, compiler, performance, and hardware engineers to improve software and system performance.

  • Study emerging machine learning workloads and contribute to the evolution of the kernel library.

  • Participate in code reviews, technical discussions, and software development processes.

  • Build an understanding of the Cerebras architecture, instruction set, memory system, and communication model.

Minimum Skills & Qualifications
  • Bachelor’s, Master’s, or PhD in Computer Science, Computer Engineering, Electrical Engineering, Mathematics, or a related field.

  • Strong programming fundamentals in C++ and familiarity with Python.

  • Understanding of foundational computer architecture concepts such as processors, memory hierarchies, instruction execution, or data movement.

  • Knowledge of data structures, algorithms, and software development fundamentals.

  • Experience debugging software through coursework, internships, research, co-op placements, or technical projects.

  • Strong analytical and problem-solving skills.

  • Interest in low-level software, parallel computing, performance optimization, or hardware/software co-design.

  • Ability to learn unfamiliar systems and collaborate effectively within a technical team.

Preferred Skills & Qualifications
  • Research, internships, or projects involving kernel development, compilers, computer architecture, HPC, or systems programming.

  • Familiarity with parallel algorithms, multithreaded programming, or distributed memory systems.

  • Exposure to programming accelerators such as GPUs, FPGAs, or other specialized processors.

  • Experience with low-level programming, assembly language, CUDA, OpenCL, or a domain-specific language.

  • Familiarity with machine learning concepts, neural networks, or frameworks such as PyTorch or TensorFlow.

  • Exposure to numerical computing, linear algebra, or HPC kernels.

  • Experience using profiling, benchmarking, or performance analysis tools.

  • Familiarity with library or API development practices.

Why Join Cerebras

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:

  1. Build a breakthrough AI platform beyond the constraints of the GPU.

  2. Publish and open source their cutting-edge AI research.

  3. Work on one of the fastest AI supercomputers in the world.

  4. Enjoy job stability with startup vitality.

  5. Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

EEO statement Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Kernel Optimzation Engineer
Staff Kernel Optimzation Engineer

Cerebras • Sterling (VA)

On-site
USD 100,000 - 150,000
Kernel Engineer
Kernel Engineer

Cerebras Systems • United States

On-site
USD 110,000 - 140,000
Non-corporate work culture
Equal opportunity employer
Continuous learning and growth opportunities
Kernel Engineer
Kernel Engineer

Cerebras • United States

On-site
USD 100,000 - 130,000
Opportunity to publish open-source AI research
Work with one of the fastest AI supercomputers
Non-corporate work culture
Kernel Engineer
Kernel Engineer

Cerebras • Raleigh (NC)

On-site
USD 100,000 - 140,000
Equal opportunity work environment
Continuous learning and support
Diverse team culture
ML Systems Integration Engineer
ML Systems Integration Engineer

Cerebras • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
ML Systems Integration Engineer
ML Systems Integration Engineer

Cerebras Systems • Sunnyvale (CA)

On-site
USD 140,000 - 190,000
ASIC Architect
ASIC Architect

Cerebras Systems, Inc. • Sunnyvale (CA)

On-site
USD 120,000 - 180,000
Commitment to diversity and inclusion
Non-corporate work culture
Job stability with startup vitality
ML Software Engineer - Integration & Quality - New Grad
ML Software Engineer - Integration & Quality - New Grad

Cerebras • Sunnyvale (CA)

Hybrid
USD 110,000 - 160,000
CoDesign & NextGen Performance Engineer
CoDesign & NextGen Performance Engineer

Cerebras • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
ML Software Tool Development Engineer
ML Software Tool Development Engineer

Cerebras • United States

On-site
USD 120,000 - 160,000
Equal opportunity employer
Non-corporate work culture
Continuous learning opportunities