Host and Network IO FPGA Engineer

Cerebras

Toronto

On-site

CAD 120,000 - 180,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cerebras Systems is seeking an experienced FPGA developer to own the in-chassis IO subsystem, including RoCE interfaces and a large programmable switching fabric, interfacing with AI application teams and embedded software. You will help optimize bandwidth and latency across multiple generations of hardware.

The role involves designing next-gen IO architectures, producing production-ready bitstreams, and leading cross-functional projects spanning hardware and software teams to deliver improved

Qualifications

  • Production FPGA development experience or advanced degree with equivalent industry experience.
  • Proficiency in Verilog in Git-based repositories.
  • Strong FPGA placement, routing, timing closure, and debugging skills.
  • Familiarity with TCP/RoCE networking and network debugging tools.

Responsibilities

  • Lead full chassis-to-wafer IO architecture and design.
  • Improve current RTL, implement next-gen FPGA design, define future IO architectures.
  • Produce production-ready bitstreams for deployment into large clusters running customer inference services.
  • Collaborate with DV team to prevent bugs and streamline debugging.
  • Optimize bandwidth/latency over FPGA datapath and all IO interfaces.
  • Interface with board team for board bringup activities.
  • Drive network performance debug of large AI clusters.
  • Integrate cutting-edge networking technologies and protocols.
  • Lead cross-functional technical projects across hardware and software teams.

Skills

FPGA development
Verilog
FPGA tooling
TCP/RoCE networking
Wireshark
AI-augmented development environment

Education

Master's/PhD in Computer or Electrical Engineering

Tools

Git
Arista
Juniper

Job description

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.

Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

About The Role

The Host and Network IO Team develops the full IO path implementation between a distributed system of server nodes, through the cluster, down to the custom RoCE network stack implemented in Cerebras' system, and over the proprietary IOs onto the WSE. As an FPGA developer on the team, you will own the in-chassis IO subsystem consisting of i) several cluster-facing RoCE v2 network interfaces via a custom implementation of the RDMA protocol; ii) a large programmable switching fabric; and iii) Serial IO communication with the Cerebras WSE via a proprietary protocol. You will interface between AI application-level IO teams, cluster architecture teams, and embedded software teams to develop solutions that optimize bandwidth and latency while minimizing congestion, pauses, pause spreading, unfairness, etc. The scope of work spans multiple generations of hardware products from improvements to presently deployed hardware, implementation of upcoming systems, and design/architecting of future next-gen architectures.

Responsibilities
  • Lead full chassis-to-wafer IO architecture and design
  • Improve current RTL, implement next-gen FPGA design, define future IO architectures
  • Produce production-ready bitstreams for deployment into large clusters running customer inference services
  • Work with DV team to prevent bug slips and simplify debug
  • Optimize bandwidth/latency over FPGA datapath and all IO interfaces
  • Interface with board team facilitating and executing board bringup
  • Drive network performance debug of large AI clusters
  • Integrate leading edge networking technologies and protocols
  • Lead cross-functional technical projects spanning multiple teams and integrating diverse software and hardware components to deliver an improved network IO solution.
  • Foster clear and effective communication across teams and stakeholders.
Skills & Qualifications
  • 5+ years industry experience generating production FPGA solutions, OR Master's/PhD in Computer or Electrical Engineering + 3 years industry experience,
  • Proficiency in Verilog development in Git based repo
  • Highly productive with FPGA placement & routing, timing closure, simulation, debug, and other FPGA tools/workflows
  • Network protocol familiarity (TCP, RoCE) and network debug tools such as Wireshark
  • Knowledge of network switch environments or willingness to learn (Arista, Juniper, etc.).
  • AI-augmented development environment
Why Join Cerebras
About

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:

  • Build a breakthrough AI platform beyond the constraints of the GPU.
  • Publish and open source their cutting-edge AI research.
  • Work on one of the fastest AI supercomputers in the world.
  • Enjoy job stability with startup vitality.
  • Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Distributed Software Engineer
Distributed Software Engineer

Cerebras Systems, Inc. • Ottawa

On-site
CAD 90,000 - 120,000
Job stability with startup vitality
Open access to cutting-edge AI research
CoDesign & NextGen Performance Engineer
CoDesign & NextGen Performance Engineer

Cerebras Systems • Toronto

On-site
CAD 120,000 - 180,000
Cluster Operations Software Engineer
Cluster Operations Software Engineer

Cerebras Systems • Toronto

On-site
CAD 120,000 - 160,000
Cluster Operations Software Engineer
Cluster Operations Software Engineer

Cerebras • Toronto

On-site
CAD 120,000 - 190,000
Senior Runtime Engineer
Senior Runtime Engineer

Cerebras • Toronto

On-site
CAD 100,000 - 150,000
Non-corporate work culture
Stability with startup vitality
Opportunities for continuous learning
Staff Software Engineer, GPU Inference
Staff Software Engineer, GPU Inference

Cerebras • Toronto

On-site
CAD 150,000 - 210,000
ML Performance Benchmarking Engineer
ML Performance Benchmarking Engineer

Cerebras Systems, Inc. • Toronto

Hybrid
CAD 80,000 - 110,000
Job stability with startup vitality
Open-source cutting-edge AI research
Non-corporate work culture
ML Systems Integration Engineer
ML Systems Integration Engineer

Cerebras • Toronto

On-site
CAD 90,000 - 150,000
Senior Software Development Engineer in Test (SDET) - AI Cluster
Senior Software Development Engineer in Test (SDET) - AI Cluster

Cerebras • Toronto

On-site
CAD 120,000 - 190,000
DevOps Engineer - New Grad 2026
DevOps Engineer - New Grad 2026

Cerebras Systems, Inc. • Toronto

On-site
CAD 70,000 - 90,000
Opportunity to work on an innovative AI platform
Diverse and inclusive work environment
Job stability with startup vitality