Cerebras Systems Inc. Logo

Cerebras Systems Inc.

FPGA Engineer

Posted 21 Days Ago
Be an Early Applicant
Hybrid
Toronto, ON, CAN
Senior level
Hybrid
Toronto, ON, CAN
Senior level
Design and deliver production FPGA solutions for chassis-to-wafer IO including RoCE v2/RDMA network interfaces, switching fabric, and serial IO. Develop RTL, produce bitstreams, optimize bandwidth and latency, debug large AI cluster networking, coordinate board bringup, and lead cross-functional projects with DV, embedded software, and architecture teams.
The summary above was generated by AI

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.
Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.

About The Role

The Host and Network IO Team develops the full IO path implementation between a distributed system of server nodes, through the cluster, down to the custom RoCE network stack implemented in Cerebras' system, and over the proprietary IOs onto the WSE. As an FPGA developer on the team, you will own the in-chassis IO subsystem consisting of i) several cluster-facing RoCE v2 network interfaces via a custom implementation of the RDMA protocol; ii) a large programmable switching fabric; and iii) Serial IO communication with the Cerebras WSE via a proprietary protocol. You will interface between AI application-level IO teams, cluster architecture teams, and embedded software teams to develop solutions that optimize bandwidth and latency while minimizing congestion, pauses, pause spreading, unfairness, etc. The scope of work spans multiple generations of hardware products from improvements to presently deployed hardware, implementation of upcoming systems, and design/architecting of future next-gen architectures.

Responsibilities
  • Lead full chassis-to-wafer IO architecture and design

  • Improve current RTL, implement next-gen FPGA design, define future IO architectures

  • Produce production-ready bitstreams for deployment into large clusters running customer inference services

  • Work with DV team to prevent bug slips and simplify debug

  • Optimize bandwidth/latency over FPGA datapath and all IO interfaces

  • Interface with board team facilitating and executing board bringup

  • Drive network performance debug of large AI clusters

  • Integrate leading edge networking technologies and protocols

  • Lead cross-functional technical projects spanning multiple teams and integrating diverse software and hardware components to deliver an improved network IO solution.

  • Foster clear and effective communication across teams and stakeholders.

Skills & Qualifications
  • 5+ years industry experience generating production FPGA solutions, OR Master's/PhD in Computer or Electrical Engineering + 3 years industry experience,

  • Proficiency in Verilog development in Git based repo

  • Highly productive with FPGA placement & routing, timing closure, simulation, debug, and other FPGA tools/workflows

  • Network protocol familiarity (TCP, RoCE) and network debug tools such as Wireshark

  • Knowledge of network switch environments or willingness to learn (Arista, Juniper, etc.).

  • AI-augmented development environment

Why Join Cerebras

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:

  1. Build a breakthrough AI platform beyond the constraints of the GPU.

  2. Publish and open source their cutting-edge AI research.

  3. Work on one of the fastest AI supercomputers in the world.

  4. Enjoy job stability with startup vitality.

  5. Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

Apply today and become part of the forefront of groundbreaking advancements in AI!

Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer. We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.

Similar Jobs

Junior
Aerospace • Security • Software
Develop, optimize, test, and verify FPGA-based solutions using Vivado, VHDL, and HLS. Integrate FPGA designs with hardware and software systems, implement video interfaces and processing, maintain reusable IP cores, troubleshoot issues in laboratory and field environments, and contribute to architecture, proposals, estimates, and technical documentation.
Top Skills: CC++DisplayportFpgaGitHigh-Level Synthesis (Hls)Logic AnalyzersOscilloscopesPythonSdiTclVerilogVhdlXilinx Vivado
Junior
Greentech • Professional Services • Software • Analytics
Develop, optimize, test, and verify FPGA-based solutions using Vivado, VHDL, and HLS. Implement video interfaces and processing, maintain reusable IP cores, integrate FPGA designs with hardware and software systems, troubleshoot issues in lab and field environments, and contribute to architecture, proposals, estimates, and technical documentation.
Top Skills: Ai/Ml ModelsCC++DisplayportFpgaFpga Prototyping BoardsGitHigh-Level Synthesis (Hls)Logic AnalysersOscilloscopesPythonSdiTclVerilogVhdlXilinx Vivado
One Month Ago
In-Office
Toronto, ON, CAN
Senior level
Senior level
Artificial Intelligence • Internet of Things • Machine Learning
Develop and optimize FPGA placement algorithms within the compiler toolchain to improve timing, congestion, and resource utilization. Integrate placement capabilities with routing, synthesis, and timing teams, analyze placement quality and bottlenecks, and enhance performance, scalability, and QoR for large designs.
Top Skills: AsicC/C++Eda ToolsFpgaPythonQuartusTclVivado

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account