Why does Cerebras's wafer-scale engine process AI workloads faster than standard GPUs?

Andrew Feldman
Replied byAndrew Feldman

Co-founder & CEO at Cerebras Systems

Niche: Technology
Revenue: Approx. $72 Million/month
Location: Sunnyvale, California, United States
Started: 2016

AI inference is dominated by moving weights from memory to compute. By building a dinner-plate-sized chip stuffed with fast SRAM instead of HBM, we keep the weights directly on the silicon. This allows us to move data to compute two thousand five hundred times faster than a GPU.

0
From the Full Interview

This answer is part of a full interview with Andrew Feldman, Co-founder & CEO at Cerebras Systems.

Share this Answer

Found this insight valuable? Share it with your network to help others learn from Andrew Feldman's experience.

Cite This Answer

Use this answer in your research, article, or academic work