Why does Cerebras's wafer-scale engine process AI workloads faster than standard GPUs?
Replied byAndrew Feldman
Co-founder & CEO at Cerebras Systems
Niche: Technology
Revenue: Approx. $72 Million/month
Location: Sunnyvale, California, United States
Started: 2016
AI inference is dominated by moving weights from memory to compute. By building a dinner-plate-sized chip stuffed with fast SRAM instead of HBM, we keep the weights directly on the silicon. This allows us to move data to compute two thousand five hundred times faster than a GPU.
0
From the Full Interview
This answer is part of a full interview with Andrew Feldman, Co-founder & CEO at Cerebras Systems.
Share this Answer
Found this insight valuable? Share it with your network to help others learn from Andrew Feldman's experience.
Cite This Answer
Use this answer in your research, article, or academic work
Related Answers
What was your worst historical investment experience, and what risk management lesson did it impart?
By Sir Paul Marshall
Finance
Not Publicly Disclosed/mo
How did you resolve your public dispute with Carl Icahn over your Herbalife short position?
By Bill Ackman
Finance
Not Publicly Disclosed/mo
What role did your early junior mining investments play in your career?
By Aaron Hoddinott
Finance
Not Publicly Disclosed /mo
How did you first connect with Rick Rule, and what role has he played in your career?
By Collin Kettell
Finance
Not Publicly Disclosed/mo
If your colleagues described you in three words, what would they say?
By Gaurav Jalan
Finance
Estimated $4M-$8M USD/mo
How did your 2012 prediction of frighteningly ambitious ideas play out with companies like OpenAI?
By Paul Graham
Finance
Not Publicly Disclosed/mo
What is your overarching outlook on macroeconomic cycles and the artificial intelligence revolution?
By Sir Paul Marshall
Finance
Not Publicly Disclosed/mo