Disclosure certainty 3/5 debate potential 2/5

Feldman: Cerebras signed an OpenAI compute deal worth over $20 billion

Andrew Feldman · Cerebras CEO: Why GPUs Can't Do Fast Inference · Jul 23, 2026 · at 9:49

Andrew Feldman, CEO of Cerebras Systems, discloses the scale of Cerebras' contract with OpenAI.

0:00 / 0:06exact quote · 7.0s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“Remember, we did a huge deal. This is probably the largest deals in Silicon Valley history north of twenty billion dollars.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Andrew Feldman

Assertion Not checkable as stated
Feldman: Nvidia CUDA lost 70% of frontier AI model training market share
“I think two years ago every state of the art model was trained in a Cuda flow. And right now, Gemini is trained without Cuda. Anthropical is trained without Cuda. Open AI as strange as could. So in a one or two year period, they lost 70% share. Of training mod…”
Andrew Feldman Jul 23, 2026 ▶ 1:01:20 Cerebras CEO: Why GPUs Can't Do Fast Inference
Opinion
Feldman: AI is causing irreparable damage to the SaaS business model
“I think the business of dashboarding and the business, the AI doesn't the damages doing the SAS is, I think, unreparable. You could ask your AI, build me a tool like Salesforce. 30 seconds later, you have a working tool.”
Andrew Feldman Jul 23, 2026 ▶ 1:10:46 Cerebras CEO: Why GPUs Can't Do Fast Inference
Assertion Contradicted
Feldman: Cerebras sales were 10x higher than Groq's at acquisition
“And we were the fastest at it, and the largest, and, you know, our sales were more than 10 times the Grox, and they paid twenty billion dollars for the number two collector.”
Andrew Feldman Jul 23, 2026 ▶ 9:01 Cerebras CEO: Why GPUs Can't Do Fast Inference
Opinion
Feldman: Chinese open-source AI models trail GPT, Anthropic, and Gemini
“They are behind in chips. But their approach was at the next level is open source models where they're producing some extraordinary models. Not as good as GPT or Anthropic or Google's Gemini, but very good.”
Andrew Feldman Jul 23, 2026 ▶ 12:59 Cerebras CEO: Why GPUs Can't Do Fast Inference
Assertion Not checkable as stated
Feldman: Agentic AI workflows are driving CPU demand through the roof
“And so, as we do more and more AI work, and more and more agentic work, we're making more and more calls to CPUs, and therefore the demand for CPUs is through the roof.”
Andrew Feldman Jul 23, 2026 ▶ 25:16 Cerebras CEO: Why GPUs Can't Do Fast Inference
Insight
Feldman: AI inference is bottlenecked by data movement, causing GPU slowness
“In inference in AI, it's the exact opposite. You move a huge amount of data, all the weights, from memory to compute, and you need one calculation to generate the next word. And then you have to do it again. So all the time is dominated by the movement of data…”
Andrew Feldman Jul 23, 2026 ▶ 32:16 Cerebras CEO: Why GPUs Can't Do Fast Inference
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.