Assertion certainty 4/5 debate potential 3/5

Feldman: Leading AI labs paused video generation development due to compute costs

Andrew Feldman · Cerebras CEO: Why GPUs Can't Do Fast Inference · Jul 23, 2026 · at 53:29

Andrew Feldman is the co-founder and CEO of Cerebras Systems. He is explaining the compute bottlenecks delaying full video integration in multimodal frontier models.

0:00 / 0:20exact quote · 20.2s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“Obviously, what follows that Is video, because a video is just a collection of images. But that takes an enormous amount of compute right now. And that's one of the reasons it's been sort of set aside by the leading labs. So unbelievably computation intensive.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Andrew Feldman

Assertion Not checkable as stated
Feldman: Nvidia CUDA lost 70% of frontier AI model training market share
“I think two years ago every state of the art model was trained in a Cuda flow. And right now, Gemini is trained without Cuda. Anthropical is trained without Cuda. Open AI as strange as could. So in a one or two year period, they lost 70% share. Of training mod…”
Andrew Feldman Jul 23, 2026 ▶ 1:01:20 Cerebras CEO: Why GPUs Can't Do Fast Inference
Opinion
Feldman: AI is causing irreparable damage to the SaaS business model
“I think the business of dashboarding and the business, the AI doesn't the damages doing the SAS is, I think, unreparable. You could ask your AI, build me a tool like Salesforce. 30 seconds later, you have a working tool.”
Andrew Feldman Jul 23, 2026 ▶ 1:10:46 Cerebras CEO: Why GPUs Can't Do Fast Inference
Assertion Contradicted
Feldman: Cerebras sales were 10x higher than Groq's at acquisition
“And we were the fastest at it, and the largest, and, you know, our sales were more than 10 times the Grox, and they paid twenty billion dollars for the number two collector.”
Andrew Feldman Jul 23, 2026 ▶ 9:01 Cerebras CEO: Why GPUs Can't Do Fast Inference
Disclosure
Feldman: Cerebras signed an OpenAI compute deal worth over $20 billion
“Remember, we did a huge deal. This is probably the largest deals in Silicon Valley history north of twenty billion dollars.”
Andrew Feldman Jul 23, 2026 ▶ 9:49 Cerebras CEO: Why GPUs Can't Do Fast Inference
Opinion
Feldman: Chinese open-source AI models trail GPT, Anthropic, and Gemini
“They are behind in chips. But their approach was at the next level is open source models where they're producing some extraordinary models. Not as good as GPT or Anthropic or Google's Gemini, but very good.”
Andrew Feldman Jul 23, 2026 ▶ 12:59 Cerebras CEO: Why GPUs Can't Do Fast Inference
Assertion Not checkable as stated
Feldman: Agentic AI workflows are driving CPU demand through the roof
“And so, as we do more and more AI work, and more and more agentic work, we're making more and more calls to CPUs, and therefore the demand for CPUs is through the roof.”
Andrew Feldman Jul 23, 2026 ▶ 25:16 Cerebras CEO: Why GPUs Can't Do Fast Inference
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.