CPU

also referred to as: cpus

10 statements across 9 episodes · 2 bullish · 2 bearish · 9 people on the record · first statement Oct 18, 2024 by Drew Houston · across every show →

Everything said about CPU, oldest first

Oct 18, 2024 positive
Insight
Houston: Humans should act as CPUs orchestrating AI systems like GPUs
“Right now we have, like, the human CPU doing a lot of, you know, silicon CPU tasks, and so you really have to, like, redesign the work thoughtfully such that, you know, probably not that different from how it's evolved in computer architecture, where the CPU i…”
Drew Houston Oct 18, 2024 ▶ 47:22 Building the Silicon Brain - Drew Houston of Dropbox
Oct 19, 2024 neutral
Assertion Supported
Hu: GPU setups showed virtually no agent performance gain over CPU-only
“They compared a CPU only setup to a GPU setup to a multi GPU setup, and it kind of made no difference really.”
Jesse Hu Oct 19, 2024 ▶ 49:40 [Paper Club] SWE-Bench [OpenAI Verified/Multimodal] + MLE-Bench with Jesse Hu
Apr 11, 2025
Insight
Conrad: Incremental GPUs always drive model performance and revenue, unlike CPUs
“Gusto isn't going to make like, you know, five percent more money. They're going to make zero, like literally zero money from every incremental GPU or CPU after a certain point. This is not the case for anyone who is training models. And it's not the case for …”
Evan Conrad Apr 11, 2025 ▶ 3:08 SF Compute: Commoditizing Compute
Jun 13, 2025
Assertion Supported
CPU latency in KV cache management bottlenecks GPU utilization
“Like your eviction policy runs on a CPU. Like that radix hashing algorithm and block hashing and all that stuff happens like primarily CPU. That's really important for performance because if you have latency in these steps, like you're not keeping your GPU uti…”
Chris Lattner Jun 13, 2025 ▶ 1:00:09 The Shape of Compute (Chris Lattner of Modular)
Jul 28, 2025 neutral
Insight
Mohan: GPU container sharing limitations leave hardware heavily idle
“For most people, one of the things about CPUs that's really nice is with containers, right? You can end up having a single node and you can place many containers on them and all the containers will slowly start eating the compute. It's not really the same with…”
Varun Mohan Jul 28, 2025 ▶ 5:13 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Feb 24, 2026 bearish
Prediction Not checkable as stated
O'Laughlin: Tech industry may face a CPU shortage from AI coding and RL
“You feel like we might actually be seeing a CPU shortage partially because of this refresh cycle, but partially also because like I legitimately believe the cloud code Cloud code is increasing software creation and then on top of that, there is real demand fro…”
Doug O'Laughlin Feb 24, 2026 ▶ 1:56:07 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
May 21, 2026 neutral
Insight
Burazin: CPU environments must spin up instantly to prevent costly GPU idle time
“The reason why a lot of people come to us is because GPUs are more expensive than CPUs, right? So you want your GPU running at what? A hundred percent the entire time. And so when you're running runs on CPUs, when the CPU cycle is like down and spinning up the…”
Ivan Burazin May 21, 2026 ▶ 27:41 AI Agents Need Computers: 74% MoM Growth, 850K/Day Runs, & New Agent Cloud — Ivan Burazin, Daytona
May 21, 2026 bearish
Prediction Not checkable as stated
Burazin: CPUs Will Become the Next Critical Bottleneck for AI Agents
“You will get to the point, and Dylan Patel was at the conference talking about, from Semi-Analysis that talks usually about GPUs, was also talking about how CPUs will now be a bottleneck because it will be the constraint. You won't be able to grow, or we won't…”
Ivan Burazin May 21, 2026 ▶ 47:35 AI Agents Need Computers: 74% MoM Growth, 850K/Day Runs, & New Agent Cloud — Ivan Burazin, Daytona
May 24, 2026 positive
Assertion Supported
Gemma 4 E2B loads only 2B of 5B parameters into GPU
“So the GEMA for model is a E to B. That means that it effectively has two billion parameters loaded into the GPU. It actually has almost five billion parameters, but those three billion parameters can be in the CPU, they can be in the disk, which means that yo…”
Omar Sanseviero May 24, 2026 ▶ 0:52 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
Aug 26, 2026 neutral
Assertion Supported
Lean faces CPU-bound scalability limits for verifying large neural networks
“So lean still has a lot of shortcomings there. It's CPU based and you know, it's not, Like, getting that onto the GPU has a lot of nuances there. So, you know, a lot of work needs to be done. So what we've started with is a framework, you know, making that mor…”
Anima Anandkumar Aug 26, 2026 ▶ 10:31 🔬 Why Transformers Hit a Wall the Moment Physics Shows Up — Anima Anandkumar, Caltech
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.