GPU

also referred to as: gpus

28 statements across 22 episodes · 13 bullish · 7 bearish · 20 people on the record · first statement Apr 25, 2023 by Jensen Huang · across every show →

Everything said about GPU, oldest first

Apr 25, 2023 positive
Assertion Supported
Huang: Andrew Ng contacted NVIDIA around 2012 to train neural nets on GPUs
“Around 20 12, I guess, and it was because simultaneously Andrew Eng reached out to Bill Daly, our chief scientist to work on a way to get the neural network model that they were working on onto GPU so that they could, instead of using thousands of CPU servers,…”
Jensen Huang Apr 25, 2023 ▶ 11:40 No Priors Ep. 13 | With Jensen Huang, Founder & CEO of NVIDIA
Apr 25, 2023 positive
Insight
Self-Attention Brought GPU Parallelism to Sequence Modeling
“The insight here was, Hey, you can use the same attention thing to like, look back at the past of the sequence that you're trying to produce. And you know, the beauty is that the, that It runs great on on GPUs and CPUs, and it's kind of parallel to, like, how …”
Noam Shazeer Apr 25, 2023 ▶ 6:21 No Priors Ep. 12 | With Noam Shazeer
Aug 24, 2023 bearish
Opinion
Uszkoreit: GPUs are not at the sweet spot for large-scale deep learning
“I don't think GPUs are at the sweet spot when it comes to large-scale deep learning with respect to exactly those trade-offs, and so it may very well be that if we actually try these combinations, we might actually even quickly find something that's better.”
Jakob Uszkoreit Aug 24, 2023 ▶ 5:45 No Priors Ep. 29 | With Inceptive CEO Jakob Uszkoreit
Sep 7, 2023 bullish
Assertion Supported
Feldman: Cerebras can run a trillion-parameter model on a single system
“And that was an idea that came from supercomputing that we knew really well, that we could organize this so you could run an arbitrarily large, now a trillion parameter network on a single system.”
Andrew Feldman Sep 7, 2023 ▶ 9:06 No Priors Ep. 31 | With Cerebras CEO Andrew Feldman
Sep 7, 2023 positive
Assertion Supported
Feldman: Cerebras hardware runs strictly data parallel without complex distributed engineering
“And if you have to spend months doing distributed compute, doing tensor model parallel distributed compute, I mean, if you look at the back of some of these papers, they're crediting 20 or 30 people, sometimes more, who helped on the distributed compute. And i…”
Andrew Feldman Sep 7, 2023 ▶ 5:31 No Priors Ep. 31 | With Cerebras CEO Andrew Feldman
Sep 14, 2023 bearish
Opinion
Polosukhin: Decentralized AI training is unrealistic due to GPU bandwidth requirements
“And the reality right now that the requirements on bandwidth, right? Like people who are training these models right now, they have like a, you know, 800 gigabit connect right between the GPUs, right? So Maybe you have a hundred megabits on between this, usual…”
Illia Polosukhin Sep 14, 2023 ▶ 22:55 No Priors Ep. 32 | With NEAR’s Illia Polosukhin
Sep 14, 2023 negative
Assertion Supported
Polosukhin: Repurposed crypto mining GPUs cannot meet frontier AI training needs
“The challenges, the GPUs there are like, not the ones that AI folks want to use, right? Like kind of all the AI is really zeroed in on like, how do we get a 100 or H 100 and the GPUs that like folks used for Ethereum mining and like similar is like older ones …”
Illia Polosukhin Sep 14, 2023 ▶ 21:55 No Priors Ep. 32 | With NEAR’s Illia Polosukhin
Oct 19, 2023 positive
Assertion Not checkable as stated
Gil: Google TPUs were dramatically more performant than GPUs for years
“And it obviously was dramatically more performant than GPU for a long time.”
Elad Gil Oct 19, 2023 ▶ 23:22 No Priors Ep. 37 | With Kawal Gandhi
Feb 29, 2024 neutral
Insight
Papermaster: Computing requires combined scalar CPU and parallel GPU architectures
“And to me, it was clear that the industry needed That powerful combination of the serial, the scalar competing of these traditional CPU workloads and the massive parallelization that you get from a GPU.”
Mark Papermaster Feb 29, 2024 ▶ 7:35 No Priors Ep. 53 | With AMD CTO Mark Papermaster
Apr 4, 2024 neutral
Disclosure
Adcock: Figure designs and builds everything in-house except battery cells, GPUs, CPUs
“Everything maybe besides like the battery cell and the GPU and CPU at this point, it feels like we've done what we're doing.”
Brett Adcock Apr 4, 2024 ▶ 21:20 No Priors Ep. 58 | The argument for humanoid robots with Brett Adcock from Figure
May 9, 2024
Insight
Sarah Guo notes training frontier models requires co-locating GPUs for data transfer
“Today to train these large models, you need all of the GPUs co-located because there is enough data transfer between different chips, right? Between your nodes. And there's a physical constraint on that in that you need to get that much power and to a data, da…”
Sarah Guo May 9, 2024 ▶ 23:38 No Priors Ep. 63 | With Sarah Guo and Elad Gil
Aug 22, 2024 negative
Assertion Not checkable as stated
Davis: AI cloud customers are forced into unwanted 3-year GPU contracts
“In AI cloud today, you're kind of forced to get really long-term reservations often three years for a fixed amount of capacity. [944] Jared Quincy Davis: No one really wants 64 GPUs for three years or a thousand GPUs for three years. [949] Jared Quincy Davis: …”
Jared Quincy Davis Aug 22, 2024 ▶ 15:37 No Priors Ep. 77 | With Foundry CEO and Founder Jared Quincy Davis
Aug 22, 2024 neutral
Assertion Not checkable as stated
Davis: AI teams hold 10% to 20% of GPUs idle as healing buffer
“And so one of the consequences of that is that it's very common now to hold aside 10 to 20% minimum of the GPUs that a team has as buffer, as healing buffer, in case of a failure so you can slot something else in to keep the training workload running, right?”
Jared Quincy Davis Aug 22, 2024 ▶ 5:01 No Priors Ep. 77 | With Foundry CEO and Founder Jared Quincy Davis
Nov 7, 2024 positive
Assertion Supported
Huang: xAI's 100,000 GPU cluster is the largest single unit built
“Decide to build this 100,000 GPU super cluster, which is, you know, the largest of its kind in, in one unit.”
Jensen Huang Nov 7, 2024 ▶ 15:24 No Priors Ep. 89 | With NVIDIA CEO Jensen Huang
Dec 12, 2024 neutral
Disclosure
Ejeckam: Akash is entering market as a server maker, not chipmaker
“We are buying chips. We're not making GPUs. We're taking the hottest chips in the world and we're cooling them down so that you can, we can open the envelope of performance for the system architect. Okay. And so we fit in, we're coming into the world as a serv…”
Felix Ejeckam Dec 12, 2024 ▶ 20:50 No Priors Ep. 93 | With Akash Systems' Felix Ejeckam and Ty Mitchell
Jan 9, 2025 negative
Opinion
Bernhardsson: Long-Term GPU Contracts Are the Wrong Model for Startups
“Means that cloud, you know, a lot of the cloud capacity is like, you know, the only way to get it is to sign long term commitments, which I think for a lot of startups is really not the right model for how things should be.”
Erik Bernhardsson Jan 9, 2025 ▶ 4:47 No Priors Ep. 96 | With Modal CEO and Founder Erik Bernhardsson
Mar 13, 2025 neutral
Insight
Dohmke: AI agent supply is infinite and constrained only by GPU capacity
“Human developers are expensive because there's limited supply agents will have infinite supply that, that will only be limited by the amount of compute capacity GPUs available in data centers.”
Thomas Dohmke Mar 13, 2025 ▶ 35:18 No Priors Ep 106 | With GitHub CEO Thomas Dohmke
Jul 31, 2025 negative
Assertion Partly supported
Krishnan: Biden Diffusion Rule Effectively Banned US GPU Exports to Allies
“Under the Biden era, there was something called the Biden diffusion rule, which basically was a 200 page document, which basically made it illegal for America to export GPUs. It was really hard for you know, if you're Jensen or if you're Lisa Su to really kind…”
Sriram Krishnan Jul 31, 2025 ▶ 23:47 No Priors Ep. 125 | With Senior White House Policy Advisor on AI Sriram Krishnan
Jul 31, 2025 bullish
Prediction Not checkable as stated
Krishnan: Exporting US GPUs Strongly Incentivizes Allies to Run American Models
“And one of the other things we're doing that is we get our GPUs over, we probably get them to run our models as opposed to models from, you know, another country, and we go from there.”
Sriram Krishnan Jul 31, 2025 ▶ 24:33 No Priors Ep. 125 | With Senior White House Policy Advisor on AI Sriram Krishnan
Aug 7, 2025 bullish
Prediction Not checkable as stated
Prince: GPUs will speed-run 30 years of CPU efficiency in ten years
“We're going to speed run the last 30 years of CPU efficiency gains. In, in the next, you know, five to 10 in, in GPUs.”
Matthew Prince Aug 7, 2025 ▶ 40:00 No Priors Ep. 126 | With Cloudfare CEO Matthew Prince
Aug 14, 2025 neutral
Assertion Supported
Patel: DeepSeek inference requires 160 GPUs and $10 million per replica
“Like the DeepSeq implementation of inference is like a 160 GPUs or something like that, like that's over ten million dollars of Hardware. And then that's just one replica.”
Dylan Patel Aug 14, 2025 ▶ 5:48 No Priors Ep. 127 | With SemiAnalysis Founder and CEO Dylan Patel
Aug 14, 2025
Assertion Supported
Patel: Meta is deploying GPUs in temporary tent structures
“Meta is literally building these like temporary, like tent structures to put GPUs in because building the building takes too long and it takes too much labor, right?”
Dylan Patel Aug 14, 2025 ▶ 31:13 No Priors Ep. 127 | With SemiAnalysis Founder and CEO Dylan Patel
Aug 14, 2025 bearish
Insight
Patel: Hyperscalers cannot replicate CPU and storage profit margins in GPUs
“Their ROIC is, like, extremely high on CPU and storage, and to assume that it can, like, translate over to GPUs is, is a bit of a fallacy, which is why a lot of these companies are moving in, right?”
Dylan Patel Aug 14, 2025 ▶ 15:54 No Priors Ep. 127 | With SemiAnalysis Founder and CEO Dylan Patel
Feb 26, 2026 positive
Assertion Not checkable as stated
Tiwari: AI Compute Debt Uses Customer Cash Flows, Not GPUs, as Primary Collateral
“And I think what got missed was the GPUs themselves were actually like the second, second or tertiary level of collateral in those instruments. The primary collateral was the contracted cash flows from investment grade counterparties.”
Neil Tiwari Feb 26, 2026 ▶ 9:32 Who's Actually Funding the AI Buildout?
Apr 3, 2026 bullish
Prediction Not checkable as stated
Fedus: Recursive self-improvement is happening in software and entering AI research
“So, In that domain, I think it's happening now-ish. And I think we'll see the same thing too for AI research. That's a slower outer loop because now the experiment isn't just checking some unit tests passing, but it's checking what was the scaling property? Di…”
Liam Fedus Apr 3, 2026 ▶ 24:26 AI for Atoms: How Periodic Labs is Revolutionizing Materials Engineering with Co-Founder Liam Fedus
May 21, 2026 bullish
Assertion Partly supported
Feldman: Cerebras is 15 to 20 times faster than GPUs at inference
“And right now we're the fastest at inference, not by a little bit, but by a lot. 1518, 20 X faster than GPUs.”
Andrew Feldman May 21, 2026 ▶ 1:37 The Story Behind Cerebras’ $63 Billion IPO with Founder and CEO Andrew Feldman
Jun 18, 2026 bullish
Assertion Not checkable as stated
Tan: AI inference is shifting CPU-to-GPU ratios toward 1:4 or 1:1
“Right now, the authentic AI and inference, CPU become, you know, highly in demand. And so, you know, versus one to eight in the training CPU to GPU, now I can see one to four, maybe one to one, and I'm delighted CPU become important.”
Lip-Bu Tan Jun 18, 2026 ▶ 5:31 Re-engineering the Semiconductor Supply Chain with Intel CEO Lip Bu Tan
Sep 3, 2026 bullish
Prediction Not checkable as stated
Haas: Edge AI Will Be a Sweet Spot for Arm Architecture
“And in fact, as you get to the smaller footprints, where more and more AI is going to take place, that's going to be a sweet spot for Arm, because the CPU's table stakes anyway, you have to have it to do all the things that are required in the edge device. But…”
Rene Haas Sep 3, 2026 ▶ 36:19 Redefining Chip Architecture with Arm CEO Rene Haas
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.