GPU

also referred to as: gpus

44 statements across 24 episodes · 21 bullish · 11 bearish · 21 people on the record · first statement May 10, 2017 by Josh Wolfe · across every show →

Everything said about GPU, oldest first

May 10, 2017 bullish
Prediction Held up
Wolfe: The future of AI and machine learning is in GPUs
“The future of AI and machine learning is going to be in GPUs.”
Josh Wolfe May 10, 2017 ▶ 14:17 20VC: The 3 Forms Of Edge A Founder Can Have, Lessons From Being on A Board With Bill Gates & Why There Are A Lot Of Tourist VCs Who Are Going To Lose A Lot Of Money, with Josh Wolfe, Co-Founder @ Lux Capital
Apr 28, 2023 negative
Opinion
Guo: Fewer than 10 AI startups should train foundation models
“The thing that is really expensive I mean, there are many things that can be expensive, but one of the things that is really expensive is I want to train a model from scratch that is very large and it's going to take me you know, low tens of people, probably 2…”
Sarah Guo Apr 28, 2023 ▶ 12:38 Sarah Guo: On Her New $101M Fund; How AI Impacts Inequality; AI Startups vs Incumbents | E1007 · 20VC with Harry Stebbings
May 17, 2023 bullish
Assertion Not checkable as stated
Mostaque: Stability AI sees zero failure rate training on Google TPUs
“TPUs are the most scalable architecture, like we have zero failure rate with our TPU language model training, whereas with GPUs it's like there's an ECC error”
Emad Mostaque May 17, 2023 ▶ 10:53 Emad Mostaque: These 5 Companies Will Win the AI War; Why We Need National Data Sets | E1015 · 20VC with Harry Stebbings
Aug 18, 2023
Insight
Socher: GPU incompatibility prevents exploration of non-transformer AI models
“Transformers are mostly amazing because they can be trained on GPUs. There are lots of other models that we're currently not exploring because they're not trainable on GPUs.”
Richard Socher Aug 18, 2023 ▶ 25:31 Richard Socher: The 3 Biggest Barriers to Building AGI; AI Startups vs Incumbents | E1050 · 20VC with Harry Stebbings
Nov 15, 2023 bearish
Prediction Held up
Traynor: Nvidia won't be the sole GPU maker in five years
“I don't think NVIDIA in five years time will be sitting here proud as the only people who can do a GPU.”
Des Traynor Nov 15, 2023 ▶ 38:33 Des Traynor: How to Survive and Thrive in a World of OpenAI | E1082 · 20VC with Harry Stebbings
Jun 12, 2024 bullish
Prediction Not checkable as stated
Wang: AI competition will shift from GPU access to proprietary data rights
“Right now if you know, in San Francisco, the big, you know, researchers and the big CEOs brag about how many GPUs they have. You know, they're the sort of the biggest indicator of how serious they are about AI is, is how many GPUs they have. But I think in the…”
Alex Wang Jun 12, 2024 ▶ 21:25 Alex Wang: Why Data Not Compute is the Bottleneck to Foundation Model Performance | E1164 · 20VC with Harry Stebbings
Aug 5, 2024 neutral
Assertion Supported
Cahn: GPUs Account for Half of Total AI Data Center Costs
“About half the cost of the data center is GPUs, half is everything else.”
David Cahn Aug 5, 2024 ▶ 41:51 David Cahn: Why Servers, Steel and Power Are the Pillars Powering the Future of AI | E1186 · 20VC with Harry Stebbings
Aug 5, 2024
Prediction Open · timeframe Aug 2029
Cahn: No AI Lab Will Train Frontier Models in the Same Data Center Twice
“No one's ever going to train a frontier model on the same data center twice, because by the time you've trained it, the GPUs will be outdated, and the data center will be too small.”
David Cahn Aug 5, 2024 ▶ 14:57 David Cahn: Why Servers, Steel and Power Are the Pillars Powering the Future of AI | E1186 · 20VC with Harry Stebbings
Oct 7, 2024
Disclosure
Kant: Poolside brought 10,000 GPUs online in summer 2024
“The 10,000 GPUs that we've now brought online this summer, you know, that came from this capital, allow us to make incredible advancements in model capabilities”
Eiso Kant Oct 7, 2024 ▶ 28:20 Eiso Kant, CTO @Poolside: Raising $600M To Compete in the Race for AGI | E1211 · 20VC with Harry Stebbings
Oct 7, 2024
Assertion Not checkable as stated
Kant: Interconnecting over 32,000 GPUs is currently extremely challenging
“Today, interconnecting more than 32,000 GPUs is extremely challenging.”
Eiso Kant Oct 7, 2024 ▶ 29:07 Eiso Kant, CTO @Poolside: Raising $600M To Compete in the Race for AGI | E1211 · 20VC with Harry Stebbings
Jan 29, 2025 positive
Prediction Held up
Ross: AI companies will use massive GPU clusters to generate synthetic data
“What you're going to see now is now that everyone has seen this deep seek architecture, they're going to go great. I have hundreds of thousands of GPUs. I'm now going to use a lot of them to create a lot of synthetic data. And then I'm going to train the bejes…”
Jonathan Ross Jan 29, 2025 ▶ 47:14 Jonathan Ross: DeepSeek Special - How Should OpenAI and the US Government Respond | E1253 · 20VC with Harry Stebbings
Feb 17, 2025 positive
Assertion Not checkable as stated
Ross: Groq LPUs Cost 5x Less Than Latest GPUs for Inference
“More than five x lower. Just the memory alone in the latest GPUs costs more than our fully loaded CapEx per chip deployed.”
Jonathan Ross Feb 17, 2025 ▶ 29:06 Jonathan Ross, Founder & CEO @ Groq: NVIDIA vs Groq - The Future of Training vs Inference | E1260 · 20VC with Harry Stebbings
Feb 17, 2025 negative
Assertion Supported
Ross: NVIDIA used misleading benchmark curves to claim 30x GPU speedup
“Like, if you look at the last GTC, there was an announcement that the latest GPUs were 30 x faster than the previous generation. And when you look at how it was done, there was this curve that looked kind of like this, and then it Basically ended here, and the…”
Jonathan Ross Feb 17, 2025 ▶ 24:48 Jonathan Ross, Founder & CEO @ Groq: NVIDIA vs Groq - The Future of Training vs Inference | E1260 · 20VC with Harry Stebbings
Feb 17, 2025 positive
Assertion Not checkable as stated
Ross: GPU Operational Cost Alone Equals Groq Total CapEx plus OpEx
“So we use about a third of the energy per token. About over a three-year period, one-third of our cost is the OpEx, which is mostly energy and data center rent, and two-thirds is the CapEx, which means that since we're one-third of the energy, the cost to run …”
Jonathan Ross Feb 17, 2025 ▶ 29:18 Jonathan Ross, Founder & CEO @ Groq: NVIDIA vs Groq - The Future of Training vs Inference | E1260 · 20VC with Harry Stebbings
Feb 17, 2025 neutral
Assertion Supported
Ross: Southeast Asian GPU deployments secretly serve Chinese firms
“One of the concerns right now is about Malaysia or Singapore, that region over there being a place where people are deploying GPUs with the wink, wink, like we're not going to rent it to China, right? But that's a belief that a lot of people are doing that. Ot…”
Jonathan Ross Feb 17, 2025 ▶ 1:01:44 Jonathan Ross, Founder & CEO @ Groq: NVIDIA vs Groq - The Future of Training vs Inference | E1260 · 20VC with Harry Stebbings
Feb 24, 2025 positive
Assertion Supported
Morin: Groq and Cerebras beat GPUs via on-chip data storage
“Actually, that's why Grok achieves, ah, not Grok, but Grok, Cerebras, and all these folks, they achieve very high performance single stream is because the data is right in the chip that doesn't have to get it from memory, which is slow, which GPU has to do.”
Steeve Morin Feb 24, 2025 ▶ 12:49 Steeve Morin: Why Google Will Win the AI Arms Race & OpenAI Will Not | E1262 · 20VC with Harry Stebbings
Feb 24, 2025 bearish
Prediction Open · timeframe Dec 2025
Morin: AI chip oversupply will lead to GPUs selling at 30% value
“I very much worry there will be an oversupply of these chips. The problem is, is that, you know, remember, the chips are the collateral. So, you know, somewhere, you know, in the US or whatever, there's going to be a data center with like a thousand GPUs that …”
Steeve Morin Feb 24, 2025 ▶ 24:31 Steeve Morin: Why Google Will Win the AI Arms Race & OpenAI Will Not | E1262 · 20VC with Harry Stebbings
Feb 24, 2025 negative
Assertion Contradicted
Morin: Doubling GPUs in AI inference yields only 10% performance gain
“If you go from one GPU to two, you don't get twice the performance. Maybe you get 10% better performance. Yeah, that's the dirty secret nobody talks about. I'm talking inference, right? So, so you go from, let's say, a hundred to a 110 by doubling the amount o…”
Steeve Morin Feb 24, 2025 ▶ 45:22 Steeve Morin: Why Google Will Win the AI Arms Race & OpenAI Will Not | E1262 · 20VC with Harry Stebbings
Feb 24, 2025 bearish
Assertion Contradicted
Morin: GPUs cannot deliver latent space AI reasoning at scale
“Fundamentally, GPUs cannot deliver, deliver this, plain and simple at scale.”
Steeve Morin Feb 24, 2025 ▶ 31:24 Steeve Morin: Why Google Will Win the AI Arms Race & OpenAI Will Not | E1262 · 20VC with Harry Stebbings
Feb 24, 2025 negative
Insight
Morin: GPUs are a clever workaround, not natively built for AI
“GPUs are, you know, are a good trick for AI, but they're not built for AI.”
Steeve Morin Feb 24, 2025 ▶ 11:13 Steeve Morin: Why Google Will Win the AI Arms Race & OpenAI Will Not | E1262 · 20VC with Harry Stebbings
Mar 24, 2025 negative
Opinion
Feldman: GPU off-chip memory architecture can be beaten in inference
“The fundamental architecture of the GPU with off-chip memory is not great for inference. Now, they will continue to do well in inference, but it can be beaten, and I think they know it.”
Andrew Feldman Mar 24, 2025 ▶ 0:17 Andrew Feldman, Cerebras Co-Founder and CEO: The AI Chip Wars & The Plan to Break Nvidia's Dominance · 20VC with Harry Stebbings
Mar 24, 2025 negative
Assertion Partly supported
Feldman: GPUs operate at only 5% to 7% utilization during inference
“In a GPU, most of the time it's doing inference, it's five or seven percent utilized. That means it's 95 or 93% wasted.”
Andrew Feldman Mar 24, 2025 ▶ 25:34 Andrew Feldman, Cerebras Co-Founder and CEO: The AI Chip Wars & The Plan to Break Nvidia's Dominance · 20VC with Harry Stebbings
Mar 24, 2025
Insight
Feldman: Data movement, not computation, is the core bottleneck in AI chips
“The hard part with AI work is results and intermediate results have to be moved a lot. And therein is the most complicated part. They have to be moved to memory and from memory. And they have to be broken up and moved among GPUs.”
Andrew Feldman Mar 24, 2025 ▶ 3:36 Andrew Feldman, Cerebras Co-Founder and CEO: The AI Chip Wars & The Plan to Break Nvidia's Dominance · 20VC with Harry Stebbings
Apr 10, 2025 positive
Assertion Supported
Boland: NVIDIA Is Re-Architecting GPUs to Optimize for Inference
“They are making a bunch of architectural changes to GPUs to make them better and better inference.”
Stan Boland Apr 10, 2025 ▶ 1:14:23 Tom Hulme & Stan Boland: Lessons from Jensen Huang & How to Fix the UK Tech Ecosystem · 20VC with Harry Stebbings
Jun 2, 2025 positive
Assertion Not checkable as stated
Windsurf's custom AI model processes hundreds of billions of tokens daily
“And now it processes hundreds of billions of tokens of code a day, that single model. And it's running on our own GPUs.”
Varun Mohan Jun 2, 2025 ▶ 29:50 Windsurf CEO & Co-Founder, Varun Mohan: AI's Biggest Acquisition to Date! · 20VC with Harry Stebbings
Aug 7, 2025 bullish
Insight
Rory O'Driscoll: GPUs and foundational models dominate the AI tech stack over applications
“I think that the two people who have, the two categories in the stack that didn't meaningfully exist at scale in SAS and cloud land that exists now are the GPUs, which is all Jensen and the models. Neither of those kind of categories even existed. And obviousl…”
Rory O'Driscoll Aug 7, 2025 ▶ 1:02:57 Figma's 250% Pop - The Greatest IPO Mispricing Ever? Meta & Microsoft Blowout Quarters: Broken Down · 20VC with Harry Stebbings
Sep 29, 2025 neutral
Assertion Not publicly verifiable
Ross: Google ran three parallel chip projects, but only TPU succeeded
“So, people look at the TPU as a big success, and what they don't realize is that there were about three chip efforts at Google at the same time and only one of them ended up outperforming GPUs.”
Jonathan Ross Sep 29, 2025 ▶ 7:13 Groq Founder, Jonathan Ross: OpenAI & Anthropic Will Build Their Own Chips & Will NVIDIA Hit $10TRN · 20VC with Harry Stebbings
Sep 29, 2025 positive
Assertion Not publicly verifiable
Ross: Microsoft withheld Azure GPUs because internal usage made more money
“Microsoft in one quarter deployed a bunch of GPUs, and then announced that they weren't going to make them available in Azure because they made more money using them themselves than renting them out.”
Jonathan Ross Sep 29, 2025 ▶ 2:01 Groq Founder, Jonathan Ross: OpenAI & Anthropic Will Build Their Own Chips & Will NVIDIA Hit $10TRN · 20VC with Harry Stebbings
Sep 29, 2025 positive
Assertion Not checkable as stated
Ross: Groq LPUs Ship in 6 Months Versus 2 Years for GPUs
“You have to write a check two years in advance to get GPUs. For us, you write us a check for a million LPUs, and the first of those LPUs starts showing up six months later.”
Jonathan Ross Sep 29, 2025 ▶ 24:14 Groq Founder, Jonathan Ross: OpenAI & Anthropic Will Build Their Own Chips & Will NVIDIA Hit $10TRN · 20VC with Harry Stebbings
Nov 27, 2025 bearish
Opinion
Jason Lemkin: Seamless switching between TPUs and GPUs undercuts Nvidia's moat
“The idea that NVIDIA is unstoppable because of the software and hardware connection, because we have to have GPUs. I know there's a lot of truth to that, but literally as an end user, I went right back and forth to them today. No, no issue. TPUs, GPUs, LLMs, a…”
Jason Lemkin Nov 27, 2025 ▶ 5:46 Anthropic Raises $30B from Microsoft & NVIDIA & NVIDIA’s Core Business Faces TPU Threat · 20VC with Harry Stebbings
May 26, 2026 positive
Disclosure
Feldman claims HBM memory shortages limit traditional GPUs but not Cerebras
“That is a limitation for all GPUs, but not us. We don't use it.”
Andrew Feldman May 26, 2026 ▶ 8:29 Cerebras CEO on the Future of Data Centres, Token Costs & Memory | Should US Companies Sell to China · 20VC with Harry Stebbings
Jul 20, 2026 positive
Assertion Not checkable as stated
Lin Qiao: Fireworks runs distributed RL across 5-6 global data center regions
“We've designed a fully distributed system. We run across five, six data center regions globally, and tap into scattered GPUs, and they are able to run massive jobs, our jobs.”
Lin Qiao Jul 20, 2026 ▶ 36:27 The Open-Source AI Reality | How Token Costs Will Fall 10X & Usage Will Explode 100X | Lin Qiao
Jul 31, 2026 bullish
Insight
Park: LLMs are the CPU of intelligence; simulation is the GPU
“What I see today that's prominent in AI space is what I consider to be the CPU of intelligence unit. You have this one language model that's really large, that's very smart, that can do very complex reasoning tasks. That's like CPU. What I see coming and what …”
Joon Sung Park Jul 31, 2026 ▶ 54:14 The AI Company Simulating the Entire Economy | Simile Co-founder & CEO, Joon Sung Park
Aug 2, 2026 bullish
Prediction Not checkable as stated
Angelopoulos: Value in AI Biology Will Accrue to Data Layer
“That's exactly one of the areas where the data layer, where you can clearly see that the data layer is where value is going to accrue. Because the GPUs Are the same GPUs in both cases. The problem is that, that data infrastructure, the flywheel, the data colle…”
Anastasios Angelopoulos Aug 2, 2026 ▶ 1:08:28 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
Aug 2, 2026 positive
Insight
Angelopoulos: Open-source AI growth reduces Nvidia's revenue concentration
“Of course, Jensen is in some sense self-serving with this letter, because the more open source models are developed, the more companies are going to be training on GPUs. They're going to be fine tuning on their own data. And it's just more and more spend. It d…”
Anastasios Angelopoulos Aug 2, 2026 ▶ 21:17 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
Aug 2, 2026 neutral
Assertion Not checkable as stated
Angelopoulos: Frontier AI labs spend 10% to 20% of GPU compute budgets on data
“Companies are spending on it, usually within Frontier Labs, at about 10 to 20% about the amount that they're spending on GPUs.”
Anastasios Angelopoulos Aug 2, 2026 ▶ 45:31 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
Aug 22, 2026 bullish
Insight
Murdock: ASIC Chips Are Ideal for Customization While GPUs Are Too Expensive
“I, look, Asics chips are really ideal if you're thinking about model customization. If you're saying, look, we're at a new phase in, in, in this AI build out, or what we really want to do is, is, is do a lot of model specialization. You don't need a GPU for th…”
Jerry Murdock Aug 22, 2026 ▶ 27:30 The AI Bubble WILL Burst | Should we be fearful of Chinese Open-Source | Jerry Murdock
Sep 3, 2026 positive
Insight
O'Driscoll: Open-Source AI Benefits GPU Vendors by Compressing Software Margins
“Open source is good for compute salespeople. If you're selling GPUs, you want everyone else's margin to be lower, so yours can be higher.”
Rory O'Driscoll Sep 3, 2026 ▶ 8:51 NVIDIA Crushes Quarter | OpenAI Cuts Off Cursor | Instinct Hits $2.5B Valuation
Sep 5, 2026 bullish
Insight
Weitzman: Networked GPUs have intrinsic utility and store value globally
“But a GPU, it has intrinsic value. Like, you can actually use that asset for something that's really, really valuable. And it doesn't matter where that GPU is. It could be in Iceland. It's still useful to anybody all over the world, as long as it's networked. …”
Cliff Weitzman Sep 5, 2026 ▶ 13:02 How to Build Your Own Data Center & Why Every Startup Should Do It
Sep 5, 2026 positive
Assertion Not checkable as stated
Weitzman says GPU analysis helped identify and solve his father's prostate cancer
“It's already solved my dad's prostate cancer, because I figured out with a bunch of help from other people how to use GPUs to identify where in his body the lesion was.”
Cliff Weitzman Sep 5, 2026 ▶ 1:03:25 How to Build Your Own Data Center & Why Every Startup Should Do It
Sep 5, 2026
Insight
Weitzman: The biggest cost of delayed GPUs is unutilized data center rent
“The most expensive part of a delivery of a GPU is if it's late, I'm still paying rent for that data center space.”
Cliff Weitzman Sep 5, 2026 ▶ 15:21 How to Build Your Own Data Center & Why Every Startup Should Do It
Sep 5, 2026
Disclosure
Weitzman: Speechify will pay $100k extra monthly for faster GPU delivery
“We're very willing to pay a hundred K per month extra to get them earlier.”
Cliff Weitzman Sep 5, 2026 ▶ 14:37 How to Build Your Own Data Center & Why Every Startup Should Do It
Sep 5, 2026
Disclosure
Weitzman: Speechify engineers concurrently run 5 to 18 autonomous coding agents
“Our engineers, really what I'm looking for is 10 really good decisions per day, which is very tiring, not like optimizing the random parts of the code. And each one has like, you know, five to 18 agents running at any point in time, doing long horizon tasks on…”
Cliff Weitzman Sep 5, 2026 ▶ 18:51 How to Build Your Own Data Center & Why Every Startup Should Do It
Sep 5, 2026 bearish
Opinion
Stebbings argues buying GPUs is a mistake due to marginal cost savings
“It is a mistake to price optimize and to spend the money to buy it versus to rent it because I get you on the optimization, but you're not saving 10 times more. It's .5 X more per year.”
Harry Stebbings Sep 5, 2026 ▶ 16:55 How to Build Your Own Data Center & Why Every Startup Should Do It
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.