GPU

also referred to as: gpus

26 statements across 17 episodes · 11 bullish · 7 bearish · 17 people on the record · first statement Apr 6, 2017 by Todd Mostak · across every show →

Everything said about GPU, oldest first

Apr 6, 2017 bullish
Prediction Not checkable as stated
Mostak in 2017: GPUs will dominate computing over the next decade
“We're at an inflection point where the growth in data is outpacing the growth in compute, and I believe, I strongly believe, I'm a little biased, but, ah, that GPUs are set to kind of become a dominant force in compute over the coming decade. Ah, they already …”
Todd Mostak Apr 6, 2017 ▶ 13:38 The Power of GPU Analytics // Todd Mostak, MapD (FirstMark's Data Driven)
Feb 1, 2024 negative
Assertion Not checkable as stated
Duderstadt: GPU scarcity in academia is driving talent flight to industry
“And even nowadays, like, you know, it's increasingly hard to get those GPUs in academic positions. And so it's making increasing amounts of sense for, I think the flight that we're seeing from academia to industry.”
Brandon Duderstadt Feb 1, 2024 ▶ 5:53 How Nomic AI Is Driving The Open Source Revolution
May 16, 2024 bullish
Opinion
Hyperscaler GPU CapEx heavily benefits early-stage AI startups and VCs
“And ideally somebody else is, it's not venture capital dollars that are being used to buy and manage those GPUs. So it's a huge benefit.”
Tomasz Tunguz May 16, 2024 ▶ 13:12 AI, Data and Blockchain: a VC perspective | Tomasz Tunguz, Founder of Theory Ventures
May 31, 2024 negative
Prediction Not checkable as stated
Scaling compute and GPUs alone won't yield autonomous enterprise AI
“I'm not an expert in the field, but I believe that this is gonna we need Another giant leap of innovation to get there. I'm not sure just throwing compute and increasing the number of GPUs without some fundamental discovery, it's gonna create this.”
Daniel Dines May 31, 2024 ▶ 55:03 From Tiny Romanian Startup to Global AI Automation Leader | Daniel Dines, CEO of UIPath
Oct 10, 2024 bearish
Insight
Socher: Nvidia will match AI hardware startups' 100x gains before startups reach scale
“By the time you get that out and it's actually scalable and it can really train it and all the software is ready for it. And I can now go on AWS and spawn up your new hardware, like, Nvidia will also have been a hundred X faster with their latest and greatest …”
Richard Socher Oct 10, 2024 ▶ 27:44 AGI, The Future of AI Agents And The Next Wave of Opportunities in AI | Richard Socher, CEO, You.com
Oct 31, 2024
Disclosure
Bernhardsson: Modal can typically scale customers to 1,000 GPUs within minutes
“And then if you one day need a thousand GPUs, we can get you, we can typically get you a thousand GPUs like pretty quickly, like talking minutes.”
Erik Bernhardsson Oct 31, 2024 ▶ 20:12 Can AI Infrastructure Work Like Magic? Erik Bernhardsson, CEO, Modal
Oct 31, 2024 positive
Disclosure
Bernhardsson: Modal containerizes Python code to scale execution to thousands of GPUs
“Modal makes it easy to build, scale and deploy applications in the data, AI, machine learning realm. So, so we basically, you can think of as like, we take, you write a little bit of Python code, and we take that code, we stick it in a container, we execute th…”
Erik Bernhardsson Oct 31, 2024 ▶ 1:51 Can AI Infrastructure Work Like Magic? Erik Bernhardsson, CEO, Modal
Nov 21, 2024
Disclosure
Writer trained its Palmyra X 004 model using just $700k in compute
“Just on GPUs, but yeah.”
May Habib Nov 21, 2024 ▶ 23:31 Building the Easy Button for Generative AI | May Habib, CEO, Writer
Oct 30, 2025
Insight
Custom AI chips become viable only after model architectures stabilize
“As soon as you get to a point where there's some convergence on an architecture that's looks like it's stable and is revenue generating and developers are coming to sort of work on it and confirm that it is like the thing, then you can flip towards doing a cus…”
Nathan Benaich Oct 30, 2025 ▶ 27:11 State of AI 2025 with Nathan Benaich: Power Deals, Reasoning Breakthroughs, Real Revenue
Nov 6, 2025 neutral
Insight
Kant: Linearly adding GPUs to train larger AI models causes exponential training delays
“It is not that I can linearly add more GPUs and train increasingly a more larger model. If I do so, the time it takes becomes exponentially longer.”
Eiso Kant Nov 6, 2025 ▶ 50:06 Intelligence Isn’t Enough: Why Energy & Compute Decide the AGI Race – Eiso Kant
Nov 26, 2025
Assertion Not checkable as stated
Kaiser: Pre-training consumes the most GPUs of any AI development stage
“Currently, pre-training just uses the most GPUs of all the parts, so it needs the most GPUs, right?”
Łukasz Kaiser Nov 26, 2025 ▶ 32:09 What’s Next for AI? OpenAI’s Łukasz Kaiser (Transformer Co-Author)
Feb 5, 2026
Assertion Not checkable as stated
OpenAI surpassed ByteDance as the largest GPU renter globally
“And so when you look at who rents the most GPUs in the world, it's three companies, right? So one of them is obviously OpenAI. Second one, actually they were bigger than OpenAI. They are bigger than OpenAI today, or no, they were bigger than OpenAI than OpenAI…”
Dylan Patel Feb 5, 2026 ▶ 25:04 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Feb 5, 2026 positive
Insight
Patel: AI workload scale enables 10x performance gains from specialized chips
“But now the workload is so large that there is room for specialization that will give you 10 X increases in certain domains, right?”
Dylan Patel Feb 5, 2026 ▶ 2:01 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Feb 12, 2026
Insight
Lacroix: Thousands-GPU AI training has a much smaller margin for error
“When you run inference on a few GPUs, or when you run small scale trainings on hundreds of GPUs, margin for error is a lot larger than when you run trainings on thousands of GPUs at the same time.”
Timothée LeCroix Feb 12, 2026 ▶ 4:18 Mistral AI vs. Silicon Valley: The Rise of Sovereign AI
May 14, 2026 bearish
Prediction Held up
Burazin: AI agent scale creates high probability of impending CPU shortages
“I don't know it goes to the extreme to where GPUs are because that is very, very, very extreme. But it is quite highly, high probability that there will be shortages of CPUs going forward.”
Ivan Borzin May 14, 2026 ▶ 1:02:36 The Agent Harness: Building Secure Sandboxes for Autonomous AI Workloads
May 14, 2026
Assertion Supported
Burazin: Firecracker microVMs cannot run sandboxes with GPUs
“Now it can't run a sandbox with a GPU. It just doesn't work so that you can't have a firecracker.”
Ivan Borzin May 14, 2026 ▶ 47:32 The Agent Harness: Building Secure Sandboxes for Autonomous AI Workloads
Jun 18, 2026 positive
Disclosure
Balaban: Lambda software allows provisioning up to 4,000 GPUs via web interface
“Lambda's designed a piece of software that allows us to give you anywhere from 16 up to You know, 4000 GPUs in a web interface.”
Stephen Balaban Jun 18, 2026 ▶ 4:56 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Jun 18, 2026 positive
Prediction Not checkable as stated
Balaban: Complex GPU financial securities may eventually emerge as compute matures
“I think that the, that market is starting to mature that, that, that may be an eventuality is having more complex securities that surround GPUs. But I think for right now, people are starting to realize that it's a great credit investment and that's what's cha…”
Stephen Balaban Jun 18, 2026 ▶ 43:30 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Jun 18, 2026 negative
Assertion Supported
Balaban: Most neocloud competitors cannot launch online clusters over 32 GPUs
“Most of the other NeoClouds either don't have the ability to launch a cluster from their website or max out, I'll say, 32 GPUs.”
Stephen Balaban Jun 18, 2026 ▶ 4:49 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Jun 18, 2026 bullish
Prediction Not checkable as stated
Balaban: Everyone in the US will eventually require at least one GPU
“I believe that in the future, everybody in the United States will need the computational power of one GPU or more to just do their daily work You know, enjoy life, whether it's getting access, whether it's getting entertained, whether it's being productive, wh…”
Stephen Balaban Jun 18, 2026 ▶ 1:11:18 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Jun 18, 2026 bullish
Assertion Not checkable as stated
Balaban: Claims that AI GPUs have a five-year lifespan are wrong
“The usable life is longer than the accounting depreciation schedule. And what really matters is the economic usable life. And so what we're starting to see is that like the people who are the naysayers, oh, this is going to be, you're going to throw these GPUs…”
Stephen Balaban Jun 18, 2026 ▶ 42:14 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Jul 2, 2026 positive
Assertion Not checkable as stated
NVIDIA DLSS Is About 10 Times More Efficient Than Traditional Rendering
“DLSS is our real-time AI for graphics, and it makes a small GPU run like a big GPU. It's about 10 times more efficient because rather than computing the color of every pixel for every frame, we use AI to infer the color.”
Bryan Catanzaro Jul 2, 2026 ▶ 18:40 Inside Nemotron & NVIDIA’s AI Lab | Bryan Catanzaro
Jul 23, 2026 negative
Assertion Not checkable as stated
Feldman: GPUs suffer from high failure rates and infant mortality
“The JPs have a huge failure rate, so I'm sure you guys have spoken about this. Infant mortality is enormous, and they fail all the time.”
Andrew Feldman Jul 23, 2026 ▶ 40:08 Cerebras CEO: Why GPUs Can't Do Fast Inference
Jul 23, 2026 bullish
Assertion Supported
Feldman: Cerebras moves weights to compute ~2,500x faster than standard GPUs
“And so the speed of moving waits to compute is about two and a half thousand times faster here than on a Wilben GP.”
Andrew Feldman Jul 23, 2026 ▶ 44:43 Cerebras CEO: Why GPUs Can't Do Fast Inference
Jul 23, 2026 bearish
Insight
Feldman: AI inference is bottlenecked by data movement, causing GPU slowness
“In inference in AI, it's the exact opposite. You move a huge amount of data, all the weights, from memory to compute, and you need one calculation to generate the next word. And then you have to do it again. So all the time is dominated by the movement of data…”
Andrew Feldman Jul 23, 2026 ▶ 32:16 Cerebras CEO: Why GPUs Can't Do Fast Inference
Jul 29, 2026 positive
Assertion Supported
Cerebras cloud achieves 10x inference speedup over fast GPUs on Gemma
“Say, if you run Gemma four, On your GP, you might get like a hundred tokens per second if you have a fast card. If you run it in their cloud, you get anywhere from 800 to 1500 tokens per second. So call it 10 X faster.”
Sanjit Biswas Jul 29, 2026 ▶ 44:58 The Biggest AI Deployment Nobody Talks About | Samsara CEO Sanjit Biswas
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.