Training

topic on 11 shows · 22 statements across 20 episodes

BG2 Pod Cheeky Pint Latent Space No Priors Catalyst the MAD Podcast the a16z Podcast Big Technology All-In TBPN 20VC

22 statements about Training, every show

Optimal inference parallelism cannot be mathematically calculated; it must be auto-tuned
“And with training, it's more of like a math, like you can run the math and see the flops and maximize it. With inference, it's more of like an auto-tuning, like GPU kernel auto-tuning... You shadow the same traffic, like real traffic, and you just see which co…”
Ali Taha Aug 3, 2026 ▶ 56:06 Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
MAD Insight
Katti: Modern AI model training consists heavily of inference workloads
“We don't like to make a distinction between Training and infants, because a lot of training is now infants. So when we train a new model, we are generating synthetic data, for example. That's inference. When we train a new model, we are doing post-train, and t…”
Sachin Katti Jul 16, 2026 ▶ 14:22 OpenAI’s Compute Chief: We Can’t Build Fast Enough | Sachin Katti
CHEEKY PINT Prediction Open · timeframe Feb 2029
Pope: MatX will sell AI inference chips first due to lower risk
“Our product is both training and inference, but I think the first sales will be an inference. That's mostly just a market effect where It's easier to buy, like, it's not as big of a risk to go to buy an inference cluster than as a training cluster.”
Reiner Pope Feb 26, 2026 ▶ 11:23 Reiner Pope of MatX on accelerating AI with transformer-optimized chips
TBPN Assertion Not checkable as stated
Bishop: China manufactures capable inference chips but cannot produce training chips
“China has enough, like they can actually make, looks like decent inference chips. They just can't make the chips they need for training, right?”
Bill Bishop Feb 10, 2026 ▶ 25:58 FULL INTERVIEW: Bill Bishop Thinks China’s Military is Still Deeply Corrupt
a16z Prediction Not checkable as stated
Altman: Society will deem AI training fair use but create IP licensing models
“So like, you'll see this continue to move, but forced guests from the position we're in today, I would say that society decides training is fair use, but There's a new model for generating content in the style of or with the IPF or something else.”
Sam Altman Oct 8, 2025 ▶ 29:18 Sam Altman on Sora, Energy, and Building an AI Empire
20VC Assertion Supported
Feldman: Vastly more people do AI inference than AI training
“To move people off GPUs in inference, and the number of people doing inference is vastly higher than the number of people doing training.”
Andrew Feldman Oct 6, 2025 ▶ 26:05 Cerebras CEO, Andrew Feldman on Why Raise $1BN and Delay the IPO & Why NVIDIA’s Worried About Growth · 20VC with Harry Stebbings
a16z Disclosure
Kupor: OPM employees granted two hours monthly for AI training
“Everybody, you know, you're entitled to two hours a month of training, using all these free resources. You can take time out of work, just clear it with your manager.”
Scott Kupor Oct 2, 2025 ▶ 35:23 The Person Who Runs HR For 2 Million Federal Workers
Guthrie: Training-only datacenters cannot serve inference without global networking
“If you are, for example, building one large data center that only does training and it's not connected to a wide area network around the world, that's close to the users, it's hard to use that same infrastructure For inferencing because you can't go faster tha…”
Scott Guthrie Oct 1, 2025 ▶ 30:00 Microsoft's Cloud & AI Head on the AI Buildout's Risks and ROI — With Scott Guthrie
20VC Insight
Ross: AI Training and Inference Form a Virtuous Hardware Demand Cycle
“The more inference you have, as mentioned before, the more you need to train the model to optimize for the inference. And the more training you have the more inference you want to deploy to optimize for the cost of that training, to amortize the cost of the tr…”
Jonathan Ross Sep 29, 2025 ▶ 44:28 Groq Founder, Jonathan Ross: OpenAI & Anthropic Will Build Their Own Chips & Will NVIDIA Hit $10TRN · 20VC with Harry Stebbings
NO PRIORS Assertion Supported
Kohli: AlphaEvolve has successfully made AI model training more computationally efficient
“What Alpha Evolve has been able to do is basically make training more efficient.”
Pushmeet Kohli Jun 26, 2025 ▶ 26:01 No Priors Ep. 120 | With Google DeepMind’s Pushmeet Kohli and Matej Balog
a16z Insight
Long: Bad security training signals that company security is unimportant
“When the training is kind of a joke at the company, I think it also sends a signal to all of the employees that the company's security posture is also not very important.”
Brian C. Long Feb 28, 2025 ▶ 14:29 How to spot an AI Deepfake
CATALYST Prediction Held up
Kimber: AI inference will ultimately draw far more power than training
“I think now what we're seeing is that inference in the aggregate is actually probably a much larger draw than the training in the long run.”
Sheldon Kimber Feb 13, 2025 ▶ 44:04 The case for colocating data centers and generation
20VC Insight
Monday Afternoon Is Optimal for Pipeline Generation Training
“Monday afternoons is a great time to do training, especially training that's related to PG skills.”
Carlos Delatorre Jan 24, 2025 ▶ 31:59 Carlos Delatorre, CRO @Harness: Why Every Sales Rep Should Do Pipeline Generation | E1251 · 20VC with Harry Stebbings
NO PRIORS Disclosure
Bernhardsson: Modal Is Expanding into Bursty Experimental AI Training
“Traditionally, most of modal has always been inference. Like that's been our main use case, but we're really interested also in training. So in particular, like probably focused more on these like shorter, like very bursty sort of experimental training runs, n…”
Erik Bernhardsson Jan 9, 2025 ▶ 6:58 No Priors Ep. 96 | With Modal CEO and Founder Erik Bernhardsson
AI inference matters more than training because it scales with global population
“Our prediction is for those kind of applications, the inference is much more important than training. Because inference scale is proportional to the upliminal world population. And training. Training scale is proportional to the number of researchers.”
Lin Qiao Nov 25, 2024 ▶ 8:59 Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI
BG2 Opinion
Huang: NVIDIA's moat in inference will be greater than in training
“And I'm sure I said it would be greater.”
Jensen Huang Oct 13, 2024 ▶ 16:09 Ep17. Welcome Jensen Huang | BG2 w/ Bill Gurley & Brad Gerstner · Bg2 Pod
ALL-IN Prediction Open · timeframe Apr 2029
Chamath predicts AI inference market will be 100x larger than training
“AI is really two markets, training and inference is going to be a hundred times bigger than training.”
Chamath Palihapitiya Apr 26, 2024 ▶ 20:51 Meta's scorched earth approach to AI, Tesla's future, TikTok bill, FTC bans noncompetes, wealth tax
ALL-IN Opinion
Chamath asserts Nvidia hardware is miscast for AI inference
“NVIDIA is really good at training and very miscast at inference.”
Chamath Palihapitiya Apr 26, 2024 ▶ 21:07 Meta's scorched earth approach to AI, Tesla's future, TikTok bill, FTC bans noncompetes, wealth tax
NO PRIORS Insight
Srivastava: Inter-rack networking matters less for AI inference than training
“Even the GPU clusters themselves, like, you know, the full training networking is a very, very important Piece to have networking on the racks themselves with inference and matters a little less because you're doing a little bit more on individual GPUs and les…”
Tuhin Srivastava Mar 21, 2024 ▶ 4:57 No Priors Ep 56 | With Baseten CEO and Co-Founder Tuhin Srivastava
NO PRIORS Insight
Srivastava: AI inference demands strict uptime, while training tolerates node terminations
“Resiliency and reliability matters a lot more. You know, downtime is unacceptable from an input perspective. Nodes get terminated all the time from a training perspective.”
Tuhin Srivastava Mar 21, 2024 ▶ 6:01 No Priors Ep 56 | With Baseten CEO and Co-Founder Tuhin Srivastava
NO PRIORS Insight
Polosukhin: Inference demands vastly more aggregate compute than AI model training
“I think an inference is really interesting because we do need so much more compute for inference than we need for training, right? Like it's a very interesting like economy of scale. You train once, like Lama trained once and then everybody runs it everywhere.”
Illia Polosukhin Sep 14, 2023 ▶ 23:23 No Priors Ep. 32 | With NEAR’s Illia Polosukhin
a16z Insight
Superior genetics beat superior training in elite sports
“I think it's probably easier to be, A world-class athlete with great genetics and bad training than it is to be a world-class athlete with bad genetics and great training, if that makes sense.”
Mark McClusky Jan 2, 2019 ▶ 5:13 a16z Podcast | Sports, Tech, and What We Can All Learn from the Latest Performance Science

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.