Post Training

topic on 11 shows · 37 statements across 33 episodes

the Y Combinator Startup Podcast Latent Space Lenny's Podcast No Priors Invest Like the Best the MAD Podcast the a16z Podcast Big Technology All-In TBPN 20VC

37 statements about Post Training, every show

20VC Prediction Not checkable as stated
Reyes: Companies will post-train internal commodity models for specialized tasks
“Businesses that will say, well, you know, we do a lot of commodity tasks, but there's a couple of very high volume specialized tasks. That only we do. And for those, your commodity model won't be good enough. Your frontier model will be too expensive. And so t…”
Eno Reyes Aug 29, 2026 ▶ 5:12 Should American Enterprises Work With Open-Source Chinese Models? | Only 10% of Neo-labs survive
MAD Insight
Trojanowski: Context engineering is runtime training with far less data
“I think the framework I like is thinking about it as training data. Except you're just training the model at runtime. It is training data. And because you, the model's learning at inference time, the total amount of training data is far lower, right? Like the …”
Mitch Trojanowski Aug 5, 2026 ▶ 43:17 How to Build Autonomous, Long-Horizon AI Agents | Basis
Kant: Big model post-training recipes transfer down to small models, not up
“It's not very helpful to have a post training recipe for a smaller model and try to apply it to a bigger model. Yeah. It just, in all cases, you're gonna have to rethink most of the recipe. But recipe for post training for a bigger model applied to a smaller m…”
Eiso Kant Jul 22, 2026 ▶ 1:44:06 The AI Frontier: from open weights to open research — Eiso Kant, Poolside AI
LATENT SPACE Disclosure
Bubna: Modal multi-node training targets post-training, not large-scale pre-training
“And we're not going for obviously like large scale pre-training runs. The thing that we've built multi-handle training for is we see a lot of smaller scale post-training like people are post-training like medium-sized fun models so they can get higher quality …”
Akshat Bubna Jul 8, 2026 ▶ 34:14 The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO
MAD Assertion Not checkable as stated
Zico Kolter: Reinforcement learning is now the foundation of all AI post-training
“RL is now the foundation of really all post training. It's all done by RL.”
Zico Kolter May 7, 2026 ▶ 1:04:27 OpenAI Board Member Zico Kolter: Modern AI Is Just 200 Lines of Code
NO PRIORS Insight
Srivastava: Startups should not do post-training before achieving product-market fit
“Hey, go find, go prove to yourself with the best in class model that you have something worth optimizing. And I think, you know, A lot of, you know, if a customer comes to us, was that meme, which was like, it was like two years ago, it feels like there's no G…”
Tuhin Srivastava May 1, 2026 ▶ 17:43 Baseten CEO Tuhin Srivastava on Custom Models, and Building the Inference Cloud
MAD Prediction Not checkable as stated
AI model progress will alternate between pre-training and post-training breakthroughs
“We're going to be having a bit of a swing back and forth between pre-training and post-training.”
Mostafa Dehghani Apr 2, 2026 ▶ 26:55 AI is Already Building AI — Google DeepMind’s Mostafa Dehghani
MAD Insight
Post-training techniques cannot compensate for a weak base AI model
“Pre-training is still the foundation and like, you can never post-train your way out of a week-based model.”
Mostafa Dehghani Apr 2, 2026 ▶ 27:01 AI is Already Building AI — Google DeepMind’s Mostafa Dehghani
Yi Tay: 'Reasoning' Technically Just Means Post-Training RL with Thinking Trajectories
“So I think the actual, like, technical definition of reasoning is making models better with thinking and post-training. Ok? Yeah. So basically, like, RL-ing the model to think better.”
Yi Tay Jan 23, 2026 ▶ 33:24 Captaining IMO Gold, Deep Think, On-Policy RL, Feeling the AGI in Singapore — Yi Tay
MAD Insight
Bourgeau: Recent continual learning progress has mostly occurred via post-training search tools
“First, I think a lot of progress has been made on this front since in the last few years. I think this is mostly around post-training, around search, use search tools and then make search calls, then they would have access to that new information.”
Sebastien Bourgeau Dec 18, 2025 ▶ 47:20 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
INVEST LIKE THE BEST Assertion Not checkable as stated
Post-training scaling laws drove all AI benchmark progress since October 2024
“And so all the progress we've had immense progress since October, 24 through today was based entirely on these two new scaling laws.”
Gavin Baker Dec 9, 2025 ▶ 9:49 GPUs, TPUs, & The Economics of AI Explained | Gavin Baker Interview · Invest Like The Best
Chen: AI post-training is an art driven by taste, not pure science
“One of the things I often think about is that there's a, it's almost like there's an art to post training. It's not purely a science. Like when you were deciding what kind of model you're trying to create and what it's good at. There's this notion of taste and…”
Edwin Chen Dec 7, 2025 ▶ 15:35 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
a16z Insight
Sherman Wu: Heavy compute for text model post-training bottlenecks verticalization
“For the text models, there's always going to be this like really big fat free training step that like you have to invest in here. And then even the post training side is like, You know, it's not the, it's not like the easiest thing. Like it's, you know we all,…”
Sherman Wu Nov 28, 2025 ▶ 41:51 How OpenAI Builds for 800 Million Weekly Users: Model Specialization and Fine-Tuning
a16z Assertion Not checkable as stated
David Owen: AI pre-training receives less focus due to post-training progress
“It seems as if pre-training is comparatively less of a focus than it was before, partly because, like, you have this exciting new direction of, well, new, newish direction of post-training where they've done so much about reasoning”
Epoch AI Researcher Nov 24, 2025 ▶ 6:16 The 2045 Superintelligence Timeline: Epoch AI’s Data-Driven Forecast
a16z Insight
David Owen: Post-training usage data generates feedback loops for pre-training
“A lot of this stuff is quite synergistic. You develop a better model. You, like, use post-training stuff to make it better. You get a load of data of the model actually being used successfully or not. A lot of that can probably go into pre-training next time.”
Epoch AI Researcher Nov 24, 2025 ▶ 6:41 The 2045 Superintelligence Timeline: Epoch AI’s Data-Driven Forecast
MAD What-if
Lambert: Scaling AI 10x alters post-training, not pre-training methods
“If like, if we were to train a model that was 10 times as big, like all this post-training stuff would change. But the pre-training And mid training and long contacts, I think would actually become looking pretty similar.”
Nathan Lambert Nov 20, 2025 ▶ 40:55 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Huyen: Internet data is maxed out, making post-training the key AI differentiator
“At some point, we are actually, like, have kind of maxed out on, like, internet data, right? And then people, like, text data, people max out. I think a lot of people are doing, like, with other data, like audios and videos, and, like, everyone's trying to thi…”
Chip Huyen Oct 23, 2025 ▶ 14:58 Al Engineering 101 with Chip Huyen (Nvidia, Stanford, Netflix)
Harris: Best AI innovations happen in post-training as data runs out
“It does seem like post training is where the best innovations are happening now and the pre-training and the amount of data, like they've, we've used up a lot of the data. They're trying to create synthetic data to try to improve model performance.”
Parker Harris Oct 14, 2025 ▶ 19:32 Will AI Kill Software? — With Salesforce Co-Founder Parker Harris
a16z Insight
Labenz: AI post-training reasoning currently yields higher ROI than raw scaling
“And it just seems like we're getting more benefit from the post training and the reasoning paradigm than scaling. But I don't think either one is I definitely don't think either one is, is dead.”
Nathan Labenz Oct 14, 2025 ▶ 13:14 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
BIG TECHNOLOGY Disclosure
Microsoft uses idle nighttime inference capacity for global AI post-training
“What's nice about post training is that you don't have to do it in one large data center in one location. And so part of the technique that we've been focused on is how do we take this inferencing capacity around the world? And a lot of it is idle at night as …”
Scott Guthrie Oct 1, 2025 ▶ 13:48 Microsoft's Cloud & AI Head on the AI Buildout's Risks and ROI — With Scott Guthrie
Anthropic's Joseph: Do Everything Possible in Post-Training Over Pre-Training
“The way I usually think about it is anything you can do in post training, you probably should, because your iteration loop, like the ability to make progress is really fast. You can try something, you can try it again, you can try it again.”
Nick Joseph Sep 30, 2025 ▶ 45:19 Anthropic Head of Pretraining on Scaling Laws, Compute, and the Future of AI · Y Combinator
Morcos: Post-training techniques are better applied in pre- and mid-training
“Most of what we do in post-training is better were done in pre and mid training and earlier on in training in general.”
Ari Morcos Aug 29, 2025 ▶ 42:16 Better Data is All You Need — Ari Morcos, Datology
Morcos: Post-training alignment is ineffective long-term compared to pre-training alignment
“Like fundamentally, I think alignment and post training doesn't really make sense as a long-term solution. If you can easily align a model through post training, you can easily misalign a model through post training. If it's easy to put it in, it's easy to tak…”
Ari Morcos Aug 29, 2025 ▶ 53:37 Better Data is All You Need — Ari Morcos, Datology
LENNY'S PODCAST Prediction Open · timeframe Aug 2030
Sharma: Industry spending on AI post-training will eventually surpass pre-training
“Like, I believe we will see, you know, just as much money spent on post-training as we will on pre-training, and in the future, more on post-training.”
Asha Sharma Aug 28, 2025 ▶ 45:03 How 80,000 companies build with AI: Products as organisms and the death of org charts | Asha Sharma
LENNY'S PODCAST Assertion Supported
Lord: AI pre-training gains asymptoted 18 to 24 months ago
“And about 18 months ago, 24 months ago, we started to really see, like, an asymptoting of gains coming from, because they had essentially, like, sucked up all of the knowledge on the internet. And so labs really shifted towards most of the gains now coming fro…”
Garrett Lord Aug 24, 2025 ▶ 6:35 Inside the expert network training every frontier AI model | Garrett Lord
a16z Insight
Kim: AI post-training functions more like art than traditional research
“For post-training, what's really f- or one of the reasons I really like post-training is it feels more like an art than maybe even, like, other areas of research, because you kind of have to make all these trade-offs, right?”
Christina Kim Aug 8, 2025 ▶ 4:39 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
BIG TECHNOLOGY Prediction Not checkable as stated
Lightcap: Post-training scaling will dominate AI development for the next 1–2 years
“You know, the O series of models, which were kind of the previous reasoning models were really just the beginning of us starting to explore what's possible in that post-training regime. And I think that's going to be kind of the dominant theme here for the nex…”
Brad Lightcap Aug 8, 2025 ▶ 7:18 OpenAI COO Brad Lightcap: GPT-5's Capabilities, Why It Matters, and Where AI Goes Next
Y COMBINATOR Prediction Not checkable as stated
Nadella: AI product creation will center on data feedback loops for post-training
“But the interesting thing is the feedback loop, the data path inside the product that then is used in order to post train, in order to be able to do the right tool you know, selection that seems to be the place where product creation is all gonna happen.”
Satya Nadella Jun 25, 2025 ▶ 6:31 Satya Nadella: Microsoft's AI Bets, Hyperscaling, Quantum Computing Breakthroughs · Y Combinator
LENNY'S PODCAST Prediction Not checkable as stated
Deng: Partnering post-training researchers with product drives AI breakthroughs
“I think that really that close, tight knit relationship between at any of these large model companies between post training and product is going to produce some really incredible stuff.”
Peter Deng Jun 22, 2025 ▶ 35:40 From ChatGPT to Instagram to Uber: The quiet architect behind the world’s most popular products
LATENT SPACE Disclosure
Brown: OpenAI models undergo mid-training and post-training before release
“For open AI models, like, they go through a mid-training step, and then they go through a post-training step, and then they're released, and they're a lot more useful. Like, frankly, if you interacted with the only pre-trained model, it would be super difficul…”
Noam Brown Jun 19, 2025 ▶ 1:11:38 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
ALL-IN Insight
Brin: Post-training and thinking models are a huge, uncapped step forward
“And more recently, the post-training, especially as the thinking models have come around. And that's been, you know, another huge step up in general in AI. So you know, we don't really know what the ceiling is.”
Sergey Brin May 20, 2025 ▶ 5:43 Sergey Brin, Google Co-Founder | All-In Live from Miami
TBPN Opinion
Kilpatrick: Pre-Training Isn't Dead; Gains Multiply Through Post-Training and RL
“And this is why, like, I don't subscribe to the, like pre-training is, you know, dead and all that stuff, because the more work that you can do at the pre-training level, those capabilities, as you do post-training and as you give the models RL capability, it'…”
Logan Kilpatrick Apr 25, 2025 ▶ 12:50 Google's AI Comeback in Their Own Words - Logan Kilpatrick
Patel: Generating pre-answer reasoning tokens yields superior AI performance
“Models now will think for some time before they answer. And this enables much better performance on all sorts of tasks, whether it be coding or math or understanding science or understanding complex Social dilemmas, right? All sorts of different topics they're…”
Dylan Patel Apr 23, 2025 ▶ 26:02 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Agarwal: Optimal Post-Training Pipeline Combines Heavy Distillation Followed by RL
“So, so I would think maybe an optimal pipeline would look like you do distillation heavily, but then you still do some RL afterwards, because maybe there's still something you can get out of your reward functions or whatever your post-training stack is.”
Rishabh Agarwal Mar 23, 2025 ▶ 8:05 The Magic of LLM Distillation — Rishabh Agarwal, Google DeepMind
Nguyen: Post-training scaling avoids data walls through infinite learnable tasks
“The scaling in post-chaining itself is not hitting the wall, and that's because Basically, we went from, like, raw data sets from pre-trained models to infinite amount of tasks that you can teach the model in the post-training world via reinforcement learning.…”
Karina Nguyen Feb 9, 2025 ▶ 9:56 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
MAD Assertion Not checkable as stated
Chip Huyen notes major AI labs keep post-training research proprietary
“And unfortunately, a lot of labs that are doing it are not quite, like, publishing papers about it.”
Chip Huyen Jan 16, 2025 ▶ 27:33 What You MUST Know About AI Engineering | Chip Huyen, Author of “AI Engineering”
MAD Insight
Chip Huyen argues post-training is what differentiates frontier AI models
“So, so I do think that post-training is what makes this, like, really big lab models are, like, different.”
Chip Huyen Jan 16, 2025 ▶ 28:37 What You MUST Know About AI Engineering | Chip Huyen, Author of “AI Engineering”

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.