post-training

also referred to as: post training

5 statements across 4 episodes · 1 bullish · 0 bearish · 4 people on the record · first statement Aug 8, 2025 by Christina Kim · across every show →

Everything said about post-training, oldest first

Aug 8, 2025
Insight
Kim: AI post-training functions more like art than traditional research
“For post-training, what's really f- or one of the reasons I really like post-training is it feels more like an art than maybe even, like, other areas of research, because you kind of have to make all these trade-offs, right?”
Christina Kim Aug 8, 2025 ▶ 4:39 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Oct 14, 2025 neutral
Insight
Labenz: AI post-training reasoning currently yields higher ROI than raw scaling
“And it just seems like we're getting more benefit from the post training and the reasoning paradigm than scaling. But I don't think either one is I definitely don't think either one is, is dead.”
Nathan Labenz Oct 14, 2025 ▶ 13:14 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
Nov 24, 2025 neutral
Assertion Not checkable as stated
David Owen: AI pre-training receives less focus due to post-training progress
“It seems as if pre-training is comparatively less of a focus than it was before, partly because, like, you have this exciting new direction of, well, new, newish direction of post-training where they've done so much about reasoning”
Epoch AI Researcher Nov 24, 2025 ▶ 6:16 The 2045 Superintelligence Timeline: Epoch AI’s Data-Driven Forecast
Nov 24, 2025 positive
Insight
David Owen: Post-training usage data generates feedback loops for pre-training
“A lot of this stuff is quite synergistic. You develop a better model. You, like, use post-training stuff to make it better. You get a load of data of the model actually being used successfully or not. A lot of that can probably go into pre-training next time.”
Epoch AI Researcher Nov 24, 2025 ▶ 6:41 The 2045 Superintelligence Timeline: Epoch AI’s Data-Driven Forecast
Nov 28, 2025
Insight
Sherman Wu: Heavy compute for text model post-training bottlenecks verticalization
“For the text models, there's always going to be this like really big fat free training step that like you have to invest in here. And then even the post training side is like, You know, it's not the, it's not like the easiest thing. Like it's, you know we all,…”
Sherman Wu Nov 28, 2025 ▶ 41:51 How OpenAI Builds for 800 Million Weekly Users: Model Specialization and Fine-Tuning
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.