pre-trained model

also referred to as: pre-trained models

3 statements across 3 episodes · 1 bullish · 0 bearish · 3 people on the record · first statement Jan 22, 2020 by Clement Delangue · across every show →

Everything said about pre-trained model, oldest first

Jan 22, 2020
Insight
Delangue: Pre-trained NLP models require only a thin software layer
“Most of the intelligence is in the models, and the software engineering layer on top of these models is really thin. Which basically led these models to go to production really, really fast, right?”
Clement Delangue Jan 22, 2020 ▶ 9:55 NLP—The Most Important Field of ML // Clement Delangue, Hugging Face (FirstMark's Data Driven NYC)
Oct 23, 2025
Insight
Schrittwieser: Raw pre-trained AI models make poor agents without RL
“Our pre-training data is not very agent-like. If you think of the pre-training data, right, there is like websites and books and, you know, all kinds of recent text that has a lot of information, but it doesn't have a lot of actions. It doesn't really capture …”
Julian Schrittwieser Oct 23, 2025 ▶ 49:17 Are We Misreading the AI Exponential? Julian Schrittwieser on Move 37 & Scaling RL (Anthropic)
Jun 4, 2026 positive
Insight
Roberts: Powerful pre-trained models are necessary for effective RL and reasoning
“If you have a powerful enough pre-trained model, then it can start to do well at RL. It can start to like think at use test time compute to for instance, solve, solve math problems that it wouldn't otherwise be able to do.”
Dan Roberts Jun 4, 2026 ▶ 27:15 OpenAI's Dan Roberts: Why AI Can Now Make Discoveries
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.