pre-train models

1 statements across 1 episodes · 0 bullish · 0 bearish · 1 people on the record · first statement Jan 16, 2026 by Arthur Mensch · across every show →

Everything said about pre-train models, oldest first

Jan 16, 2026 neutral
Insight
Mensch: AI pre-training saturates at 10^26 FLOPs due to data limits
“Basically you have a saturation effect when you pre-train models around 10 to the power of 26 flops. The reason for that is that there's only that much data you can find to compress when you pre-train models.”
Arthur Mensch Jan 16, 2026 ▶ 26:08 Who Wins if AI Models Commoditize? — With Mistral CEO Arthur Mensch
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.