Pre Train Models

topic on 1 show · 1 statements across 1 episodes

Big Technology

1 statements about Pre Train Models, every show

Mensch: AI pre-training saturates at 10^26 FLOPs due to data limits
“Basically you have a saturation effect when you pre-train models around 10 to the power of 26 flops. The reason for that is that there's only that much data you can find to compress when you pre-train models.”
Arthur Mensch Jan 16, 2026 ▶ 26:08 Who Wins if AI Models Commoditize? — With Mistral CEO Arthur Mensch

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.