training

5 statements across 4 episodes · 3 bullish · 0 bearish · 4 people on the record · first statement Sep 14, 2023 by Illia Polosukhin · across every show →

Everything said about training, oldest first

Sep 14, 2023 bullish
Insight
Polosukhin: Inference demands vastly more aggregate compute than AI model training
“I think an inference is really interesting because we do need so much more compute for inference than we need for training, right? Like it's a very interesting like economy of scale. You train once, like Lama trained once and then everybody runs it everywhere.”
Illia Polosukhin Sep 14, 2023 ▶ 23:23 No Priors Ep. 32 | With NEAR’s Illia Polosukhin
Mar 21, 2024 neutral
Insight
Srivastava: AI inference demands strict uptime, while training tolerates node terminations
“Resiliency and reliability matters a lot more. You know, downtime is unacceptable from an input perspective. Nodes get terminated all the time from a training perspective.”
Tuhin Srivastava Mar 21, 2024 ▶ 6:01 No Priors Ep 56 | With Baseten CEO and Co-Founder Tuhin Srivastava
Mar 21, 2024 neutral
Insight
Srivastava: Inter-rack networking matters less for AI inference than training
“Even the GPU clusters themselves, like, you know, the full training networking is a very, very important Piece to have networking on the racks themselves with inference and matters a little less because you're doing a little bit more on individual GPUs and les…”
Tuhin Srivastava Mar 21, 2024 ▶ 4:57 No Priors Ep 56 | With Baseten CEO and Co-Founder Tuhin Srivastava
Jan 9, 2025 positive
Disclosure
Bernhardsson: Modal Is Expanding into Bursty Experimental AI Training
“Traditionally, most of modal has always been inference. Like that's been our main use case, but we're really interested also in training. So in particular, like probably focused more on these like shorter, like very bursty sort of experimental training runs, n…”
Erik Bernhardsson Jan 9, 2025 ▶ 6:58 No Priors Ep. 96 | With Modal CEO and Founder Erik Bernhardsson
Jun 26, 2025 positive
Assertion Supported
Kohli: AlphaEvolve has successfully made AI model training more computationally efficient
“What Alpha Evolve has been able to do is basically make training more efficient.”
Pushmeet Kohli Jun 26, 2025 ▶ 26:01 No Priors Ep. 120 | With Google DeepMind’s Pushmeet Kohli and Matej Balog
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.