Dense Models

topic on 3 shows · 3 statements across 3 episodes

Latent Space No Priors the MAD Podcast

3 statements about Dense Models, every show

MAD Insight
Catanzaro: Dense models outperform MoE models under strict memory constraints
“You know, they take a lot more memory. If you have a very small amount of memory, a dense model is going to be smarter.”
Bryan Catanzaro Jul 2, 2026 ▶ 47:06 Inside Nemotron & NVIDIA’s AI Lab | Bryan Catanzaro
He: Upcycling dense models to MoE beats continuing dense training per FLOP
“By training these upcycled models, you can achieve better accuracy than simply training the dense model further for the same number of flops.”
Ethan He Oct 29, 2024 ▶ 20:32 [Paper Club] Upcycling Large Language Models into Mixture of Experts
NO PRIORS Insight
Guu: AI Researchers Should Prioritize Making Dense Representations More Controllable
“If I were kind of advising a student or something on looking into knowledge representations, at the moment I would say there's a lot of momentum on, on dense models continuing to capture more and more of the different applications, and so if there is a way tha…”
Kelvin Guu May 4, 2023 ▶ 28:39 No Priors Ep. 15 | With Kelvin Guu, Staff Research Scientist, Google Brain

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.