Recurrent Neural Networks

topic on 3 shows · 10 statements across 8 episodes

the Y Combinator Startup Podcast Latent Space No Priors

10 statements about Recurrent Neural Networks, every show

Chaubard: Standard LLMs sacrifice latent time compression unlike RNNs
“What you actually paid for that you have to give up is this latent reasoning thing and this compression in the time direction. There is no compression in LMs. Every single decode that I do, I still have to retain the entire, you know, Shakespeare novel just to…”
Francois Chaubard May 1, 2026 ▶ 3:23 Recursion Is The Next Scaling Law In AI · Y Combinator
Chaubard: Hierarchical Reasoning Models Offer Little Novelty Over Standard RNNs
“The, this is directly in the lineage of RNNs. There's not that much novel from, like, the RNN standpoint at least in my opinion.”
Francois Chaubard May 1, 2026 ▶ 7:45 Recursion Is The Next Scaling Law In AI · Y Combinator
Morris: ChatGPT could likely have been built using RNNs instead of Transformers
“And I think like, we honestly probably could have gotten this with RNNs. I know like the scaling laws paper shows that RNNs have worse curves for scaling, but probably people would have been like, I bet you could have built chat GPT with a very sophisticated R…”
Jack Morris Jul 2, 2025 ▶ 1:10:28 Information Theory for Language Models: Jack Morris
NO PRIORS Assertion Supported
Vinyals: RNNs and LSTMs Never Remembered Beyond a Few Hundred Words
“We come from a world where we had recurrent neural networks and LSTMs that actually had infinite memory, although it was not very capable, right? You, the models in, in practice, they never remember more than a few hundred words or so.”
Oriol Vinyals Aug 1, 2024 ▶ 8:51 No Priors Ep. 74 | With Google DeepMind VP of Research Oriol Vinyals
LATENT SPACE Assertion Supported
RWKV Trains in Parallel Across GPUs, Unlike Traditional RNNs
“And in practice, once you start cascading there, you just saturate the GPU, and that's how it starts being paralysable trained. You no longer need to train in slices like traditional RNNs.”
Eugene Cheah Aug 31, 2023 ▶ 58:45 RWKV: Reinventing RNNs for the Transformer Era
LATENT SPACE Prediction Not checkable as stated
RNNs will outperform Transformers in batch generation and long sequences
“I am personally bullish on, on, on RNNs. I think RNNs they don't, they essentially summarize the past into a state vector. They have fixed size, so the size doesn't grow with the history. So that means that you don't need as much memory to keep around all the …”
Tri Dao Aug 3, 2023 ▶ 49:14 FlashAttention-2: Making Transformers 800% faster AND exact
NO PRIORS Insight
Transformers Beat Recurrent Models by Processing Whole Sequences at Once
“The magic of transformer kind of like convolutions is that you get to process the whole sequence at once. I mean, it still talks, you know, it's still a function of like the, you know, the predictions for the later words are dependent on what the earlier words…”
Noam Shazeer Apr 25, 2023 ▶ 4:50 No Priors Ep. 12 | With Noam Shazeer
Sidor: Rigorous baseline optimization will advance AI more than complex architectures
“And you know, it's not the kind of sexy research that people want to see, where you have like some hierarchy of big RNN, but it actually, this kind of research, I think at this point will advance field the most.”
Szymon Sidor Nov 8, 2017 ▶ 2:33 Building Dota Bots That Beat Pros - OpenAI's Greg Brockman, Szymon Sidor, and Sam Altman · Y Combinator
Eck: Google Magenta's early music generation models were primitive RNNs
“We put out some models that were, you know, by any reasonable account primitive, I mean, kind of very simple recurrent neural networks that that generate MIDI from MIDI”
Doug Eck Jul 21, 2017 ▶ 3:54 Making Music and Art Through Machine Learning - Doug Eck of Magenta · Y Combinator
Y COMBINATOR Assertion Supported
Eck: Google AI Duet relies on simple 2002 RNN technology
“And what I noticed, even with AI Duet, which is this like web-based, like, it's a simple RNN. It's like, I can lay claim. It's technology that was published in 2002. It's really a very simple, really simple.”
Doug Eck Jul 21, 2017 ▶ 13:16 Making Music and Art Through Machine Learning - Doug Eck of Magenta · Y Combinator

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.