Gradient Descent

topic on 3 shows · 5 statements across 5 episodes

the Y Combinator Startup Podcast Latent Space No Priors

5 statements about Gradient Descent, every show

Chollet: Gradient Descent Fails at Reasoning by Defaulting to Pattern Matching
“You could not really get Gradient descent to encode sort of like reasoning style algorithms. It was not because the models could not represent these algorithms. It was because gradient descent could not find them, right? So the problem was that it wasn't about…”
François Chollet Mar 27, 2026 ▶ 14:11 François Chollet: Why Scaling Alone Isn’t Enough for AGI · Y Combinator
Yi Tay: Gradient descent learning paradigm is AI's bottleneck, not architecture
“It's not architecture itself. That's, that there's a problem that we, that is more of like the learning paradigm itself rather than the architecture itself. I think the architecture is just basically like the interface between the learning algorithm and the to…”
Yi Tay Jan 23, 2026 ▶ 49:07 Captaining IMO Gold, Deep Think, On-Policy RL, Feeling the AGI in Singapore — Yi Tay
Y COMBINATOR Assertion Not checkable as stated
Chollet: Gradient descent requires 3 to 4 orders of magnitude more data than humans
“Gradient descent requires vast amounts of data to distill simple abstractions. Many orders of magnitude more data than what humans need. Roughly three to four orders of magnitude more.”
François Chollet Jul 3, 2025 ▶ 24:13 François Chollet: How We Get To AGI · Y Combinator
Yao: Reflexion replaces scalar RL rewards with verbal gradient descent
“I think one way to think of reflection is that the traditional idea of reinforcement learning is you have a scalar reward, and then you somehow back propagate the signal of the scalar reward. To the rest of your neural network through whatever algorithm, like …”
Shunyu Yao Sep 27, 2024 ▶ 15:35 Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph
NO PRIORS Insight
Aravind Srinivas argues AI application builders understand LLMs better than trainers
“The people who use the LLMs for building stuff, Understand it better than the people who actually did gradient descent and train these models.”
Aravind Srinivas Apr 25, 2023 ▶ 10:15 No Priors Ep. 9 | With Perplexity AI’s Aravind Srinivas and Denis Yarats

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.