Insight certainty 4/5 debate potential 2/5

Gupta: Recursive architectures achieve compute depth without parameter depth

Ankit Gupta · Recursion Is The Next Scaling Law In AI · Y Combinator · May 1, 2026 · at 24:32

Ankit Gupta breaks down the structural advantages of Tiny Recursive Models (TRMs) versus standard scaling of transformer layers.

0:00 / 0:13exact quote · 13.1s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“Recursion advantage now gives you a bunch of advantages over transformers where rather than having, you know, 500 or a thousand or a million or whatever transformer layers and having tons and tons of parameters, you get compute depth basically without this parameter depth.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Ankit Gupta

Insight
Gupta: Machine learning progresses by abandoning bio-plausibility for computational efficiency
“I think machine learning tends to have a long history of people starting with bio-plausible arguments, and then realizing that there's some variant of them that seems highly bio-implausible that actually works better.”
Ankit Gupta May 1, 2026 ▶ 14:35 Recursion Is The Next Scaling Law In AI · Y Combinator
Insight
Gupta: Recursive Latent Hidden States Function Like a Turing Machine Tape
“I kind of think of this set of hidden states or carry as akin to a Turing machine tape or akin to the radix sort, ah, memory bank, where you can basically train a model to use this memory cache in an intelligent way in a single forward pass so that you can get…”
Ankit Gupta May 1, 2026 ▶ 17:03 Recursion Is The Next Scaling Law In AI · Y Combinator
Insight
Gupta: Chain of Thought Operates in Token Space, Not Model Recursion
“There already exists some type of recursion that people are used to in LLMs, which is a chain of thought we mentioned earlier, but that is a recursion that's happening in the token space of the model's outputs, not inherent to the model itself. That's sort of …”
Ankit Gupta May 1, 2026 ▶ 19:01 Recursion Is The Next Scaling Law In AI · Y Combinator
Assertion Supported
Gupta: Current TRMs and HRMs are task-specific, not general-purpose
“One of the things that's really interesting about these TRMs and HRMs is they're not general purpose models, right? These were Task specific models, right? The model trained to do Sudoku cannot do ArcPrize inherently. It has to be trained on the ArcPrize set t…”
Ankit Gupta May 1, 2026 ▶ 36:19 Recursion Is The Next Scaling Law In AI · Y Combinator
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.