Karpathy
1 statements across 1 episodes · 0 bullish · 1 bearish · 1 people on the record · first statement Feb 5, 2026 by Didi Das · across every show →
Everything said about Karpathy, oldest first
Feb 5, 2026 bearish
Das: Reinforcement learning is an inefficient paradigm requiring massive sample sizes
“One is RL's kind of a shitty paradigm to learn. Karpathy obviously talks about this a lot. It takes a lot of samples to learn some very basic stuff because you only get a reward at the end. You don't actually understand things as it's happening.”