AI reasoning models
1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement May 21, 2026 by Yann Dubois · across every show →
Everything said about AI reasoning models, oldest first
May 21, 2026 positive
Dubois: RL allows AI reasoning models to backtrack wrong paths earlier
“Part of it is the model knowing when it's going down the wrong path. But this is also something that we can that the model can be trained for with reinforcement learning is like knowing, okay, like that seems like not a great path. Let me backtrack and let me …”