reinforcement learning
also referred to as: rl
9 statements across 5 episodes · 5 bullish · 2 bearish · 5 people on the record · first statement Aug 21, 2017 by Daniel Gross · across every show →
Everything said about reinforcement learning, oldest first
Aug 21, 2017 bullish
Feb 20, 2025 bullish
Hiremath: Shift to reinforcement learning will create domain-specific reasoning AI
“The whole market is shifting to reinforcement learning, right? You're already seeing this with O-one, O-three, the deep seek models. And as a result, I think we're going to see really, really powerful models in specific domains that can reason extremely well.”
Jul 18, 2025 positive
Nov 3, 2025 negative
Pineau: Using reinforcement learning to teach AI social behavior remains unsolved
“RL, to shape the behavior of models, to get them to be social creatures, that we have no idea how to do. I mean, I don't know if you have children, but like shaping their behaviors, you know, the number of times you can repeat the same thing, and still they do…”
Nov 3, 2025 bearish
Pineau: Out-of-the-box reinforcement learning will not deliver AGI
“Now, you know, where we're maybe getting a little bit ahead is thinking that just RL out of the box is gonna give us AGI. That part, a lot less so. You know, if you look at the curve of progress, RL is terribly inefficient, and so the amount of signal you need…”
Nov 3, 2025 bullish
Pineau: Reinforcement learning is fundamental to AI and will not disappear
“Oh, I'm still super bullish on RL in that, like, the concept itself is so fundamental. You know, this idea of training through a system of rewards, of indicating what's valuable and what's not valuable through numerical values, like, that is so fundamental. It…”
Nov 3, 2025 positive