reinforcement learning
also referred to as: rl
8 statements across 7 episodes · 7 bullish · 0 bearish · 7 people on the record · first statement Feb 9, 2025 by Karina Nguyen · across every show →
Everything said about reinforcement learning, oldest first
Feb 9, 2025 bullish
Nguyen: Post-training scaling avoids data walls through infinite learnable tasks
“The scaling in post-chaining itself is not hitting the wall, and that's because Basically, we went from, like, raw data sets from pre-trained models to infinite amount of tasks that you can teach the model in the post-training world via reinforcement learning.…”
Mar 13, 2025 bullish
LLMs improve faster at coding due to its deterministic execution
“Software is deterministic. When you write code and you hit run, either runs or it doesn't. And that is why that's the key insight Anthropic really had. They just went deep. And this is what they're doing. It's just reinforcement learning. I'm basically permuta…”
May 4, 2025 positive
May 4, 2025 bullish
Jul 20, 2025 bullish
Aug 28, 2025 bullish
Sharma: Reinforcement learning will become a critical product technique
“With the advent of agents and products that think and can act and reason, there's going to be this kind of new wave around RL, and I have a deep belief that that, that will become one of the most important product techniques, kind of the next season, or at lea…”
Sep 18, 2025 bullish
Sep 6, 2026 neutral
Enterprises will split work between cheap open-weight and expensive frontier models
“I think what we're going to see is a split between job functions that demand kind of mid IQ intelligence, and those will often be open weight, sort of biased with reinforcement learning, you know, things that make the models even cheaper, more performant for a…”