reinforcement learning

5 statements across 3 episodes · 2 bullish · 0 bearish · 3 people on the record · first statement Mar 31, 2026 by Sergey Levine · across every show →

Everything said about reinforcement learning, oldest first

Mar 31, 2026 positive
Insight
Levine: Robots Surpass Human Speed by Editing Out Cognitive Pauses
“It turns out to be like pretty straightforward to go in and like find all those pauses and remove them. And you can speed things up further, so you can get a task where a person demonstrates what it means to succeed, and then you can have the robot practice th…”
Sergey Levine Mar 31, 2026 ▶ 38:54 World's Top Researcher on AI, LLMs, and Robot Intelligence · Invest Like The Best
Mar 31, 2026 positive
Disclosure
Physical Intelligence aims to fuse generative AI prior knowledge with reinforcement learning
“So, I think the big challenge, and this is kind of what I'm leaning up to, and what I hope to, ah, that we'll figure out here at Physical Intelligence is how to combine those threads. How to bring in all of that knowledge that you get with generative AI, but a…”
Sergey Levine Mar 31, 2026 ▶ 16:38 World's Top Researcher on AI, LLMs, and Robot Intelligence · Invest Like The Best
Mar 31, 2026
Disclosure
Levine: Physical Intelligence trained espresso-making robot using repeated RL practice
“And for example, we had this demo on, ah, making espresso. That system practiced making those espressos many, many times and used that to improve robustness, improve speed, improve throughput.”
Sergey Levine Mar 31, 2026 ▶ 18:30 World's Top Researcher on AI, LLMs, and Robot Intelligence · Invest Like The Best
Apr 23, 2026 neutral
Insight
Patel: RL Simulation Environments Run on CPUs, Not GPUs or ASICs
“So the environments can get more and more complex, and those environments run on CPUs. They don't run on GPUs. They don't run on ASICs. The ASICs run the model,”
Dylan Patel Apr 23, 2026 ▶ 38:44 The Supply and Demand of AI Tokens | Dylan Patel Interview · Invest Like The Best
May 13, 2026 neutral
Insight
Rao: More efficient inference directly increases reinforcement learning efficiency
“If we're doing reinforcement learning on the model, it's basically inference within a sandbox with a reward function, right? And so if the model's better at more efficient inference, that RL is more efficient as well.”
Krishna Rao May 13, 2026 ▶ 9:44 Inside Anthropic's $100 Billion Al Compute Commitment | CFO Krishna Rao · Invest Like The Best
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 60 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.