reinforcement learning

2 statements across 2 episodes · 2 bullish · 0 bearish · 2 people on the record · first statement Apr 15, 2025 by Bret Taylor · across every show →

Everything said about reinforcement learning, oldest first

Apr 15, 2025 positive
Disclosure
Taylor: OpenAI o1 used reinforcement learning on chains of thought
“What at OpenAI, what we did with the O-one model, which is to do some reinforcement learning those chains of thought to really reach new levels of intelligence.”
Bret Taylor Apr 15, 2025 ▶ 46:44 Bret Taylor: A Vision for AI’s Next Frontier
Apr 22, 2026 positive
Assertion Not checkable as stated
Brockman: OpenAI's 10-year roadmap focused on RL, unsupervised learning, then complexity
“We came up with what I would Really say is almost the technical plan that we have pursued for the past 10 years. Number one, solve reinforcement learning. Number two, solve unsupervised learning. And number three was gradually learn more complicated, in quotes…”
Greg Brockman Apr 22, 2026 ▶ 4:00 Ai Goes Parabolic | OpenAI Co-Founder Greg Brockman
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.