RL scaling

2 statements across 2 episodes · 1 bullish · 0 bearish · 2 people on the record · first statement Oct 2, 2025 by Sholto Douglas · across every show →

Everything said about RL scaling, oldest first

Oct 2, 2025 positive
Assertion Not checkable as stated
Douglas: OpenAI's o1 established test-time compute and RL as a scaling axis
“And I think OpenAI deserves a lot of credit for you know, releasing the first, like, serious RL plus LLMs release with O-one. And I think this really kicked off a pretty, you know, substantial change because it opened up a new axis of scaling, right? There was…”
Sholto Douglas Oct 2, 2025 ▶ 50:01 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Dec 18, 2025 neutral
Insight
Bourgeau: Vibe coding performance stems mostly from RL scaling and post-training
“I think this is, yeah, this is in general for vibe coding specifically, I think that's maybe more of an RL scaling and post training thing where, where you can actually get quite a lot of data and train them all to do that really well.”
Sebastien Bourgeau Dec 18, 2025 ▶ 46:15 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.