RL Scaling

topic on 2 shows · 3 statements across 3 episodes

the MAD Podcast Big Technology

3 statements about RL Scaling, every show

MAD Insight
Bourgeau: Vibe coding performance stems mostly from RL scaling and post-training
“I think this is, yeah, this is in general for vibe coding specifically, I think that's maybe more of an RL scaling and post training thing where, where you can actually get quite a lot of data and train them all to do that really well.”
Sebastien Bourgeau Dec 18, 2025 ▶ 46:15 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
MAD Assertion Not checkable as stated
Douglas: OpenAI's o1 established test-time compute and RL as a scaling axis
“And I think OpenAI deserves a lot of credit for you know, releasing the first, like, serious RL plus LLMs release with O-one. And I think this really kicked off a pretty, you know, substantial change because it opened up a new axis of scaling, right? There was…”
Sholto Douglas Oct 2, 2025 ▶ 50:01 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
BIG TECHNOLOGY Assertion Not checkable as stated
Patel: RL scaling is outpacing pre-training scaling
“RL scaling is happening much faster than even overall training scaling.”
Dwarkesh Patel Jun 18, 2025 ▶ 13:38 Dwarkesh Patel: AI Continuous Improvement, Intelligence Explosion, Memory, Frontier Lab Competition

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.