Assertion certainty 4/5 debate potential 2/5

Lambert: Larger pre-trained base models are easier to improve with RL

Nathan Lambert · Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking" · Nov 20, 2025 · at 44:29

Nathan Lambert, research scientist at AI2, explains why pre-training base models remains essential for post-training reinforcement learning.

0:00 / 0:04exact quote · 4.7s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“A better base model and a bigger base model is much easier to improve with RL.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Nathan Lambert

Assertion Not checkable as stated
Lambert: Chinese open AI models currently do not contain backdoors
“Like, you can't prove that the models aren't doing certain backdoors, where I'm fairly certain they definitely aren't now.”
Nathan Lambert Nov 20, 2025 ▶ 17:52 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Prediction Not checkable as stated
Lambert: AI progress will yield steady improvements rather than rapid singularity
“I think these researchers are going to grind out improvements for multiple years, but never in a way that results in this kind of accelerating well that we get drawn into.”
Nathan Lambert Nov 20, 2025 ▶ 1:21:34 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Assertion Not checkable as stated
Lambert: OLMo 3 32B base model matches Qwen 2.5 32B quality
“This base model is similar in quality to the best available, which is like Quinn's 2.5, 32 B is, was still the best base model.”
Nathan Lambert Nov 20, 2025 ▶ 4:06 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Assertion Not checkable as stated
Lambert: OLMo 3 7B outperforms Meta's Llama 3.1 8B in internal tests
“And I just think of this cause like Lama 3.1 AP is one of the most used models and hugging base of all time. And this should be better. We're, In our measurements, we see it as being better than Llama.”
Nathan Lambert Nov 20, 2025 ▶ 4:54 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Assertion Partly supported
Lambert: 80% of a16z's open-model portfolio startups use Alibaba's Qwen
“80% of companies building with open models are using Quinn, which is like 16 to 24% of his portfolio, which is still a lot.”
Nathan Lambert Nov 20, 2025 ▶ 17:01 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Assertion Not checkable as stated
Lambert: Chinese companies with $1B+ valuations routinely pirate SaaS software
“Mediumly large, like billion dollar plus valuation companies in China will just like pirate SaaS software.”
Nathan Lambert Nov 20, 2025 ▶ 18:47 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.