o3

9 statements across 6 episodes · 6 bullish · 0 bearish · 6 people on the record · first statement Jan 24, 2025 by Shawn Wang · said 39 times in 14 episodes since 2025 · across every show →

Mentions by year

brought up most by Greg Brockman (7), Noam Brown (6), Raaz Dwivedi (5), Nathan Lambert (5), Will Brown (4), Alex Duffy (4), Alessio Fanelli (2), Stephanie Palazzolo (1)

tap a year for its mentions
00208401520252026episodesmentions
081520252026episodes it came up in
001.57.531520252026episodesmentions per episode
2026 1 mention in 1 episode
2025 38 mentions in 13 episodes 3 per episode

every mention, scene by scene, with the transcript →

Everything said about o3, oldest first

Jan 24, 2025 neutral
Opinion
OpenAI likely reaches frontier capabilities via search, then distills into mini models
“The only way you reach the frontier with the full size models of O-one and O-three is with that stuff. And then you can distill to the minis, the O-one mini, O-three mini. So in my writeup, I said like, maybe this is the formula for O-one mini, O-three mini. T…”
Shawn Wang Jan 24, 2025 ▶ 11:05 The Unreasonable Effectiveness of Reasoning Distillation: using DeepSeek R1 to beat OpenAI o1
Jan 26, 2025 bullish
Insight
Beauchamp: AI intelligence is generative LLMs combined with tree search
“I think if you want to talk about what would intelligence look like, it looks much more like tree search. Combining the generative nature of these LLMs with a really good tree search. And that's what opening I've done with O-one and O-three.”
William Beauchamp Jan 26, 2025 ▶ 1:10:38 Outlasting Noam Shazeer, Crowdsourcing Chai AI w/ 1.4m DAU — with William Beauchamp, Chai Research
Mar 19, 2025 positive
Assertion Not checkable as stated
OpenAI's o3 outperforms GPT-4o on Convex evals by a small margin
“You know, oh, three does do better than four. Oh, I mean, we use brain trust for tracking all this quantitatively, but I can't remember off the top of my head, but it's not like a slam dunk.”
Sujay Jayakar Mar 19, 2025 ▶ 17:56 Fullstack-Bench: The Eval for Coding Agents — with Sujay Jayakar, Chief Scientist, Convex
Jun 19, 2025 bullish
Prediction Held up
OpenAI's technology will surpass o3 within six months
“I think that Oh, three is not where the technology will be in six months.”
Noam Brown Jun 19, 2025 ▶ 38:35 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
Jun 19, 2025 positive
Disclosure
Noam Brown: OpenAI o3 has basically replaced Google Search for me
“Like I've been using it day to day. It's basically replaced Google search for me. Like I just use it all the time.”
Noam Brown Jun 19, 2025 ▶ 35:46 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
Jun 19, 2025 neutral
Assertion Not checkable as stated
Brown: OpenAI's o3 Gets 'Not Very Far' Playing Pokémon Unharnessed
“How far does O three get without any harness? How far does it get playing Pokemon? And the answer is like, not very far, you know?”
Noam Brown Jun 19, 2025 ▶ 14:33 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
Jul 31, 2025 neutral
Opinion
Lambert: Deep Research relies on modular RL tasks rather than end-to-end outcomes
“I think the deep research blog post kind of hints that they do a bunch of small scale RL and then poof, the system works. Which I think is much more of what's happening is people train on a bunch of small things and they do some prompting and they see that whe…”
Nathan Lambert Jul 31, 2025 ▶ 7:35 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Aug 15, 2025 positive
Assertion Not checkable as stated
Brockman: OpenAI's 80% o3 price cut yielded neutral or positive revenue
“And you can see it with O three, I think we did like an 80% price cut and actually the usage grew such that it was like, I think in the revenue, it either was neutral or positive.”
Greg Brockman Aug 15, 2025 ▶ 44:15 Greg Brockman on OpenAI's Road to AGI
Aug 15, 2025 positive
Assertion Not checkable as stated
Brockman: Wet lab tests of o3 produced mid-tier journal-level work
“We have wet lab scientists who took models like O-three, ask it for some hypotheses of, here's an experimental setup, what should I do? They have five ideas, They tried these five ideas out, four of them don't work, but one of them does. And the kind of feedba…”
Greg Brockman Aug 15, 2025 ▶ 12:01 Greg Brockman on OpenAI's Road to AGI
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.