OpenAI O3

product on 21 shows · 9 statements across 8 episodes · said 254 times in 107 episodes since 2024

Latent Space 51 Big Technology 49 TBPN 38 the MAD Podcast 22 the Y Combinator Startup Podcast 20 No Priors 18 the a16z Podcast 12 BG2 Pod 7 Lenny's Podcast 6 the Startup Ideas Podcast 6 the Knowledge Project 5 All-In 4 20VC 4 Cheeky Pint 3 We Live to Build 2 My First Million 2 Sourcery 2 In Depth 1 BigDeal 1 WTF is with Nikhil Kamath 1 the Official SaaStr Podcast

Mentions by year, every show

tap a year for its mentions
0012550250100202420252026episodesmentions
050100202420252026episodes it came up in
002504100202420252026episodesmentions per episode

Latent Space 51Big Technology 49TBPN 38the MAD Podcast 22the Y Combinator Startup Podcast 20No Priors 18the a16z Podcast 12BG2 Pod 712 more shows

2026 23 mentions in 12 episodes 2 per episode
2025 219 mentions in 92 episodes 2 per episode
2024 12 mentions in 3 episodes 4 per episode

every mention on every show, scene by scene, with the transcript →

9 statements about OpenAI O3, every show

MAD Insight
Trojanowski: OpenAI o3 proved post-training improves per-token reasoning quality
“And then I think after a one, it was oh three, because I think oh three helped prove that not only could you scale the amount of reasoning at inference time, but with better training, with more compute, better data, et cetera, in the post-training phase, you c…”
Mitch Trojanowski Aug 5, 2026 ▶ 16:40 How to Build Autonomous, Long-Horizon AI Agents | Basis
LATENT SPACE Assertion Not checkable as stated
Swix: OpenAI Deep Research was built by three people as an o3 wrapper
“As far as I know, it's three people did it. It was Isa and like the two other collaborators that she had. I don't know if they did that much on top of all three, like every indication I've had from over the eye is that deep research is more or less a thin wrap…”
Shawn Wang Jul 31, 2025 ▶ 25:36 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
SAASTR Insight
Levie: Upgrading to reasoning models instantly replaces a year of custom scaffolding
“The amount of work you would have to do to pack into, let's say, a non-reasoning model for giving it exactly the right sort of context, instructions, kind of hacking and tool use, versus today, just having O-Tree go do that, or having the new Lama IV go and do…”
Aaron Levie Jul 2, 2025 ▶ 15:46 Why Enterprise AI Adoption Is Moving 5-10X Faster Than Cloud with Box 's Aaron Levie and IBM's VP AI
Brown: Reasoning models improve primarily through compute efficiency rather than longer thinking duration
“These models are becoming more efficient in the way they're thinking, as they're able to do more with the same amount of test time compute, and I think that's a very underappreciated point, that it's not just that we're getting these models to think for longer…”
Noam Brown Jun 19, 2025 ▶ 1:08:26 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
BIG TECHNOLOGY Prediction Not checkable as stated
Patel: Reinforcement learning 10x compute scaling can only continue for a year
“So already within the course of six months. RL compute has 10 X. That pace can only continue for a year, even if you build up all the RL environments before it's, you know, you're like, you've, you're at the frontier of training compute for these systems overa…”
Dwarkesh Patel Jun 18, 2025 ▶ 15:09 Dwarkesh Patel: AI Continuous Improvement, Intelligence Explosion, Memory, Frontier Lab Competition
Programmatic tools beat end-to-end image generation for multimodal reasoning
“Where I think for a while some people were, like, speculating, like, oh, what if you have the model, like, generate images in its chain of thought reasoning where everything is, like end-to-end multimodal input and output, and it seems like you don't really ne…”
Will Brown May 9, 2025 ▶ 6:01 ⚡️Open Questions in Agentic RL — Will Brown (Prime Intellect)
TBPN Insight
Will Brown: OpenAI o3 succeeds because reinforcement learning enables tool use
“The reason O-three is good is because it's trained to use tools. The way you train a model to use the right tool for the job is reinforcement learning. And they've said as much, like, deep research, reinforcement learning.”
Will Brown Apr 26, 2025 ▶ 7:33 Why Humor Is the True Test of AI Intelligence | Will Brown on TBPN
MAD Opinion
Chollet: OpenAI o3 is the most advanced test-time adaptation model
“And OSTRI best I can tell is the most advanced the most successful test and adaptation model out there at this time.”
Francois Chollet Apr 3, 2025 ▶ 17:53 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop
MAD Assertion Not checkable as stated
Chollet: OpenAI o3 cost $10k–$20k per ARC puzzle on maximum compute
“For instance OpenAI O.S. On the highest compute settings that we tried it on for Arc, it was consuming somewhere between, like, 10,000 dollars to 20,000 dollars per task, like, for one little puzzle, which you could normally solve with a base of an API for a f…”
Francois Chollet Apr 3, 2025 ▶ 20:00 Chasing Real AGI: Inside ARC Prize 2025 with Chollet & Knoop

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.