Trojanowski: Nothing Paradigm-Shifting Has Changed Since o3
“Things have really not changed since oh three, I would say. Almost everything since oh three has been relatively on, I don't want to say on trend and then like I knew this exact trend, but I would say it's all within the same paradigm. Like nothing paradigm sh…”
Brockman: Doctors used OpenAI's o3 to diagnose a 20-year mysterious ailment
“Like, for example, today we announced I we have in peer reviewed literature, people, doctors who are using O three. Remember O three? It was like forever ago now. Right. It was like one of our earliest reasoning models using that to find diagnoses for People w…”
Brockman: Doctors use OpenAI's o3 model to diagnose long-unsolved medical cases
“Today we announced I we have in peer-reviewed literature people, doctors who are using O. Three. Remember O. Three? I was like forever ago now, right? It was like one of our earliest reasoning models using that to find diagnoses for People who I had no, no ans…”
French-Owen: OpenAI used reinforcement learning on o3 checkpoints for coding tasks
“Well, basically we kind of had like a checkpoint. For I think it was O three, like one of the reasoning models. And then we did a bunch of fine tuning on it and reinforcement learning where it's like, oh, you're given a bunch of questions to like solve these c…”
Tworek: GPT-5 can effectively be considered an iteration like 'o3.1'
“Like GPT-Five in some way I can be considered as like, oh, 3.1. It's a little bit of like, you know, iteration of like the same thing and the same concept”
Wu: OpenAI deployed o3 on an air-gapped Los Alamos supercomputer
“We actually did a custom on-prem deployment with them onto one of their supercomputers called Venado. And so this actually involves a bunch of, you know very bespoke work with some FDs also with a lot of our developer team. To actually bring one of our reasoni…”
Brockman: Wet lab tests of o3 produced mid-tier journal-level work
“We have wet lab scientists who took models like O-three, ask it for some hypotheses of, here's an experimental setup, what should I do? They have five ideas, They tried these five ideas out, four of them don't work, but one of them does. And the kind of feedba…”
Brockman: OpenAI's 80% o3 price cut yielded neutral or positive revenue
“And you can see it with O three, I think we did like an 80% price cut and actually the usage grew such that it was like, I think in the revenue, it either was neutral or positive.”
Hays: An o3-level reasoning model was open-sourced within a year of o1
“So less than one year between O one announced, which was September of 20, 24. And we have an O three level model open sourced that's runnable on consumer hardware, wild progress.”
Sonwalkar: GPT-5 costs half as much as o3
“Also it's half the cost of O three. So it's much cheaper. So it helps you, helps your margins.”
Kim: GPT-5 front-end coding is a massive leap over o3
“If you compare it to O three's front end coding capability, this is just totally next level.”
Lambert: Deep Research relies on modular RL tasks rather than end-to-end outcomes
“I think the deep research blog post kind of hints that they do a bunch of small scale RL and then poof, the system works. Which I think is much more of what's happening is people train on a bunch of small things and they do some prompting and they see that whe…”
Liu: Everyone should run their lab diagnostics through AI like ChatGPT
“I mean, I think, like, today, it's already better to, I mean, everybody should be taking all of their, you know, personal, like health diagnostics, you get a blood test or whatever, and, like, running that through ChatGPT, like, upload it, Asko three, like, He…”
Altman: OpenAI's o3 model cost dropped by 5x in one week
“And also, like last week, O-three cost five times as much as it did this week, and that's gonna keep going.”
Altman: AI models like o3 can act as junior employees executing multi-hour tasks
“But now you start to see things where you can like really give a task to Codex, for example or to deep research, and you have this thing go off and do a bunch of stuff and come back to you with like a proposal. It's like a very junior employee that can work on…”
Brown: OpenAI's o3 Gets 'Not Very Far' Playing Pokémon Unharnessed
“How far does O three get without any harness? How far does it get playing Pokemon? And the answer is like, not very far, you know?”
Noam Brown: OpenAI o3 has basically replaced Google Search for me
“Like I've been using it day to day. It's basically replaced Google search for me. Like I just use it all the time.”
OpenAI's technology will surpass o3 within six months
“I think that Oh, three is not where the technology will be in six months.”
Patel: OpenAI's o3 is currently the smartest AI model on the market
“I do think O three is the smartest model on the market right now.”
Taggar: o3 follows rubrics rigidly while Gemini 2.5 Pro reasons flexibly
“Oh, three was very rigid, actually. Like it really sticks to the rubric. It's heavily penalizes for anything that doesn't fit like the rubric that you've given it. Whereas Gemini 2.5 pro was actually quite good at being flexible in that it would apply the rubr…”
Courtney: OpenAI o3 Recreated a $35K, 3-Month Whitepaper in 10 Minutes
“We spent. 35,000 dollars to get a white paper created three years ago, which we used in a lot of marketing. Basically we wanted this company to go and research all of these things that were happening in businesses so that we could have all of these crazy quote…”
Marcus: OpenAI's o3 hallucinates more than preceding models
“I'll give you just one more example is O three apparently hallucinates more than the models that came before it.”
Mitchell: o3 autonomously executes multi-step tasks using integrated tools
“Not only is the model it's on its own smarter than our previous O series models, which is great, but it's also able to use all these tools that like further enhance its abilities and whether that's doing like research on something where you want up-to-date inf…”
McKinzie: Reinforcement learning is the key differentiator behind o3 reasoning
“I guess the short answer is reinforcement learning is, is the biggest one. So yeah, rather than just having to predict the next token and some large pre-training corpus from, you know you know, everywhere essentially now we have a more focused goal of the mode…”
Brown: OpenAI's o3 qualifies as 10-minute AGI
“I'm happy to call O three, 10 minute AGI. And I think like framing AGI in terms of like length of time, it takes a human to do a task is like more reasonable than like a global framing. Like, sure. There's a bar of like drop and replace for a human that we are…”
Brown: OpenAI's o3 already achieves practical program synthesis
“Like when people say program synthesis, like we're already there, like O three is program synthesis, but the programs are like JSON and Python.”
Srinivas: Most AI foundation model investment will go toward reinforcement learning
“I think RL is the place where most investments are gonna go to especially with models like O-three that are able to do tool calls pretty natively rather than being prompt engineered to do that”
Srinivas: Perplexity replaced a three-model agent pipeline with a single model
“Before O-three, the way we built, like, our agents, is there would be one model that gave, came up with a plan for the query, another model that would execute the plan by converting the plan to, like, Smaller queries, filtering links, calling searches, and the…”
Fulford: OpenAI Built Deep Research by Fine-Tuning o3
“Yeah, I think also the base model, or the model that we started fine tuning from O three is just a very capable model. It's trained on many different data sets, including a lot of coding and reasoning and math tasks.”
Patel: GPT-5 will simultaneously scale pre-training and post-training reasoning
“And so now GPT-Five, as Sam calls it, is, is gonna be a model that has huge pre-training scale, right? Like GPT-Five, but also huge post-training scale, Like O-one and O-three and continuing to scale that up, right? This would be the first time we see a model …”
Chollet: GPT-4 lacks fluid intelligence, but OpenAI's o3 model has it
“GPT-IV does not have fluid intelligence, for instance, but O-III does.”
OpenAI's o3 outperforms GPT-4o on Convex evals by a small margin
“You know, oh, three does do better than four. Oh, I mean, we use brain trust for tracking all this quantitatively, but I can't remember off the top of my head, but it's not like a slam dunk.”
Fernando: OpenAI's upcoming o3 model will take the industry by storm
“And so it's going to be really exciting when they actually dropped O three. I think a lot of people are going to be taken by storm of like, what's actually really going to come out from them. It's going to be a really, really big leap.”
Coogan: OpenAI's o3 High-Compute Mode Spent $3,000 to Solve a Benchmark Task
“Oh three, which isn't out yet, but is even more advanced in terms of reasoning. They have a high compute model. That spends almost 3000 dollars per task. And it just thinks for hours and hours and hours basically, and it was able to break arc that that AI, AGI…”
Beauchamp: AI intelligence is generative LLMs combined with tree search
“I think if you want to talk about what would intelligence look like, it looks much more like tree search. Combining the generative nature of these LLMs with a really good tree search. And that's what opening I've done with O-one and O-three.”
OpenAI likely reaches frontier capabilities via search, then distills into mini models
“The only way you reach the frontier with the full size models of O-one and O-three is with that stuff. And then you can distill to the minis, the O-one mini, O-three mini. So in my writeup, I said like, maybe this is the formula for O-one mini, O-three mini. T…”
Coogan: OpenAI o3 high-compute configuration costs $2,000 per solve
“O-three is a reasoning model. The high version costs 2000 dollars per solve.”
Coogan: OpenAI o3 is underpriced and represents a major breakthrough
“Oh, three, I think is still underpriced. It's a big deal. Very, very big deal.”
Friedman: Sam Altman Says OpenAI o2 and o3 Are Not Far Behind
“Sam was just telling us that like O-two and O-three are not far behind.”