why aren't all 26 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 1 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Marcus: OpenAI's Project Orion failed and became GPT-4.5
“So OpenAI tried to build GPT-V and they had a thing called Project Orion and it actually failed. And eventually got released as GPT four and a half. So what they thought was going to be GPT five just didn't meet expectations.”
Assertion Not checkable as stated
Lightcap: GPT-5 is four to five times more accurate than predecessor models
“GBD-V, I think depends on how you measure it, but it's, you know, four to five times more accurate than its predecessors.”
Opinion
Lightcap: GPT-5 does not qualify as an AGI system
“And so I do, I think we're at a system that I would call AGI. No. But I think we see, we start to see the traces and the pieces of that overall system for generalized learning start to come together in models like GPT-V and I suspect suspect in its successors.”
Opinion
Lightcap: GPT-5 capability overhang would fuel ten years of product building
“I think you could pause AI progress right here for 10 years, and you'd still have about a decade worth of new products to get built, of new ways that people figure out how to use the models even at a GPT-V level model in interesting products and interesting pr…”
Prediction Held up
Morris: OpenAI will not declare GPT-5 as AGI
“I hate to make bold predictions, especially in tech, because you can often be spectacularly wrong, but I do not think they're going to say GPT-V is anything approaching AGI, however you choose to define it.”
Assertion Not checkable as stated
Patel: OpenAI's Orion training run failed to reach GPT-5 performance levels
“There were hopes that Orion could be used for GPT-V but its improvement was, like, not enough to be, like, really a GPT-V. Furthermore, it was trained on the classical method, which is, like which is a ton of pre-training, and then some reinforcement learning …”
Opinion
Saunders: GPT-4 is safe, but GPT-5 or later might be the Titanic
“I don't think that I was working on the Titanic. I don't think that GPT four was the Titanic. I'm more, I'm afraid that like GPT five or GPT six or GPT seven might be the Titanic in, in, in this analogy.”
Prediction Open · timeframe May 2029
Training next-generation frontier AI models will soon require gigawatts of power
“And then how much, basically, would it cost in terms of energy to train a GPT four level model, a 4.5 level model, five, whatever. And you get into the gigawatts pretty soon.”
Prediction Held up
GPT-5 will enable autonomous agents to execute UI tasks for hours
“Then we're gonna see much more multimodal data, and I think that'll look a lot like the equivalent of supervised fine-tuning, but for a bunch of people recording their screen and doing workflows with their screen, navigating UIs. So I think you'll have agents …”
Opinion
Roy: GPT-5 felt like a massive dud compared to expectations
“Again, GPT-V, we all know, felt like a massive dud that was supposed to be that magic moment for everyone”
Assertion Not checkable as stated
OpenAI's latest model leveled Anthropic's dominant lead among Cursor developers
“GPT five, the model opening I released a month ago at this point, I believe two months ago, it really changed that. You know, people have talked to a cursor say that, you know, it completely Leveled the kind of usage between OpenAI and Anthropics models in a w…”
Assertion Supported
Lightcap: GPT-5 bakes in tool use and longer-horizon reasoning
“So using tools, for example, is something that really thinks really important for overall intelligence, GPT two and three couldn't really do that as well. GPT-IV could do it in a more nascent way. And now GPT-V, you get that baked in with the benefit of these …”
Assertion Supported
Lightcap: Base GPT-5 beats GPT-4o even without added reasoning time
“Even though if you don't allow any thinking time you still get a typically net better answer than you would for one of our non-thinking models like GPT-IV-I.”
Insight
Lightcap: Post-training and test-time compute act as force multipliers
“And that continues to hold true, but we now have this kind of other category of training, which is post-training and being able to use test time compute in more interesting ways than we used to as almost kind of a second stage of training. And so we think that…”
Assertion Supported
Lightcap: GPT-5 beats previous models on SWE-bench and health benchmarks
“It scores better on things like Sweebench. It scores better on all the kind of academic evals that we put it through. This one in particular, we actually made a real emphasis to have it score better on certain health benchmarks. So It's better at medical reaso…”
Prediction Not checkable as stated
Kantrowitz: There is a decent chance OpenAI calls GPT-5 AGI
“I think there is a decent chance. I'm not saying it's for sure going to happen. I think there's a decent chance that they are going to say it's AGI that GPT five is AGI and by they, I mean, open AI.”
Prediction Partly held up
Patel: GPT-5 will simultaneously scale pre-training and post-training reasoning
“And so now GPT-Five, as Sam calls it, is, is gonna be a model that has huge pre-training scale, right? Like GPT-Five, but also huge post-training scale, Like O-one and O-three and continuing to scale that up, right? This would be the first time we see a model …”
Prediction Not checkable as stated
Chen: GPT-5 could combine unsupervised scaling with reasoning paradigms
“And so I think, like GPT-V really could be the culmination of a lot of these things coming together.”
Prediction Didn’t hold up
Kantrowitz: OpenAI is unlikely to release GPT-5 in 2025 or ever
“Again, I doubt we're seeing GPT-V this year, or maybe ever.”
Prediction Held up
Roy: OpenAI will release GPT-5 before the end of 2025
“I think we see GPT-V this year.”
Disclosure
OpenAI Tested GPT-5 With Uber, Amgen, Cursor, and JetBrains Pre-Release
“We've worked with large enterprises and small startups and the entire spectrum in between on testing these models and GPT-V specifically before release. And we get a lot of feedback from companies like Uber and Amgen and Harvey and Cursor lovable you know JetB…”
Prediction Not checkable as stated
Lightcap: GPT-5 will feel dramatically different to average users, not power users
“And so we expect that like for, yeah, for the average user, it will feel dramatically different. Maybe for the kind of upper echelon of power user, it may not feel as different.”
Disclosure
Lightcap: OpenAI is bringing GPT-5 to ChatGPT's free tier
“And we're bringing GPT-V to our free tier”
Prediction Not checkable as stated
Levie: 10x cheaper AI models will drive 100x more usage
“If you could make, again, kind of wave magic wand and you say like we have GPT five or GPT six, and it costs like a 10th
Of what today GPT-IV costs I would argue that you'll probably get a hundred X more usage of AI, not, not just 10 X, you know, more usage of…”
Prediction Held up
GPT-5 will have noticeably improved reasoning capabilities from step-by-step training
“I think it'll definitely be better at reasoning, which is trivial to say, because the training methods that we've seen them talk about, like you, I'm sure you heard the talk about Q star and what it seems to be is training the model to rewarding it on getting …”
Disclosure
Lightcap: OpenAI prioritized healthcare applications during GPT-5 training
“We focused on health a lot with this release because that was one of the consistently common things that we heard from people as a starting point for how they've used powerful AI was in, when they're navigating a health journey. And so we really wanted to make…”