why aren't all 8 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Marcus: OpenAI's Project Orion failed and became GPT-4.5
“So OpenAI tried to build GPT-V and they had a thing called Project Orion and it actually failed. And eventually got released as GPT four and a half. So what they thought was going to be GPT five just didn't meet expectations.”
Prediction Held up
Morris: OpenAI will not declare GPT-5 as AGI
“I hate to make bold predictions, especially in tech, because you can often be spectacularly wrong, but I do not think they're going to say GPT-V is anything approaching AGI, however you choose to define it.”
Prediction Held up
GPT-5 will enable autonomous agents to execute UI tasks for hours
“Then we're gonna see much more multimodal data, and I think that'll look a lot like the equivalent of supervised fine-tuning, but for a bunch of people recording their screen and doing workflows with their screen, navigating UIs. So I think you'll have agents …”
Assertion Supported
Lightcap: GPT-5 bakes in tool use and longer-horizon reasoning
“So using tools, for example, is something that really thinks really important for overall intelligence, GPT two and three couldn't really do that as well. GPT-IV could do it in a more nascent way. And now GPT-V, you get that baked in with the benefit of these …”
Assertion Supported
Lightcap: Base GPT-5 beats GPT-4o even without added reasoning time
“Even though if you don't allow any thinking time you still get a typically net better answer than you would for one of our non-thinking models like GPT-IV-I.”
Assertion Supported
Lightcap: GPT-5 beats previous models on SWE-bench and health benchmarks
“It scores better on things like Sweebench. It scores better on all the kind of academic evals that we put it through. This one in particular, we actually made a real emphasis to have it score better on certain health benchmarks. So It's better at medical reaso…”
Prediction Held up
Roy: OpenAI will release GPT-5 before the end of 2025
“I think we see GPT-V this year.”
Prediction Held up
GPT-5 will have noticeably improved reasoning capabilities from step-by-step training
“I think it'll definitely be better at reasoning, which is trivial to say, because the training methods that we've seen them talk about, like you, I'm sure you heard the talk about Q star and what it seems to be is training the model to rewarding it on getting …”