why aren't all 12 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 2 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Marcus: OpenAI's Project Orion failed and became GPT-4.5
“So OpenAI tried to build GPT-V and they had a thing called Project Orion and it actually failed. And eventually got released as GPT four and a half. So what they thought was going to be GPT five just didn't meet expectations.”
Assertion Not checkable as stated
Chen: GPT-4.5 performance jump matches leap from GPT-3.5 to GPT-4
“It signifies an order of magnitude improvement over the last models, kind of commensurate with the jump from 3.5 to four.”
Assertion Not checkable as stated
Chen: GPT-4.5 scaling returns remain consistent with OpenAI's prior projections
“You know, we are seeing the same returns, and I do want to stress that GPT-D 4.5 is that next point on this unsupervised learning paradigm, and, you know, we're very rigorous about how we do this. We make projections based on all the models we've trained befor…”
Assertion Not checkable as stated
Chen: GPT-4.5 hits expected benchmark progression consistent with OpenAI's trajectory
“Well, I really don't think that the accurate characterization is that it doesn't hit the benchmarks that, that we expect it to. So when you look at kind of the development of three to 3.5 to four to 4.5 this does hit the benchmarks that we expect.”
Prediction Open · timeframe May 2029
Training next-generation frontier AI models will soon require gigawatts of power
“And then how much, basically, would it cost in terms of energy to train a GPT four level model, a 4.5 level model, five, whatever. And you get into the gigawatts pretty soon.”
Assertion Open · timeframe Aug 2026
OpenAI's GPT-4.5 Was Built as a Trillion-Parameter Dense Model
“GPT, 4.5 was an experimental model from OpenAI. It was the idea, let's train a trillion parameter dense model, meaning it is not sparse, meaning all the token, all the neurons are activated on every request. And it was so slow. It's really hard to run these th…”
Disclosure
Chen: Focus on reasoning models caused the longer gap before GPT-4.5
“Why there seems to be, you know, a little bit bigger of a gap in release time between four and 4.5, we've been really largely focused on developing the reasoning parallel paradigm as well.”
Assertion Partly supported
Chen: Users prefer GPT-4.5 over GPT-4o by 60% to 70% margins
“When we look at, kind of, comparisons against GPT-FORO you'll see that everyday use cases, people prefer, you know, by a margin of 60% for actually productivity and knowledge work against GPT-FORO, there's almost like a 70% preference rate.”
Prediction Held up
Srinivas: OpenAI will stay ahead of competitors with GPT-4.5 or GPT-5
“I'm sure there's a GPT 4.5 or five that will stay ahead. So it really is going to be a cat and mouse game there where Anthropics playing catch up and OpenAI is ahead through multimodal capabilities, reasoning capabilities, and things, things like that.”
Opinion
Chen: GPT-4.5 outshines reasoning models like o1 in creative writing
“And, you know, we find that in a lot of areas like creative writing, for instance
Again, this is stuff that we want to test over the next one or two months but we find that there are areas like creative writing where this model outshines reasoning models.”
Assertion Not checkable as stated
Chen: GPT-4.5 creates ASCII art almost flawlessly, unlike previous models
“If you ask any of the previous models to create ASCII art for you, right? Actually, they mostly just fall down. This one can do it Almost flawless.”
Assertion Not checkable as stated
Chen: Pausing and restarting training runs is standard across OpenAI models
“Actually, so I think it's interesting that this gets is a point that's attributed to this model because actually in, in, in developing all of our foundation models, right they're all experiments, right? I think you know, running all of the foundation models of…”