why aren't all 7 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Not checkable as stated
Patel: $10B AI data centers aim to automate software engineering, not chatbots
“No one is trying to make with these, you know, with these ten billion dollar data centers, they're not trying to make chat models, right? They're not trying to make models that people chat with, just to be clear, right? They're trying to solve things like soft…”
Assertion Not checkable as stated
Patel: OpenAI's Orion training run failed to reach GPT-5 performance levels
“There were hopes that Orion could be used for GPT-V but its improvement was, like, not enough to be, like, really a GPT-V. Furthermore, it was trained on the classical method, which is, like which is a ton of pre-training, and then some reinforcement learning …”
Assertion Supported
Patel: Model inference costs dropped 60x from GPT-4 to DeepSeek-V3
“And likewise, when we look at from GPT-IV to DeepSeq VIII it's fallen roughly 600 X in cost. Right. So we're not quite at that 1200 X, but it has fallen 600 X in cost from 60 dollars to less than you know, to about a dollar. Right. Or to less than a dollar. So…”
Assertion Not checkable as stated
Patel: Frontier AI cluster costs have scaled from $100M to $10B
“For GPT-IV, it was a few hundred million dollars and it's one building full of GPUs, too. GPT-IV 4.5 and the reasoning models, like, oh, one, oh, three were done in a, in three buildings on the same site, and, you know, billions of dollars to, hey, these next …”
Prediction Partly held up
Patel: GPT-5 will simultaneously scale pre-training and post-training reasoning
“And so now GPT-Five, as Sam calls it, is, is gonna be a model that has huge pre-training scale, right? Like GPT-Five, but also huge post-training scale, Like O-one and O-three and continuing to scale that up, right? This would be the first time we see a model …”
Prediction Held up
Patel: Meta's next Llama model will match DeepSeek-V3's cost efficiency
“And Meta's Meta is going to release their new llama soon enough. Right. And that one is going to be, you know, a similar level of cost decrease probably similar areas, deep seek V three.”
Assertion Supported
Patel: AI inference costs for GPT-3-level performance have dropped 1,200x
“So when we looked at GPT-III, the cost fell of 1200 X from GPT-III's initial cost to what you can get LLAMA three point two three B today, right?”