why aren't all 7 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Marcus: Models like o1 are not systematically better than GPT-4
“Models like O-I are not systematically better than GPT-IV. There's, they're better in certain use cases. Especially ones where you can create data in advance.”
Prediction Held up
Suleyman: GPT-4 and GPT-4o Efficiency Will Improve 100x
“So I expect that to happen for GPT-IV, GPT-IV-O and all of the other models down the road.”
Prediction Held up
Srinivas: Anthropic will release a model better than GPT-4 in 2024
“I actually think they will end up creating a model better than GPT-IV this year. Like, it, it's sort of almost guaranteed to happen. So I believe it's going to happen with CLAW-III.”
Assertion Supported
Ramaswamy: GPT-4 and Claude are a clear step ahead of open source models
“The blunt truth is that the very best of the models out there, whether it's GPT-IV or Claude's biggest model, are a clear step ahead of the pack when it comes to quality. When it comes to reasoning, when it comes to the quality of the text that they produce th…”
Assertion Supported
Lightcap: GPT-5 bakes in tool use and longer-horizon reasoning
“So using tools, for example, is something that really thinks really important for overall intelligence, GPT two and three couldn't really do that as well. GPT-IV could do it in a more nascent way. And now GPT-V, you get that baked in with the benefit of these …”
Assertion Supported
Patel: GPT-4-Level Training Costs Have Dropped 10x to 100x
“If you look at what it costs to train GBT for originally, I think it was like 20,008, 100 over the course of a hundred days. So I think it costs on the order of like half a million to a hundred million dollars, somewhere in that range. And I think you could tr…”
Assertion Supported
Patel: Model inference costs dropped 60x from GPT-4 to DeepSeek-V3
“And likewise, when we look at from GPT-IV to DeepSeq VIII it's fallen roughly 600 X in cost. Right. So we're not quite at that 1200 X, but it has fallen 600 X in cost from 60 dollars to less than you know, to about a dollar. Right. Or to less than a dollar. So…”