why aren't all 7 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Marcus: Models like o1 are not systematically better than GPT-4
“Models like O-I are not systematically better than GPT-IV. There's, they're better in certain use cases. Especially ones where you can create data in advance.”
Assertion Partly supported
DeepSeek used model distillation to match OpenAI's o1 at lower cost
“And, you know, they've just used the process of distillation to, you know, effectively bring those bigger versions of the sort of state-of-the-art models and distill them down into you know, smaller models, which eventually led to this R-one, you know, the equ…”
Assertion Supported
Wang: China's DeepSeek produced the first replication of OpenAI's o1 model
“OpenAI released O-one and released the O-one preview a number of months ago... Yeah, this is OpenAI's advanced reasoning model, which is great at sort of scientific reasoning and mathematical reasoning and reasoning and code, et cetera. And the very first repl…”
Prediction Partly held up
Patel: GPT-5 will simultaneously scale pre-training and post-training reasoning
“And so now GPT-Five, as Sam calls it, is, is gonna be a model that has huge pre-training scale, right? Like GPT-Five, but also huge post-training scale, Like O-one and O-three and continuing to scale that up, right? This would be the first time we see a model …”
Assertion Supported
OpenAI o1 Benchmarks Show Major Gains Across Math, Coding, and Science
“So if you think about, ah, competition math, ah, GPT-IV-O was getting a 13.4 accuracy score on competition math. But this thing and the way that it can think through the different problems is getting 83.3, ah, score on accuracy. Ah, with competition math. So i…”
Assertion Supported
OpenAI Notably Avoided Using the Word 'Agent' in o1 Announcement
“Open AI hasn't actually used the word agent once in its announcement, and it doesn't sound like any of its scientists have talked about that either.”
Assertion Supported
Kantrowitz: OpenAI's Blog Post States o1 Does Not Solve Hallucinations
“Yeah, and OpenAI even in its blog post says that it does not solve, this does not solve hallucinations, and then you can”