why aren't all 7 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Not checkable as stated
Combining reinforcement learning with pre-training outperforms scaling pre-training alone
“If you were just trying to scale pre-training, you wouldn't get anywhere near as far as also trying to scale RL on top of pre-training, which is what we do now.”
Assertion Supported
ChatGPT disproved an Erdős conjecture using cross-disciplinary mathematical reasoning
“The big result was that this conjecture of this lower bound for the number of pairs that you can make is, is false. Not only is it false, it was false due to a really interesting connection to another field of mathematics.”
Assertion Not checkable as stated
Current AI models lack research taste and problem formulation ability
“There's part of the scientific process. I think that the models haven't been imbued with yet. And I'm sure people are thinking about how to do that. You know, like what, trying to get to what is the right question as opposed to here's a well-defined thing and …”
Prediction Not checkable as stated
Dan Roberts expects more AI-driven math and science breakthroughs within six months
“For the next six months. Like, I think we'll see more of these sorts of math and science breakthroughs.”
Prediction Not checkable as stated
AI systems evolving into fully fledged scientists will be a gradual transition
“The, there's no sharp point, or I don't think there will be a sharp point where we'll say that systems didn't, weren't able to be useful for scientific, the scientific process to their fully fledged scientists. There'll be sort of a gradual shift.”
Prediction Not checkable as stated
OpenAI will release reinforcement learning products for consulting, banking, and legal
“I definitely think OpenAI will have amazing products that will be relevant in those domains, and some amount of RL will play a role in there.”
Prediction Not checkable as stated
Continuous AI improvements render multi-year autonomous agent runs highly inefficient
“In, in general, we're not just going to like set up a system and let it think autonomously for eight years, if anything, because like the system's eight years. After will be so much more powerful that it probably doesn't make sense to let a system think for a …”