why aren't all 16 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 1 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Prediction Not checkable as stated
Hendrycks: U.S. and China may launch preemptive cyberattacks on data centers
“I think that later on, it becomes so destabilizing that China just says, we're going to do something preemptive, like do a cyber attack on your data center. And the U.S. Might do that to China.”
Prediction Not checkable as stated
Hendrycks: Extreme AI export controls make a Taiwan invasion more likely
“If you turn the pain dial all the way up for China in export controls and if AI chips are the currency of economic power in the future, then this increases the probability that they want to invade Taiwan.”
Insight
Hendrycks: Voluntary AI Pauses Without Enforcement Only Benefit Bad Actors
“If you do it voluntarily, you just make yourself less powerful and you let the worst actors get ahead of you. You could say, well, we'll try and try to sign a treaty. We will not assume that the treaty will be followed. Like that would be very imprudent. You w…”
Insight
Hendrycks: AI Safety Is a Geopolitical Problem, Not Primarily a Technical One
“Safety isn't, as I've been I'm trying to reinforce not really that much of a technical problem. This is more of a complex geopolitical problem with technical aspects.”
Prediction Not checkable as stated
Hendrycks: China will steal model weights even if chip controls succeed
“Even so, I still think if you really tighten the export controls, you made it so that China can't get any of those chips at all, and this is your, one of your biggest priorities, they're just going to steal the weights anyway.”
Prediction Not checkable as stated
Hendrycks: US-China race will force rapid, high-risk military AI integration
“China can have AIs that are totally aligned with them. The U S can have AIs that are totally aligned with them. You still are going to have a strategic competition between the two. This is going to they're going to need to integrate it in their militaries. The…”
Assertion Not checkable as stated
Hendrycks: Reasoning models have reached expert-level virology capabilities
“The AIs are getting very good at STEM PhD level types of topics, and that includes virology. So I think that they are sort of rounding the corner on being able to provide expert level capabilities in terms of their knowledge of the literature, Or even helping …”
Assertion Not checkable as stated
Hendrycks: Over 30% of Staff at Some Top AI Labs Are Chinese Nationals
“I mean, we have, you know, some places, you know, 30% plus of the employees at these top AI companies are like Chinese nationals.”
Assertion Not checkable as stated
Hendrycks: Adversaries Can Spy on Top AI Labs via Slack Zero-Day Exploits
“All they need to do is do a zero day on Slack. And then they can know what DeepMind is up to in very high fidelity and OpenAI and XAI and others.”
Insight
Hendrycks: AI alignment is only a subset of AI safety
“So I view the distinction between alignment and safety as alignment as being a sort of subset of safety. Obviously you want the value systems of the AIs to be in keeping with or compatible with say the US public for USAIs or for you as an individual, but that …”
Opinion
Hendrycks: Current AI cannot enable devastating grid cyberattacks
“For cyber, I don't think AIs are that relevant for Being able to pull off a devastating cyber attack on the grid by a malicious actor currently.”
Insight
Hendrycks: Deterrence can restrict superpower intent, but not AI capabilities
“So you can restrict their intent, which is what deterrence does, but I don't think you can reliably or robustly restrict their capabilities.”
Prediction Open · timeframe Mar 2030
Hendrycks: Nations Will Race on AI Drones Without Coordinating Restraints
“No, like, super weapon type stuff, but more conventional type of warfare, like drones and things like that, I expect that they'll continue to race and Probably not, maybe not even coordinate on anything like that, but that's just how things would go.”
Prediction Not checkable as stated
Hendrycks: Saturating Humanity's Last Exam benchmark will signal superhuman STEM AI
“When performance is near the ceiling, I think that'd basically be an indication that, like, you have something like a superhuman mathematician or a superhuman STEM scientist for, in many ways, for when they're, when closed-ended questions are very useful, such…”
Opinion
Hendrycks: AI Models Remain 'Extremely Defective' as Autonomous Agents
“This could possibly change overnight, but it's still near the floor. I think they're still extremely defective as agents.”
Opinion
Hendrycks: AI tail risks are systematically under-addressed
“Since it'd be such a big deal we'd need to make sure that we can think about it properly channel it in a productive direction, and take care of some sort of tail risks which will, are generally systematically under addressed.”