Everything Dan Hendrycks said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Hendrycks: U.S. and China may launch preemptive cyberattacks on data centers
“I think that later on, it becomes so destabilizing that China just says, we're going to do something preemptive, like do a cyber attack on your data center. And the U.S. Might do that to China.”
Hendrycks: Superintelligence race will incentivize preemptive sabotage of rival AI projects
“I think that form of competition would be very dangerous, and because there's a risk of loss of control, and because it might incentivize states to engage in preventive sabotage or preemptive sabotage to disable these sorts of projects.”
Hendrycks: Russia and China may threaten US data centers over superintelligence progress
“If the US were pulling ahead both Russia and China may have a substantial interest in saying, hey, cut this out. Pulling ahead to develop super intelligence, which could give it a huge advantage and an ability to crush crush them. They'd say, you don't get to …”
Hendrycks: Extreme AI export controls make a Taiwan invasion more likely
“If you turn the pain dial all the way up for China in export controls and if AI chips are the currency of economic power in the future, then this increases the probability that they want to invade Taiwan.”
Hendrycks: Voluntary AI Pauses Without Enforcement Only Benefit Bad Actors
“If you do it voluntarily, you just make yourself less powerful and you let the worst actors get ahead of you. You could say, well, we'll try and try to sign a treaty. We will not assume that the treaty will be followed. Like that would be very imprudent. You w…”
Hendrycks: AI Safety Is a Geopolitical Problem, Not Primarily a Technical One
“Safety isn't, as I've been I'm trying to reinforce not really that much of a technical problem. This is more of a complex geopolitical problem with technical aspects.”
Hendrycks: China will steal model weights even if chip controls succeed
“Even so, I still think if you really tighten the export controls, you made it so that China can't get any of those chips at all, and this is your, one of your biggest priorities, they're just going to steal the weights anyway.”
Hendrycks: Automated AI R&D could compress a decade of progress into one year
“If you could have one AI do that, then you could make, you know, a 100,000 copies of these and have them perform research simultaneously. So that could lead to some very substantial acceleration in the rate of development. You might get a decade's worth of AI …”
Hendrycks: Fast-moving automated AI R&D loops are nearly impossible to de-risk
“Meanwhile, some automated AI research and development loop that goes extremely quickly with very little human oversight, that seems hard to de-risk and get the risks to be at a negligible level, just because it's so fast and so ah, and there's so little human …”
Hendrycks: AI poses no immediate extinction risk because it cannot make PowerPoints
“So, I don't think AI poses a risk of extinction like today, ok? I don't think that they're powerful enough to do that. They, because they can't make PowerPoints yet, right? They don't have Agential skills. They can't accomplish tasks that require many hours to…”
Hendrycks: Cyberattacks and bioweapons are top malicious AI risks within two years
“In the shorter term when AIs get more agential, I'd be concerned about AIs causing cyber attacks on critical infrastructure, possibly by as directed by a rogue actor. There'd also be the risk of AIs facilitating the development of bioweapons in particular pand…”
Hendrycks: Loss-of-Control Risk Stems from Competitive Pressure to Fully Automate AI R&D
“Loss of control risks, which I think primarily stem from people, an AI company trying to automate all of AI research and development, and they can't have humans check in on that process because that would slow them down too much. If you have a human do a week …”
Hendrycks: Recent AI reasoning models score in 90th percentile on wet-lab guidance
“We are finding that with the most recent reasoning models quite unlike the models from two years ago, like the initial GPT-IV, the most recent reasoning models are getting around 90th percentile compared to these expert level virologists in their area of exper…”
Hendrycks: AI Has Brainstormed Ways to Make Viruses More Dangerous for Over a Year
“Them doing brainstorming to come up with ways to make viruses more dangerous, I think that's a capability that they've had for over a year, the brainstorming part, but the implementation part seems to be fairly different.”
Hendrycks: Consensus on Expert-Level AI Bio Capabilities Coming Within Months
“So, I think in bio, actually, the I would not be surprised if in a few months there's a consensus that there expert level in many relevant ways in that we need to be doing something about that.”
Hendrycks: Mastering Humanity's Last Exam will signal superhuman math capabilities
“When there's very high performance on, on that benchmark that would be suggestive of something that has, say, in the ballpark of superhuman mathematician capabilities. And so I think that would revolutionize the, academy quite substantially. Because all the th…”
Hendrycks: The next-token prediction paradigm in AI is running out of steam
“That sort of paradigm does seem like it's running out of steam. It has held for many, many orders of magnitude but the returns on doing that are lower.”
Hendrycks: RL-based reasoning models are improving faster than pre-training did
“That is separate from the new reasoning paradigm that has emerged in the past year which is where you train models to on math and coding types of questions with reinforcement learning, and that has a very steep slope, and I don't see any signs of that slowing …”
Hendrycks: Biological AI capabilities are offense-dominant over defense
“So in, in bio, I think that is offense dominant. If somebody creates a virus, there's not necessarily a cure that it will immediately find for it. If it would help a rogue actor make A somewhat compelling virus. Now that could be enough to cause many millions …”
Hendrycks: Cyber AI in critical infrastructure is offense-dominant due to patching barriers
“Where in the context of critical infrastructure there the software is not updated rapidly. So even if you identify various vulnerabilities, there will not necessarily be a patch, because the system needs to always be on, or there are interoperability constrain…”
Hendrycks: AI labs spend over 90% of intellectual energy on scaling compute
“Over, you know, 90% of the intellectual energies that they're going to spend is actually, how can we afford the 10 X larger supercomputer? And, ah, that means being very competitive, speeding this up and making safety be some priority, but not necessarily a su…”
Hendrycks: Unilateral AI development pause makes no game theoretic sense
“An individual company pausing their development while others race ahead doesn't make game theoretic sense.”
Hendrycks: Elon Musk's X successfully reduced self-censorship in US discourse
“I think that overall in terms of cultural influence and people being more disagreeable and doing less self-censoring has been has been successful. I think that was the main objective of it. And so I think I think that X had a large role to play there. So I don…”
Hendrycks: Automated AI R&D creates an insurmountable lead for early starters
“When you get to a different paradigm, like automated AI R&D, the slope might be extremely high, such that if the competitor starts to do automated AI R&D a year later, they may never catch up just because you're so far ahead and your gains are compounding on y…”