Dan Hendrycks outlines why recursive self-improvement and rapid automated AI research pose severe loss-of-control risks compared to today's manageable AI misbehavior.
Prediction Not checkable as stated
Hendrycks: Superintelligence race will incentivize preemptive sabotage of rival AI projects
“I think that form of competition would be very dangerous, and because there's a risk of loss of control, and because it might incentivize states to engage in preventive sabotage or preemptive sabotage to disable these sorts of projects.”
Prediction Not checkable as stated
Hendrycks: Russia and China may threaten US data centers over superintelligence progress
“If the US were pulling ahead both Russia and China may have a substantial interest in saying, hey, cut this out. Pulling ahead to develop super intelligence, which could give it a huge advantage and an ability to crush crush them. They'd say, you don't get to …”
Prediction Not checkable as stated
Hendrycks: Automated AI R&D could compress a decade of progress into one year
“If you could have one AI do that, then you could make, you know, a 100,000 copies of these and have them perform research simultaneously. So that could lead to some very substantial acceleration in the rate of development. You might get a decade's worth of AI …”
Opinion
Hendrycks: AI poses no immediate extinction risk because it cannot make PowerPoints
“So, I don't think AI poses a risk of extinction like today, ok? I don't think that they're powerful enough to do that. They, because they can't make PowerPoints yet, right? They don't have Agential skills. They can't accomplish tasks that require many hours to…”
Prediction Not checkable as stated
Hendrycks: Cyberattacks and bioweapons are top malicious AI risks within two years
“In the shorter term when AIs get more agential, I'd be concerned about AIs causing cyber attacks on critical infrastructure, possibly by as directed by a rogue actor. There'd also be the risk of AIs facilitating the development of bioweapons in particular pand…”
Insight
Hendrycks: Loss-of-Control Risk Stems from Competitive Pressure to Fully Automate AI R&D
“Loss of control risks, which I think primarily stem from people, an AI company trying to automate all of AI research and development, and they can't have humans check in on that process because that would slow them down too much. If you have a human do a week …”