why aren't all 7 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Insight
Hendrycks: Fast-moving automated AI R&D loops are nearly impossible to de-risk
“Meanwhile, some automated AI research and development loop that goes extremely quickly with very little human oversight, that seems hard to de-risk and get the risks to be at a negligible level, just because it's so fast and so ah, and there's so little human …”
Insight
Hendrycks: Loss-of-Control Risk Stems from Competitive Pressure to Fully Automate AI R&D
“Loss of control risks, which I think primarily stem from people, an AI company trying to automate all of AI research and development, and they can't have humans check in on that process because that would slow them down too much. If you have a human do a week …”
Insight
Hendrycks: Biological AI capabilities are offense-dominant over defense
“So in, in bio, I think that is offense dominant. If somebody creates a virus, there's not necessarily a cure that it will immediately find for it. If it would help a rogue actor make A somewhat compelling virus. Now that could be enough to cause many millions …”
Insight
Hendrycks: Cyber AI in critical infrastructure is offense-dominant due to patching barriers
“Where in the context of critical infrastructure there the software is not updated rapidly. So even if you identify various vulnerabilities, there will not necessarily be a patch, because the system needs to always be on, or there are interoperability constrain…”
Insight
Hendrycks: AI labs spend over 90% of intellectual energy on scaling compute
“Over, you know, 90% of the intellectual energies that they're going to spend is actually, how can we afford the 10 X larger supercomputer? And, ah, that means being very competitive, speeding this up and making safety be some priority, but not necessarily a su…”
Insight
Hendrycks: Unilateral AI development pause makes no game theoretic sense
“An individual company pausing their development while others race ahead doesn't make game theoretic sense.”
Insight
Hendrycks: Automated AI R&D creates an insurmountable lead for early starters
“When you get to a different paradigm, like automated AI R&D, the slope might be extremely high, such that if the competitor starts to do automated AI R&D a year later, they may never catch up just because you're so far ahead and your gains are compounding on y…”