Everything Dan Hendrycks said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Hendrycks: Geopolitical deterrence will delay superintelligence development
“This is why I think it might take a while for superintelligence to be developed, because there'll be deterrence around it later on.”
Hendrycks: AI models lied 20% to 60% of the time under pressure
“So we have a paper out last week, we're just measuring the extent to which they're deceptive. And in the scenarios we have, like all the models were in these sorts of scenarios under, you know, slight pressure to lie, not being told to lie, but just some sligh…”
Hendrycks: AI labs will break voluntary commitments when facing commercial competition
“I think voluntary commitments from AI companies are also a distraction because the companies will, you should expect most of them by default to just break those sorts of commitments if they end up going up against economic competitiveness.”
Hendrycks: Open-weight AI releases should be restricted for critical infrastructure cyber risks
“If for instance they have these cyber capabilities later on yeah, I think that, or I think that would be a potential place to be drawing the line on on open weight releases personally. In particular the ones that could cause damage to critical infrastructure.”
Hendrycks: US-China race will force rapid, high-risk military AI integration
“China can have AIs that are totally aligned with them. The U S can have AIs that are totally aligned with them. You still are going to have a strategic competition between the two. This is going to they're going to need to integrate it in their militaries. The…”
Hendrycks: Reasoning models have reached expert-level virology capabilities
“The AIs are getting very good at STEM PhD level types of topics, and that includes virology. So I think that they are sort of rounding the corner on being able to provide expert level capabilities in terms of their knowledge of the literature, Or even helping …”
Hendrycks: Over 30% of Staff at Some Top AI Labs Are Chinese Nationals
“I mean, we have, you know, some places, you know, 30% plus of the employees at these top AI companies are like Chinese nationals.”
Hendrycks: Adversaries Can Spy on Top AI Labs via Slack Zero-Day Exploits
“All they need to do is do a zero day on Slack. And then they can know what DeepMind is up to in very high fidelity and OpenAI and XAI and others.”
Hendrycks: Malicious human misuse is a bigger near-term AI risk than rogue AI
“I think those risks would potentially grow in time. I don't think they're as substantial now compared to just the malicious use sorts of risks.”
Hendrycks: AI alignment is only a subset of AI safety
“So I view the distinction between alignment and safety as alignment as being a sort of subset of safety. Obviously you want the value systems of the AIs to be in keeping with or compatible with say the US public for USAIs or for you as an individual, but that …”
Hendrycks: Current AI cannot enable devastating grid cyberattacks
“For cyber, I don't think AIs are that relevant for Being able to pull off a devastating cyber attack on the grid by a malicious actor currently.”
Hendrycks: Deterrence can restrict superpower intent, but not AI capabilities
“So you can restrict their intent, which is what deterrence does, but I don't think you can reliably or robustly restrict their capabilities.”
Hendrycks: Top AI models score only 10% to 20% on Humanity's Last Exam
“They're in the ballpark of, like, 10 to 20% overall. They're the very best models.”
Hendrycks: Reasoning training on code and math generalizes to domains like law
“There has been a fair amount of generalization from training on coding and mathematics to other sorts of domains like law, for instance.”
Hendrycks: I earn $1 a year from xAI and $12 from Scale
“My appointment at XAI, I get a dollar a year. At Scale, I've at Scale AI, I've increased my salary exponentially to where I get 12 dollars a year, a dollar per month from Scale.”
Hendrycks: Nations Will Race on AI Drones Without Coordinating Restraints
“No, like, super weapon type stuff, but more conventional type of warfare, like drones and things like that, I expect that they'll continue to race and Probably not, maybe not even coordinate on anything like that, but that's just how things would go.”
Hendrycks: Saturating Humanity's Last Exam benchmark will signal superhuman STEM AI
“When performance is near the ceiling, I think that'd basically be an indication that, like, you have something like a superhuman mathematician or a superhuman STEM scientist for, in many ways, for when they're, when closed-ended questions are very useful, such…”
Hendrycks: AI Models Remain 'Extremely Defective' as Autonomous Agents
“This could possibly change overnight, but it's still near the floor. I think they're still extremely defective as agents.”
Hendrycks: AI tail risks are systematically under-addressed
“Since it'd be such a big deal we'd need to make sure that we can think about it properly channel it in a productive direction, and take care of some sort of tail risks which will, are generally systematically under addressed.”