Mar 5, 2025 · 36m · no-priors

No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks

Dan Hendrycks · 29m spoken Sarah Guo · 5m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

Center for AI Safety Director Dan Hendrycks discusses catastrophic risk mitigation, proposing that AI safety requires geopolitical deterrence frameworks, compute governance, and rigorous capability evaluations rather than narrow laboratory alignment.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. The hosts hold 14.7% of the talking time here. How this is scored →

The hosts as informed peer 5.5 Guest teaching 5.3 Guest disagreement 3.7 The hosts pushing back 3.3
05100:0010:0020:0030:001:24–4:48 · The hosts as informed peer 4/10 Frontier Lab Constraints and Alignment Versus Safety Guo questions Hendrycks' dismissive stance toward frontier lab safety efforts and probes the semantic distinction between alignment and safety. Hendrycks reframes alignment as merely a small, obedient subset of safety that fails to resolve overarching structural and geopolitical competition.4:48–10:07 · The hosts as informed peer 6/10 National Security Implications and Biosecurity Safeguards Guo cites specific venture portfolio investments across cybersecurity and biotech to probe the trade-off between safety and competitive benefits. Hendrycks bluntly dismisses the notion of an intractable trade-off, arguing that dangerous biosecurity capabilities can simply be gated behind enterprise sales tiers.10:08–17:50 · The hosts as informed peer 6/10 Military AI Applications, Espionage, and Competition Realities Guo contributes technical domain knowledge regarding electronic warfare, autonomous systems, and radar communications in modern combat. Hendrycks educates on game theory, dismantling simplistic race narratives by citing corporate espionage realities and high proportions of Chinese national researchers at US labs.17:51–25:35 · The hosts as informed peer 5/10 Mutually Assured AI Malfunction and Deterrence Policy Guo asks Hendrycks to break down his Mutually Assured AI Malfunction framework and its nuclear parallels into concrete policy steps. Hendrycks details offensive cyber deterrence, CIA espionage cells, and licensing regimes analogous to fissile material tracking.25:36–30:37 · The hosts as informed peer 7/10 Compute Security, Algorithmic Efficiency, and Export Controls Guo presses Hendrycks on whether DeepSeek and compute-efficient pretraining undermine hardware-based export controls and compute tracking. Hendrycks argues DeepSeek actually validates the need for intent deterrence rather than capability restriction, while Guo challenges the assumption that rival powers would ever forego commercial capabilities.30:37–35:58 · The hosts as informed peer 5/10 The Frontier of AI Evals and Autonomous Agents Guo queries how evaluation methodology must evolve once models surpass human benchmarks. Hendrycks outlines the jagged frontier between high-level reasoning/proof verification and defective autonomous agent capabilities.1:24–4:48 · Guest teaching 5/10 Frontier Lab Constraints and Alignment Versus Safety Guo questions Hendrycks' dismissive stance toward frontier lab safety efforts and probes the semantic distinction between alignment and safety. Hendrycks reframes alignment as merely a small, obedient subset of safety that fails to resolve overarching structural and geopolitical competition.4:48–10:07 · Guest teaching 5/10 National Security Implications and Biosecurity Safeguards Guo cites specific venture portfolio investments across cybersecurity and biotech to probe the trade-off between safety and competitive benefits. Hendrycks bluntly dismisses the notion of an intractable trade-off, arguing that dangerous biosecurity capabilities can simply be gated behind enterprise sales tiers.10:08–17:50 · Guest teaching 6/10 Military AI Applications, Espionage, and Competition Realities Guo contributes technical domain knowledge regarding electronic warfare, autonomous systems, and radar communications in modern combat. Hendrycks educates on game theory, dismantling simplistic race narratives by citing corporate espionage realities and high proportions of Chinese national researchers at US labs.17:51–25:35 · Guest teaching 6/10 Mutually Assured AI Malfunction and Deterrence Policy Guo asks Hendrycks to break down his Mutually Assured AI Malfunction framework and its nuclear parallels into concrete policy steps. Hendrycks details offensive cyber deterrence, CIA espionage cells, and licensing regimes analogous to fissile material tracking.25:36–30:37 · Guest teaching 5/10 Compute Security, Algorithmic Efficiency, and Export Controls Guo presses Hendrycks on whether DeepSeek and compute-efficient pretraining undermine hardware-based export controls and compute tracking. Hendrycks argues DeepSeek actually validates the need for intent deterrence rather than capability restriction, while Guo challenges the assumption that rival powers would ever forego commercial capabilities.30:37–35:58 · Guest teaching 5/10 The Frontier of AI Evals and Autonomous Agents Guo queries how evaluation methodology must evolve once models surpass human benchmarks. Hendrycks outlines the jagged frontier between high-level reasoning/proof verification and defective autonomous agent capabilities.1:24–4:48 · Guest disagreement 4/10 Frontier Lab Constraints and Alignment Versus Safety Guo questions Hendrycks' dismissive stance toward frontier lab safety efforts and probes the semantic distinction between alignment and safety. Hendrycks reframes alignment as merely a small, obedient subset of safety that fails to resolve overarching structural and geopolitical competition.4:48–10:07 · Guest disagreement 5/10 National Security Implications and Biosecurity Safeguards Guo cites specific venture portfolio investments across cybersecurity and biotech to probe the trade-off between safety and competitive benefits. Hendrycks bluntly dismisses the notion of an intractable trade-off, arguing that dangerous biosecurity capabilities can simply be gated behind enterprise sales tiers.10:08–17:50 · Guest disagreement 4/10 Military AI Applications, Espionage, and Competition Realities Guo contributes technical domain knowledge regarding electronic warfare, autonomous systems, and radar communications in modern combat. Hendrycks educates on game theory, dismantling simplistic race narratives by citing corporate espionage realities and high proportions of Chinese national researchers at US labs.17:51–25:35 · Guest disagreement 3/10 Mutually Assured AI Malfunction and Deterrence Policy Guo asks Hendrycks to break down his Mutually Assured AI Malfunction framework and its nuclear parallels into concrete policy steps. Hendrycks details offensive cyber deterrence, CIA espionage cells, and licensing regimes analogous to fissile material tracking.25:36–30:37 · Guest disagreement 4/10 Compute Security, Algorithmic Efficiency, and Export Controls Guo presses Hendrycks on whether DeepSeek and compute-efficient pretraining undermine hardware-based export controls and compute tracking. Hendrycks argues DeepSeek actually validates the need for intent deterrence rather than capability restriction, while Guo challenges the assumption that rival powers would ever forego commercial capabilities.30:37–35:58 · Guest disagreement 2/10 The Frontier of AI Evals and Autonomous Agents Guo queries how evaluation methodology must evolve once models surpass human benchmarks. Hendrycks outlines the jagged frontier between high-level reasoning/proof verification and defective autonomous agent capabilities.1:24–4:48 · The hosts pushing back 4/10 Frontier Lab Constraints and Alignment Versus Safety Guo questions Hendrycks' dismissive stance toward frontier lab safety efforts and probes the semantic distinction between alignment and safety. Hendrycks reframes alignment as merely a small, obedient subset of safety that fails to resolve overarching structural and geopolitical competition.4:48–10:07 · The hosts pushing back 3/10 National Security Implications and Biosecurity Safeguards Guo cites specific venture portfolio investments across cybersecurity and biotech to probe the trade-off between safety and competitive benefits. Hendrycks bluntly dismisses the notion of an intractable trade-off, arguing that dangerous biosecurity capabilities can simply be gated behind enterprise sales tiers.10:08–17:50 · The hosts pushing back 3/10 Military AI Applications, Espionage, and Competition Realities Guo contributes technical domain knowledge regarding electronic warfare, autonomous systems, and radar communications in modern combat. Hendrycks educates on game theory, dismantling simplistic race narratives by citing corporate espionage realities and high proportions of Chinese national researchers at US labs.17:51–25:35 · The hosts pushing back 2/10 Mutually Assured AI Malfunction and Deterrence Policy Guo asks Hendrycks to break down his Mutually Assured AI Malfunction framework and its nuclear parallels into concrete policy steps. Hendrycks details offensive cyber deterrence, CIA espionage cells, and licensing regimes analogous to fissile material tracking.25:36–30:37 · The hosts pushing back 6/10 Compute Security, Algorithmic Efficiency, and Export Controls Guo presses Hendrycks on whether DeepSeek and compute-efficient pretraining undermine hardware-based export controls and compute tracking. Hendrycks argues DeepSeek actually validates the need for intent deterrence rather than capability restriction, while Guo challenges the assumption that rival powers would ever forego commercial capabilities.30:37–35:58 · The hosts pushing back 2/10 The Frontier of AI Evals and Autonomous Agents Guo queries how evaluation methodology must evolve once models surpass human benchmarks. Hendrycks outlines the jagged frontier between high-level reasoning/proof verification and defective autonomous agent capabilities.

speaking balance: gold is the hosts, purple is the guest (3 minute bins)

0:00 · the hosts 21% · guest 79%0:00 · the hosts 21% · guest 79%3:00 · the hosts 15.4% · guest 84.6%3:00 · the hosts 15.4% · guest 84.6%6:00 · the hosts 20.5% · guest 79.5%6:00 · the hosts 20.5% · guest 79.5%9:00 · the hosts 20.4% · guest 79.6%9:00 · the hosts 20.4% · guest 79.6%12:00 · the hosts 9.9% · guest 90.1%12:00 · the hosts 9.9% · guest 90.1%15:00 · the hosts 12.2% · guest 87.8%15:00 · the hosts 12.2% · guest 87.8%18:00 · the hosts 8.2% · guest 91.8%18:00 · the hosts 8.2% · guest 91.8%21:00 · the hosts 8.4% · guest 91.6%21:00 · the hosts 8.4% · guest 91.6%24:00 · the hosts 19.9% · guest 80.1%24:00 · the hosts 19.9% · guest 80.1%27:00 · the hosts 12% · guest 88%27:00 · the hosts 12% · guest 88%30:00 · the hosts 11.3% · guest 88.7%30:00 · the hosts 11.3% · guest 88.7%33:00 · the hosts 8.6% · guest 91.4%33:00 · the hosts 8.6% · guest 91.4%36:00 · the hosts 92.5% · guest 7.5%36:00 · the hosts 92.5% · guest 7.5%
Sharpest disagreement ▶ 7:28 Hendrycks rejects safety trade-off framing

Hendrycks forcefully dismisses Guo's premise regarding complex trade-offs between AI progress and safety, arguing that high-risk biology access is trivial to gate via enterprise sales.

Hardest push from the hosts ▶ 25:35 Guo challenges compute security with DeepSeek efficiency

Guo directly challenges Hendrycks' compute security strategy, arguing that recent breakthroughs by DeepSeek prove algorithmic efficiency outpaces hardware export controls.

Biggest teaching moment ▶ 13:52 Hendrycks on espionage and multinational AI talent reality

Hendrycks breaks down the geopolitical impracticality of isolating US frontier AI development, pointing out the reliance on Chinese nationals and vulnerability to zero-day espionage.

The host holds their own ▶ 11:31 Guo details electronic warfare dynamics in Ukraine

Guo demonstrates deep practical expertise in defense tech by explaining how battlefield AI depends fundamentally on radio frequency and radar communications systems.

the scores for every segment, with the reasoning behind each
ChapterTopicThe hosts as informed peerGuest teachingGuest disagreementThe hosts pushing backWhy
Frontier Lab Constraints and Alignment Versus Safety 4544 Guo questions Hendrycks' dismissive stance toward frontier lab safety efforts and probes the semantic distinction between alignment and safety. Hendrycks reframes alignment as merely a small, obedient subset of safety that fails to resolve overarching structural and geopolitical competition.
National Security Implications and Biosecurity Safeguards 6553 Guo cites specific venture portfolio investments across cybersecurity and biotech to probe the trade-off between safety and competitive benefits. Hendrycks bluntly dismisses the notion of an intractable trade-off, arguing that dangerous biosecurity capabilities can simply be gated behind enterprise sales tiers.
Military AI Applications, Espionage, and Competition Realities 6643 Guo contributes technical domain knowledge regarding electronic warfare, autonomous systems, and radar communications in modern combat. Hendrycks educates on game theory, dismantling simplistic race narratives by citing corporate espionage realities and high proportions of Chinese national researchers at US labs.
Mutually Assured AI Malfunction and Deterrence Policy 5632 Guo asks Hendrycks to break down his Mutually Assured AI Malfunction framework and its nuclear parallels into concrete policy steps. Hendrycks details offensive cyber deterrence, CIA espionage cells, and licensing regimes analogous to fissile material tracking.
Compute Security, Algorithmic Efficiency, and Export Controls 7546 Guo presses Hendrycks on whether DeepSeek and compute-efficient pretraining undermine hardware-based export controls and compute tracking. Hendrycks argues DeepSeek actually validates the need for intent deterrence rather than capability restriction, while Guo challenges the assumption that rival powers would ever forego commercial capabilities.
The Frontier of AI Evals and Autonomous Agents 5522 Guo queries how evaluation methodology must evolve once models surpass human benchmarks. Hendrycks outlines the jagged frontier between high-level reasoning/proof verification and defective autonomous agent capabilities.

Statements from this episode (15)

Opinion
Hendrycks: AI tail risks are systematically under-addressed
“Since it'd be such a big deal we'd need to make sure that we can think about it properly channel it in a productive direction, and take care of some sort of tail risks which will, are generally systematically under addressed.”
Dan Hendrycks Mar 5, 2025 ▶ 1:05
Insight
Hendrycks: AI alignment is only a subset of AI safety
“So I view the distinction between alignment and safety as alignment as being a sort of subset of safety. Obviously you want the value systems of the AIs to be in keeping with or compatible with say the US public for USAIs or for you as an individual, but that …”
Dan Hendrycks Mar 5, 2025 ▶ 3:46
Prediction Not checkable as stated
Hendrycks: US-China race will force rapid, high-risk military AI integration
“China can have AIs that are totally aligned with them. The U S can have AIs that are totally aligned with them. You still are going to have a strategic competition between the two. This is going to they're going to need to integrate it in their militaries. The…”
Dan Hendrycks Mar 5, 2025 ▶ 4:13
Opinion
Hendrycks: Current AI cannot enable devastating grid cyberattacks
“For cyber, I don't think AIs are that relevant for Being able to pull off a devastating cyber attack on the grid by a malicious actor currently.”
Dan Hendrycks Mar 5, 2025 ▶ 5:21
Assertion Not checkable as stated
Hendrycks: Reasoning models have reached expert-level virology capabilities
“The AIs are getting very good at STEM PhD level types of topics, and that includes virology. So I think that they are sort of rounding the corner on being able to provide expert level capabilities in terms of their knowledge of the literature, Or even helping …”
Dan Hendrycks Mar 5, 2025 ▶ 5:43
Prediction Not checkable as stated
Hendrycks: Extreme AI export controls make a Taiwan invasion more likely
“If you turn the pain dial all the way up for China in export controls and if AI chips are the currency of economic power in the future, then this increases the probability that they want to invade Taiwan.”
Dan Hendrycks Mar 5, 2025 ▶ 9:18
Insight
Hendrycks: Voluntary AI Pauses Without Enforcement Only Benefit Bad Actors
“If you do it voluntarily, you just make yourself less powerful and you let the worst actors get ahead of you. You could say, well, we'll try and try to sign a treaty. We will not assume that the treaty will be followed. Like that would be very imprudent. You w…”
Dan Hendrycks Mar 5, 2025 ▶ 13:02
Assertion Not checkable as stated
Hendrycks: Over 30% of Staff at Some Top AI Labs Are Chinese Nationals
“I mean, we have, you know, some places, you know, 30% plus of the employees at these top AI companies are like Chinese nationals.”
Dan Hendrycks Mar 5, 2025 ▶ 14:20
Assertion Not checkable as stated
Hendrycks: Adversaries Can Spy on Top AI Labs via Slack Zero-Day Exploits
“All they need to do is do a zero day on Slack. And then they can know what DeepMind is up to in very high fidelity and OpenAI and XAI and others.”
Dan Hendrycks Mar 5, 2025 ▶ 19:56
Prediction Open · timeframe Mar 2030
Hendrycks: Nations Will Race on AI Drones Without Coordinating Restraints
“No, like, super weapon type stuff, but more conventional type of warfare, like drones and things like that, I expect that they'll continue to race and Probably not, maybe not even coordinate on anything like that, but that's just how things would go.”
Dan Hendrycks Mar 5, 2025 ▶ 22:35
Insight
Hendrycks: AI Safety Is a Geopolitical Problem, Not Primarily a Technical One
“Safety isn't, as I've been I'm trying to reinforce not really that much of a technical problem. This is more of a complex geopolitical problem with technical aspects.”
Dan Hendrycks Mar 5, 2025 ▶ 25:07
Insight
Hendrycks: Deterrence can restrict superpower intent, but not AI capabilities
“So you can restrict their intent, which is what deterrence does, but I don't think you can reliably or robustly restrict their capabilities.”
Dan Hendrycks Mar 5, 2025 ▶ 26:33
Prediction Not checkable as stated
Hendrycks: China will steal model weights even if chip controls succeed
“Even so, I still think if you really tighten the export controls, you made it so that China can't get any of those chips at all, and this is your, one of your biggest priorities, they're just going to steal the weights anyway.”
Dan Hendrycks Mar 5, 2025 ▶ 27:43
Prediction Not checkable as stated
Hendrycks: Saturating Humanity's Last Exam benchmark will signal superhuman STEM AI
“When performance is near the ceiling, I think that'd basically be an indication that, like, you have something like a superhuman mathematician or a superhuman STEM scientist for, in many ways, for when they're, when closed-ended questions are very useful, such…”
Dan Hendrycks Mar 5, 2025 ▶ 32:09
Opinion
Hendrycks: AI Models Remain 'Extremely Defective' as Autonomous Agents
“This could possibly change overnight, but it's still near the floor. I think they're still extremely defective as agents.”
Dan Hendrycks Mar 5, 2025 ▶ 33:05
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.