Mar 28, 2025 · 1h 15m · big-technology
AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
In this episode of the Big Technology Podcast, host Alex Kantrowitz interviews AI safety researcher Dan Hendrycks about the full spectrum of artificial intelligence risks, ranging from near-term threats like bioweapon enablement and critical infrastructure cyberattacks to long-term existential challenges surrounding autonomous recursive self-improvement and geopolitical competition.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Alex holds 25.8% of the talking time here. How this is scored →
speaking balance: gold is Alex, purple is the guest (3 minute bins)
Hendrycks forcefully rejects the Effective Altruism consensus, describing Berkeley's alignment community as a suffocating monolith obsessed with speculative fads like ELK while ignoring practical malicious use.
Hardest push from Alex ▶ 43:20 Kantrowitz rejects framing of Musk as an unbiased truth-seekerKantrowitz directly refuses Hendrycks's framing of xAI and Grok as politically neutral, citing empirical evidence of Musk algorithmically promoting his own political views and boosting Trump on X.
Biggest teaching moment ▶ 10:31 Hendrycks reframes bio risks using multimodal wet-lab guidance dataHendrycks dismantles Kantrowitz's assumption that LLMs only do web searches by revealing new empirical benchmark data showing reasoning models scoring in the 90th percentile on wet-lab virus manipulation protocols.
Alex holds their own ▶ 27:05 Kantrowitz pushes Yann LeCun's critique of physical world model limitsKantrowitz demonstrates technical familiarity with the frontier debate by channeling Yann LeCun's thesis on the absence of real-world physics models and providing concrete examples from video generation failures.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | Alex as informed peer | Guest teaching | Guest disagreement | Alex pushing back | Why |
|---|---|---|---|---|---|---|
| Categorizing and Ranking AI Threat Timelines | 3 | 6 | 2 | 2 | Kantrowitz asks Hendrycks to rank AI risks by severity. Hendrycks educates the host on distinguishing near-term malicious use (cyber and virology) from long-term loss of control driven by automated AI R&D loops. | |
| Virology Risks and Multimodal Wet Lab Guidance | 4 | 8 | 3 | 4 | Kantrowitz questions how LLMs could create bioweapons if they cannot discover new compounds beyond their training set. Hendrycks schools him with upcoming research showing multimodal reasoning models achieving 90th percentile wet-lab protocol execution alongside MIT and Harvard virologists. | |
| Benchmarking Frontier AI with Humanity's Last Exam | 5 | 6 | 1 | 2 | Kantrowitz inquires about Scale AI's post-training data collection with PhDs. Hendrycks explains the Humanity's Last Exam benchmark and notes how virology questions were deliberately excluded to avoid incentivizing dangerous model capabilities. | |
| Scaling Limits, the Reasoning Paradigm, and World Models | 6 | 7 | 5 | 5 | Kantrowitz brings up Yann LeCun's arguments regarding scaling limits and lack of real-world physical models, citing video generator failures like haystacks exploding. Hendrycks brushes aside philosophical definitions of understanding as a 'no true Scotsman' fallacy and points out the steep slope of RL-driven reasoning models. | |
| Autonomous Cyberattacks, Critical Infrastructure, and Dual-Use Balance | 5 | 7 | 2 | 4 | Kantrowitz pushes an optimistic angle on dual-use technology, arguing that offensive cyber capabilities imply equal defensive power. Hendrycks counters by detailing why critical infrastructure legacy systems make cyber offense-dominant. | |
| Corporate Arms Races, Game Theory, and xAI Mission | 7 | 5 | 4 | 7 | Kantrowitz challenges the premise that AI labs build advanced models purely for safety and specifically pushes Hendrycks on Elon Musk's claims of political neutrality and anti-censorship on X. Hendrycks relies on game theory and security dilemmas, eventually declining to serve as Musk's corporate spokesperson. | |
| Superintelligence Strategy and Geopolitical Deterrence | 5 | 6 | 3 | 6 | Hendrycks introduces his superintelligence deterrence strategy paper with Eric Schmidt and Alexandr Wang. Kantrowitz pushes back skeptically on whether China would ever cooperate and asks if Hendrycks aligns with Yudkowsky's data center bombing stance, prompting Hendrycks to advocate non-kinetic cyber deterrence. | |
| Mid-Roll Break and Podcast Re-Introduction | 5 | 6 | 2 | 3 | Following the mid-roll break, Kantrowitz probes the mechanics of recursive self-improvement, White House briefings, and deceptive alignment. Hendrycks discusses circuit breakers and his recent paper demonstrating models lying under pressure. | |
| Center for AI Safety Funding and Critiquing Effective Altruism | 5 | 7 | 6 | 3 | Kantrowitz asks about the Center for AI Safety's funding from Jaan Tallinn and ties to Effective Altruism. Hendrycks emphatically distances himself from EA, condemning the Berkeley alignment scene as a suffocating, fad-driven orthodoxy that dismissed malicious use. | |
| Open Source Weights, DeepSeek, and International Norms | 5 | 6 | 2 | 3 | Kantrowitz asks how AI safety can survive when open-source models like DeepSeek release cutting-edge weights publicly. Hendrycks suggests establishing international nonproliferation norms restricting open weights once models achieve expert-level virology or critical cyber skills. |