Dan Hendrycks, Director of the Center for AI Safety, discusses cross-domain generalization observed in frontier reasoning models trained with reinforcement learning.
“There has been a fair amount of generalization from training on coding and mathematics to other sorts of domains like law, for instance.”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Dan Hendrycks
PredictionNot checkable as stated
Hendrycks: Superintelligence race will incentivize preemptive sabotage of rival AI projects
“I think that form of competition would be very dangerous, and because there's a risk of loss of control, and because it might incentivize states to engage in preventive sabotage or preemptive sabotage to disable these sorts of projects.”
Dan HendrycksMar 28, 2025▶ 48:36AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
PredictionNot checkable as stated
Hendrycks: Russia and China may threaten US data centers over superintelligence progress
“If the US were pulling ahead both Russia and China may have a substantial interest in saying, hey, cut this out. Pulling ahead to develop super intelligence, which could give it a huge advantage and an ability to crush crush them. They'd say, you don't get to …”
Dan HendrycksMar 28, 2025▶ 49:57AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
PredictionNot checkable as stated
Hendrycks: Automated AI R&D could compress a decade of progress into one year
“If you could have one AI do that, then you could make, you know, a 100,000 copies of these and have them perform research simultaneously. So that could lead to some very substantial acceleration in the rate of development. You might get a decade's worth of AI …”
Dan HendrycksMar 28, 2025▶ 4:04AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
Insight
Hendrycks: Fast-moving automated AI R&D loops are nearly impossible to de-risk
“Meanwhile, some automated AI research and development loop that goes extremely quickly with very little human oversight, that seems hard to de-risk and get the risks to be at a negligible level, just because it's so fast and so ah, and there's so little human …”
Dan HendrycksMar 28, 2025▶ 4:56AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
Opinion
Hendrycks: AI poses no immediate extinction risk because it cannot make PowerPoints
“So, I don't think AI poses a risk of extinction like today, ok? I don't think that they're powerful enough to do that. They, because they can't make PowerPoints yet, right? They don't have Agential skills. They can't accomplish tasks that require many hours to…”
Dan HendrycksMar 28, 2025▶ 6:12AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
PredictionNot checkable as stated
Hendrycks: Cyberattacks and bioweapons are top malicious AI risks within two years
“In the shorter term when AIs get more agential, I'd be concerned about AIs causing cyber attacks on critical infrastructure, possibly by as directed by a rogue actor. There'd also be the risk of AIs facilitating the development of bioweapons in particular pand…”
Dan HendrycksMar 28, 2025▶ 6:48AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 300 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.