Insight certainty 4/5 debate potential 3/5

Hendrycks: AI alignment is only a subset of AI safety

Dan Hendrycks · No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks · Mar 5, 2025 · at 3:46

Dan Hendrycks, director of the Center for AI Safety, defines how technical alignment differs from comprehensive safety and risk management.

0:00 / 0:21exact quote · 21.8s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“So I view the distinction between alignment and safety as alignment as being a sort of subset of safety. Obviously you want the value systems of the AIs to be in keeping with or compatible with say the US public for USAIs or for you as an individual, but that doesn't make it necessarily safe.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Dan Hendrycks

Prediction Not checkable as stated
Hendrycks: U.S. and China may launch preemptive cyberattacks on data centers
“I think that later on, it becomes so destabilizing that China just says, we're going to do something preemptive, like do a cyber attack on your data center. And the U.S. Might do that to China.”
Dan Hendrycks Oct 31, 2025 ▶ 7:24 No Priors Ep. 138 | The Best of 2025 (So Far) with Sarah Guo and Elad Gil
Prediction Not checkable as stated
Hendrycks: Extreme AI export controls make a Taiwan invasion more likely
“If you turn the pain dial all the way up for China in export controls and if AI chips are the currency of economic power in the future, then this increases the probability that they want to invade Taiwan.”
Dan Hendrycks Mar 5, 2025 ▶ 9:18 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
Insight
Hendrycks: Voluntary AI Pauses Without Enforcement Only Benefit Bad Actors
“If you do it voluntarily, you just make yourself less powerful and you let the worst actors get ahead of you. You could say, well, we'll try and try to sign a treaty. We will not assume that the treaty will be followed. Like that would be very imprudent. You w…”
Dan Hendrycks Mar 5, 2025 ▶ 13:02 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
Insight
Hendrycks: AI Safety Is a Geopolitical Problem, Not Primarily a Technical One
“Safety isn't, as I've been I'm trying to reinforce not really that much of a technical problem. This is more of a complex geopolitical problem with technical aspects.”
Dan Hendrycks Mar 5, 2025 ▶ 25:07 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
Prediction Not checkable as stated
Hendrycks: China will steal model weights even if chip controls succeed
“Even so, I still think if you really tighten the export controls, you made it so that China can't get any of those chips at all, and this is your, one of your biggest priorities, they're just going to steal the weights anyway.”
Dan Hendrycks Mar 5, 2025 ▶ 27:43 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
Prediction Not checkable as stated
Hendrycks: US-China race will force rapid, high-risk military AI integration
“China can have AIs that are totally aligned with them. The U S can have AIs that are totally aligned with them. You still are going to have a strategic competition between the two. This is going to they're going to need to integrate it in their militaries. The…”
Dan Hendrycks Mar 5, 2025 ▶ 4:13 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.