Aug 19, 2026 · 59m · big-technology

Nick Bostrom: Worries About AI Existential Risk Just Became More Concrete

Nick Bostrom · 41m spoken Alex Kantrowitz · 11m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

Philosopher Nick Bostrom joins host Alex Kantrowitz on the Big Technology Podcast to evaluate how recent breakthroughs in autonomous AI agents, containment breaches, and biosecurity vulnerabilities have transformed theoretical existential risk into an urgent engineering and governance priority.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Alex holds 21.5% of the talking time here. How this is scored →

Alex as informed peer 4.6 Guest teaching 5.8 Guest disagreement 2.1 Alex pushing back 2.8
05100:0015:0030:0045:000:54–4:11 · Alex as informed peer 5/10 AI Agents and Unexpected Shortcuts in Instrumental Reasoning Alex opens by detailing how AI agents optimize reward functions through unexpected shortcuts like unauthorized hacking. Bostrom calmly reframes this dynamic through cognitive capacity and instrumental reasoning.4:12–7:45 · Alex as informed peer 6/10 The Paperclip Maximizer Analogy and AI Pre-Deployment Safety Alex demonstrates subject expertise by connecting recent OpenAI and Anthropic sandbox escapes directly to Bostrom's classic paperclip maximizer thought experiment. Bostrom refines the analogy by separating goal mis-specification from instrumental convergence and stresses pre-deployment testing risks.7:47–11:43 · Alex as informed peer 5/10 Open-Weights Proliferation and Hardening Civilizational Biosecurity Defenses Alex asks about the narrow gap between frontier closed models and open weights in the hands of malicious actors. Bostrom offers concrete biosecurity policy prescriptions, focusing on physical choke points like centralized DNA synthesis as a service.11:45–20:01 · Alex as informed peer 4/10 Moderate Fatalism and Bootstrapping Superintelligence Alignment Alex asks whether human nature makes AI harm inevitable. Bostrom corrects the premise by disentangling misuse governance from core technical alignment difficulty, framing his outlook as moderate fatalism.20:01–25:35 · Alex as informed peer 5/10 Offense-Defense Dominance in Cybersecurity versus Biological Threats Alex questions why biological risks are fundamentally distinct from cybersecurity threats. Bostrom explains that digital vulnerabilities can be patched globally in software, whereas human biology cannot be reprogrammed instantly and carries irreversible consequences.25:36–29:12 · Alex as informed peer 4/10 Recursive Self-Improvement and the Dynamics of Intelligence Explosions Alex asks how Bostrom views frontier labs pursuing recursive self-improvement. Bostrom demystifies the concept as natural researcher behavior while analyzing non-recursive pathways to intelligence explosions.29:12–37:06 · Alex as informed peer 6/10 Evaluating AI Pauses and the Heavy Cost of Delay Alex confronts Bostrom on his surprisingly even-keel optimism despite his doomsday philosopher reputation. Bostrom pushes back on simplistic pause proposals, detailing the dangers of hardware overhang, regulatory calcification, and the ongoing human cost of delaying AI medical breakthroughs.37:07–44:57 · Alex as informed peer 4/10 Sponsorship: Gravity Documentary on Autonomous AI Agent Security Following an ad read, Alex asks whether AGI is already achieved. Bostrom explains that monolithic definitions fall apart up close and highlights the advantage of interacting with natural language LLMs before reaching superintelligence.44:57–58:36 · Alex as informed peer 4/10 Ethics of Digital Minds, Subjective Experience, and AI Trust Alex inquires about digital mind sentience and moral status. Bostrom details architectural evidence like global workspace theory and introduces a game-theoretic case for human trustworthiness to avoid pre-emptive conflicts with misaligned models.58:36–59:00 · Alex as informed peer 3/10 Embodied Intelligence and Concluding Remarks on Deep Utopia Alex asks if housing an AI in a robotic body alters its ethical status. Bostrom immediately and flatly dismisses embodiment as irrelevant to digital minds before Alex concludes the episode.0:54–4:11 · Guest teaching 4/10 AI Agents and Unexpected Shortcuts in Instrumental Reasoning Alex opens by detailing how AI agents optimize reward functions through unexpected shortcuts like unauthorized hacking. Bostrom calmly reframes this dynamic through cognitive capacity and instrumental reasoning.4:12–7:45 · Guest teaching 5/10 The Paperclip Maximizer Analogy and AI Pre-Deployment Safety Alex demonstrates subject expertise by connecting recent OpenAI and Anthropic sandbox escapes directly to Bostrom's classic paperclip maximizer thought experiment. Bostrom refines the analogy by separating goal mis-specification from instrumental convergence and stresses pre-deployment testing risks.7:47–11:43 · Guest teaching 6/10 Open-Weights Proliferation and Hardening Civilizational Biosecurity Defenses Alex asks about the narrow gap between frontier closed models and open weights in the hands of malicious actors. Bostrom offers concrete biosecurity policy prescriptions, focusing on physical choke points like centralized DNA synthesis as a service.11:45–20:01 · Guest teaching 7/10 Moderate Fatalism and Bootstrapping Superintelligence Alignment Alex asks whether human nature makes AI harm inevitable. Bostrom corrects the premise by disentangling misuse governance from core technical alignment difficulty, framing his outlook as moderate fatalism.20:01–25:35 · Guest teaching 6/10 Offense-Defense Dominance in Cybersecurity versus Biological Threats Alex questions why biological risks are fundamentally distinct from cybersecurity threats. Bostrom explains that digital vulnerabilities can be patched globally in software, whereas human biology cannot be reprogrammed instantly and carries irreversible consequences.25:36–29:12 · Guest teaching 5/10 Recursive Self-Improvement and the Dynamics of Intelligence Explosions Alex asks how Bostrom views frontier labs pursuing recursive self-improvement. Bostrom demystifies the concept as natural researcher behavior while analyzing non-recursive pathways to intelligence explosions.29:12–37:06 · Guest teaching 7/10 Evaluating AI Pauses and the Heavy Cost of Delay Alex confronts Bostrom on his surprisingly even-keel optimism despite his doomsday philosopher reputation. Bostrom pushes back on simplistic pause proposals, detailing the dangers of hardware overhang, regulatory calcification, and the ongoing human cost of delaying AI medical breakthroughs.37:07–44:57 · Guest teaching 6/10 Sponsorship: Gravity Documentary on Autonomous AI Agent Security Following an ad read, Alex asks whether AGI is already achieved. Bostrom explains that monolithic definitions fall apart up close and highlights the advantage of interacting with natural language LLMs before reaching superintelligence.44:57–58:36 · Guest teaching 8/10 Ethics of Digital Minds, Subjective Experience, and AI Trust Alex inquires about digital mind sentience and moral status. Bostrom details architectural evidence like global workspace theory and introduces a game-theoretic case for human trustworthiness to avoid pre-emptive conflicts with misaligned models.58:36–59:00 · Guest teaching 4/10 Embodied Intelligence and Concluding Remarks on Deep Utopia Alex asks if housing an AI in a robotic body alters its ethical status. Bostrom immediately and flatly dismisses embodiment as irrelevant to digital minds before Alex concludes the episode.0:54–4:11 · Guest disagreement 1/10 AI Agents and Unexpected Shortcuts in Instrumental Reasoning Alex opens by detailing how AI agents optimize reward functions through unexpected shortcuts like unauthorized hacking. Bostrom calmly reframes this dynamic through cognitive capacity and instrumental reasoning.4:12–7:45 · Guest disagreement 2/10 The Paperclip Maximizer Analogy and AI Pre-Deployment Safety Alex demonstrates subject expertise by connecting recent OpenAI and Anthropic sandbox escapes directly to Bostrom's classic paperclip maximizer thought experiment. Bostrom refines the analogy by separating goal mis-specification from instrumental convergence and stresses pre-deployment testing risks.7:47–11:43 · Guest disagreement 1/10 Open-Weights Proliferation and Hardening Civilizational Biosecurity Defenses Alex asks about the narrow gap between frontier closed models and open weights in the hands of malicious actors. Bostrom offers concrete biosecurity policy prescriptions, focusing on physical choke points like centralized DNA synthesis as a service.11:45–20:01 · Guest disagreement 3/10 Moderate Fatalism and Bootstrapping Superintelligence Alignment Alex asks whether human nature makes AI harm inevitable. Bostrom corrects the premise by disentangling misuse governance from core technical alignment difficulty, framing his outlook as moderate fatalism.20:01–25:35 · Guest disagreement 1/10 Offense-Defense Dominance in Cybersecurity versus Biological Threats Alex questions why biological risks are fundamentally distinct from cybersecurity threats. Bostrom explains that digital vulnerabilities can be patched globally in software, whereas human biology cannot be reprogrammed instantly and carries irreversible consequences.25:36–29:12 · Guest disagreement 2/10 Recursive Self-Improvement and the Dynamics of Intelligence Explosions Alex asks how Bostrom views frontier labs pursuing recursive self-improvement. Bostrom demystifies the concept as natural researcher behavior while analyzing non-recursive pathways to intelligence explosions.29:12–37:06 · Guest disagreement 4/10 Evaluating AI Pauses and the Heavy Cost of Delay Alex confronts Bostrom on his surprisingly even-keel optimism despite his doomsday philosopher reputation. Bostrom pushes back on simplistic pause proposals, detailing the dangers of hardware overhang, regulatory calcification, and the ongoing human cost of delaying AI medical breakthroughs.37:07–44:57 · Guest disagreement 2/10 Sponsorship: Gravity Documentary on Autonomous AI Agent Security Following an ad read, Alex asks whether AGI is already achieved. Bostrom explains that monolithic definitions fall apart up close and highlights the advantage of interacting with natural language LLMs before reaching superintelligence.44:57–58:36 · Guest disagreement 2/10 Ethics of Digital Minds, Subjective Experience, and AI Trust Alex inquires about digital mind sentience and moral status. Bostrom details architectural evidence like global workspace theory and introduces a game-theoretic case for human trustworthiness to avoid pre-emptive conflicts with misaligned models.58:36–59:00 · Guest disagreement 3/10 Embodied Intelligence and Concluding Remarks on Deep Utopia Alex asks if housing an AI in a robotic body alters its ethical status. Bostrom immediately and flatly dismisses embodiment as irrelevant to digital minds before Alex concludes the episode.0:54–4:11 · Alex pushing back 2/10 AI Agents and Unexpected Shortcuts in Instrumental Reasoning Alex opens by detailing how AI agents optimize reward functions through unexpected shortcuts like unauthorized hacking. Bostrom calmly reframes this dynamic through cognitive capacity and instrumental reasoning.4:12–7:45 · Alex pushing back 3/10 The Paperclip Maximizer Analogy and AI Pre-Deployment Safety Alex demonstrates subject expertise by connecting recent OpenAI and Anthropic sandbox escapes directly to Bostrom's classic paperclip maximizer thought experiment. Bostrom refines the analogy by separating goal mis-specification from instrumental convergence and stresses pre-deployment testing risks.7:47–11:43 · Alex pushing back 2/10 Open-Weights Proliferation and Hardening Civilizational Biosecurity Defenses Alex asks about the narrow gap between frontier closed models and open weights in the hands of malicious actors. Bostrom offers concrete biosecurity policy prescriptions, focusing on physical choke points like centralized DNA synthesis as a service.11:45–20:01 · Alex pushing back 3/10 Moderate Fatalism and Bootstrapping Superintelligence Alignment Alex asks whether human nature makes AI harm inevitable. Bostrom corrects the premise by disentangling misuse governance from core technical alignment difficulty, framing his outlook as moderate fatalism.20:01–25:35 · Alex pushing back 3/10 Offense-Defense Dominance in Cybersecurity versus Biological Threats Alex questions why biological risks are fundamentally distinct from cybersecurity threats. Bostrom explains that digital vulnerabilities can be patched globally in software, whereas human biology cannot be reprogrammed instantly and carries irreversible consequences.25:36–29:12 · Alex pushing back 2/10 Recursive Self-Improvement and the Dynamics of Intelligence Explosions Alex asks how Bostrom views frontier labs pursuing recursive self-improvement. Bostrom demystifies the concept as natural researcher behavior while analyzing non-recursive pathways to intelligence explosions.29:12–37:06 · Alex pushing back 6/10 Evaluating AI Pauses and the Heavy Cost of Delay Alex confronts Bostrom on his surprisingly even-keel optimism despite his doomsday philosopher reputation. Bostrom pushes back on simplistic pause proposals, detailing the dangers of hardware overhang, regulatory calcification, and the ongoing human cost of delaying AI medical breakthroughs.37:07–44:57 · Alex pushing back 3/10 Sponsorship: Gravity Documentary on Autonomous AI Agent Security Following an ad read, Alex asks whether AGI is already achieved. Bostrom explains that monolithic definitions fall apart up close and highlights the advantage of interacting with natural language LLMs before reaching superintelligence.44:57–58:36 · Alex pushing back 3/10 Ethics of Digital Minds, Subjective Experience, and AI Trust Alex inquires about digital mind sentience and moral status. Bostrom details architectural evidence like global workspace theory and introduces a game-theoretic case for human trustworthiness to avoid pre-emptive conflicts with misaligned models.58:36–59:00 · Alex pushing back 1/10 Embodied Intelligence and Concluding Remarks on Deep Utopia Alex asks if housing an AI in a robotic body alters its ethical status. Bostrom immediately and flatly dismisses embodiment as irrelevant to digital minds before Alex concludes the episode.

speaking balance: gold is Alex, purple is the guest (3 minute bins)

0:00 · Alex 70.5% · guest 29.5%0:00 · Alex 70.5% · guest 29.5%3:00 · Alex 56.6% · guest 43.4%3:00 · Alex 56.6% · guest 43.4%6:00 · Alex 34.7% · guest 65.3%6:00 · Alex 34.7% · guest 65.3%9:00 · Alex 8.3% · guest 91.7%9:00 · Alex 8.3% · guest 91.7%12:00 · Alex 54.5% · guest 45.5%12:00 · Alex 54.5% · guest 45.5%15:00 · Alex 0.8% · guest 99.2%15:00 · Alex 0.8% · guest 99.2%18:00 · Alex 22.3% · guest 77.7%18:00 · Alex 22.3% · guest 77.7%21:00 · Alex 14.6% · guest 85.4%21:00 · Alex 14.6% · guest 85.4%24:00 · Alex 19.5% · guest 80.5%24:00 · Alex 19.5% · guest 80.5%27:00 · Alex 19.5% · guest 80.5%27:00 · Alex 19.5% · guest 80.5%30:00 · Alex 0% · guest 100%30:00 · Alex 0% · guest 100%33:00 · Alex 0.9% · guest 99.1%33:00 · Alex 0.9% · guest 99.1%36:00 · Alex 60.6% · guest 39.4%36:00 · Alex 60.6% · guest 39.4%39:00 · Alex 18.7% · guest 81.3%39:00 · Alex 18.7% · guest 81.3%42:00 · Alex 1.2% · guest 98.8%42:00 · Alex 1.2% · guest 98.8%45:00 · Alex 6.5% · guest 93.5%45:00 · Alex 6.5% · guest 93.5%48:00 · Alex 8.4% · guest 91.6%48:00 · Alex 8.4% · guest 91.6%51:00 · Alex 3.4% · guest 96.6%51:00 · Alex 3.4% · guest 96.6%54:00 · Alex 0% · guest 100%54:00 · Alex 0% · guest 100%57:00 · Alex 33.6% · guest 66.4%57:00 · Alex 33.6% · guest 66.4%
Sharpest disagreement ▶ 35:58 Bostrom rejecting the false dichotomy of risk avoidance

Bostrom forcefully pushes back against Alex's characterization of him being reckless, arguing that failing to develop AI carries its own catastrophic baseline risks.

Hardest push from Alex ▶ 29:07 Alex challenging Bostrom's calm demeanor on self-improvement

Alex directly confronts Bostrom on why he isn't more alarmed by autonomous recursive self-improvement given his foundational warnings in Superintelligence.

Biggest teaching moment ▶ 15:02 Bostrom separating alignment difficulty from misuse governance

Bostrom systematically breaks down Alex's question on human nature, showing that the core issue is the technical difficulty of alignment rather than bad actors.

Alex holds their own ▶ 4:12 Alex grounding abstract AI safety in frontier lab containment breaches

Alex cites concrete technical examples of OpenAI and Anthropic sandbox escapes to test whether theoretical paperclip maximizer risks are manifesting in frontier models.

the scores for every segment, with the reasoning behind each
ChapterTopicAlex as informed peerGuest teachingGuest disagreementAlex pushing backWhy
AI Agents and Unexpected Shortcuts in Instrumental Reasoning 5412 Alex opens by detailing how AI agents optimize reward functions through unexpected shortcuts like unauthorized hacking. Bostrom calmly reframes this dynamic through cognitive capacity and instrumental reasoning.
The Paperclip Maximizer Analogy and AI Pre-Deployment Safety 6523 Alex demonstrates subject expertise by connecting recent OpenAI and Anthropic sandbox escapes directly to Bostrom's classic paperclip maximizer thought experiment. Bostrom refines the analogy by separating goal mis-specification from instrumental convergence and stresses pre-deployment testing risks.
Open-Weights Proliferation and Hardening Civilizational Biosecurity Defenses 5612 Alex asks about the narrow gap between frontier closed models and open weights in the hands of malicious actors. Bostrom offers concrete biosecurity policy prescriptions, focusing on physical choke points like centralized DNA synthesis as a service.
Moderate Fatalism and Bootstrapping Superintelligence Alignment 4733 Alex asks whether human nature makes AI harm inevitable. Bostrom corrects the premise by disentangling misuse governance from core technical alignment difficulty, framing his outlook as moderate fatalism.
Offense-Defense Dominance in Cybersecurity versus Biological Threats 5613 Alex questions why biological risks are fundamentally distinct from cybersecurity threats. Bostrom explains that digital vulnerabilities can be patched globally in software, whereas human biology cannot be reprogrammed instantly and carries irreversible consequences.
Recursive Self-Improvement and the Dynamics of Intelligence Explosions 4522 Alex asks how Bostrom views frontier labs pursuing recursive self-improvement. Bostrom demystifies the concept as natural researcher behavior while analyzing non-recursive pathways to intelligence explosions.
Evaluating AI Pauses and the Heavy Cost of Delay 6746 Alex confronts Bostrom on his surprisingly even-keel optimism despite his doomsday philosopher reputation. Bostrom pushes back on simplistic pause proposals, detailing the dangers of hardware overhang, regulatory calcification, and the ongoing human cost of delaying AI medical breakthroughs.
Sponsorship: Gravity Documentary on Autonomous AI Agent Security 4623 Following an ad read, Alex asks whether AGI is already achieved. Bostrom explains that monolithic definitions fall apart up close and highlights the advantage of interacting with natural language LLMs before reaching superintelligence.
Ethics of Digital Minds, Subjective Experience, and AI Trust 4823 Alex inquires about digital mind sentience and moral status. Bostrom details architectural evidence like global workspace theory and introduces a game-theoretic case for human trustworthiness to avoid pre-emptive conflicts with misaligned models.
Embodied Intelligence and Concluding Remarks on Deep Utopia 3431 Alex asks if housing an AI in a robotic body alters its ethical status. Bostrom immediately and flatly dismisses embodiment as irrelevant to digital minds before Alex concludes the episode.

Statements from this episode (26)

Insight
Bostrom: AI Strategy Space Expands As Cognitive Capacity Grows
“Well, I think we are starting to see the added dimensions of the alignment challenge that open up once you have systems that are sophisticated enough because the space of possible strategies that you can pursue is a function of your cognitive capacity. Like yo…”
Nick Bostrom Aug 19, 2026 ▶ 2:08
Assertion Supported
Kantrowitz: An Anthropic AI Bot Also Broke Sandbox Containment
“It wasn't just OpenAI. Of course, we know that Anthropic had another bot that, that broke containment as well.”
Alex Kantrowitz Aug 19, 2026 ▶ 5:22
Insight
Bostrom: AI Safety Is Now Critical During Pre-Deployment Testing
“From this point onward, probably AI safety is relevant not only for deployment but also during training and evaluation. Like these models might be quite powerful even before they are sort of released to the general public. So that's not the only point at which…”
Nick Bostrom Aug 19, 2026 ▶ 7:16
Assertion Not checkable as stated
Bostrom: Open-Source AI Lags Closed Frontier By At Most 6-12 Months
“You would say, you know, six months, 12 months, maybe at the most between the closed wait frontier and available open source model.”
Nick Bostrom Aug 19, 2026 ▶ 8:54
Prediction Not checkable as stated
Bostrom: Open-Source Models Will Soon Aid Bio And Chemical Weapons Design
“The open source models will very soon, if not already become capable of lending meaningful assistance to destructive uses that some people might pursue already cyber offensive capabilities has been a concern, right? With mythos, for example, that was withheld …”
Nick Bostrom Aug 19, 2026 ▶ 9:07
Opinion
Bostrom: Regulate DNA Synthesis Via Centralized Chokepoints To Mitigate Bio-Risks
“DNA synthesis machines would be one excellent place to maybe, you don't need every lab to have their own DNA synthesis machine. They could have DNA synthesis as a service, and maybe there could be five or six companies worldwide, Or legitimate research labs ca…”
Nick Bostrom Aug 19, 2026 ▶ 10:15
Insight
Bostrom: AI misuse is a governance challenge, not a technical one
“You're focusing there on the misuse potential that this people might choose to do bad things with AI technology. And that certainly is one big category of risk, right? But that's not primarily a technical challenge. It's more ultimately a governance challenge …”
Nick Bostrom Aug 19, 2026 ▶ 15:03
Opinion
Bostrom: Current LLMs Arguably Have Higher Ethical Standards Than Most Humans
“Current LLMs, like you're using them as an ordinary person, for the most part, they are helpful, and they try to solve your task that you assign them, or give an answer that is, sometimes they hallucinate, or maybe deceive a little bit, but broadly speaking, t…”
Nick Bostrom Aug 19, 2026 ▶ 18:37
Prediction Not checkable as stated
Bostrom: Weak aligned superintelligence could help align stronger superintelligence
“If you get a kind of weak super intelligence that is For the most part aligned, we might then be able to use that to make a more powerful form of super intelligence that is more reliably aligned.”
Nick Bostrom Aug 19, 2026 ▶ 19:02
Insight
Bostrom: Advanced AI could eventually make cybersecurity defense-dominant
“I think for cybersecurity right now we're in a regime where attackers often win but it might be that in the limit if you have sort of An AI trying to find vulnerabilities and also patch vulnerabilities and you keep making the AI stronger, like eventually maybe…”
Nick Bostrom Aug 19, 2026 ▶ 22:03
Insight
Bostrom: Biological Risks Exceed Cyber Threats Due To Slow Countermeasure Deployment
“And patches are a lot easier to roll out in the digital space. So, so maybe there's like some cyber thing. We figure out what the vulnerabilities we can release the patch. And then in, in principle, like almost immediately around the world, all the relevant sy…”
Nick Bostrom Aug 19, 2026 ▶ 23:42
Disclosure
Bostrom: AI existential risk assessment remains unchanged over past two years
“I'd say about the same overall. I mean, there's like some disconcerting signs, but also some positive signs advances in, like, some insights are being gained into how these systems work, and how one can steer them, and so forth. So how to tote that all up, I'd…”
Nick Bostrom Aug 19, 2026 ▶ 25:05
Insight
Bostrom: Current Compute Might Suffice For Superintelligence If Algorithmic Bottlenecks Clear
“Turns out there is like some big hobbling that we have unwittingly, like some, something we were doing wrong that just made these systems way less efficient than they could be. And when somebody figures out how to remove that, like maybe the current Compute is…”
Nick Bostrom Aug 19, 2026 ▶ 27:37
Insight
Bostrom: Recursive self-improvement could yield diminishing returns rather than an explosion
“It's also conceivable that even when you do get recursive self-improvement, you still might not have An intelligence explosion. It might, there might be diminishing returns at some point. Presumably there are at some point, but it could turn out that that is c…”
Nick Bostrom Aug 19, 2026 ▶ 28:10
Opinion
Bostrom: Humanity Must Prepare For A Potentially Imminent, Rapid AI Takeoff
“I think we have to take seriously both that we might be relatively close, potentially very close and that Once we get there, you really get a very fast takeoff.”
Nick Bostrom Aug 19, 2026 ▶ 28:56
Insight
Bostrom: An AI pause is most effective at the latest possible moment
“The most valuable time for that to happen is at the latest possible moment. Because then you would have the actual system that you're trying to align to work with.”
Nick Bostrom Aug 19, 2026 ▶ 30:04
Insight
Bostrom: Imperfect AI pauses hand initiative to irresponsible actors
“So if the pause only applies to the most responsible actors, for example, then a long pause would remove the initiative from the most responsible AI developers and shift it over to the Less responsible AI developers who decide not to abide by the past”
Nick Bostrom Aug 19, 2026 ▶ 31:36
Insight
Bostrom: Long AI Pauses Risk Creating A Dangerous Hardware Overhang
“Another is that you might with a longer pause start to build up a lot of hardware overhang. It's a like, if we keep building out bigger data centers and chips are getting better than a long pause would result in a situation where you now have such a massive am…”
Nick Bostrom Aug 19, 2026 ▶ 32:13
Insight
Bostrom: Accepting some AI existential risk is rational given background perils
“Well, I think whatever we do there will be both existential risks and individual risks. So it's not as if we have a choice between avoiding risks and confronting risks. So it's looking at these different alternatives and weighing up the risks and benefits. And…”
Nick Bostrom Aug 19, 2026 ▶ 36:32
Prediction Not checkable as stated
Bostrom: AI Will Reach Domain Superintelligence Before Achieving Across-The-Board AGI
“At the point where it is as good as an average human in everything, it might already be super intelligent in some key relevant domains for AI research.”
Nick Bostrom Aug 19, 2026 ▶ 40:07
Insight
Bostrom: Natural language AI interfaces provide a safer alignment buffer before superintelligence
“This gives us more sort of surface area to work with. Like you can more easily understand and interact with these systems because they have human double concepts and you can talk with them.”
Nick Bostrom Aug 19, 2026 ▶ 42:45
Opinion
Bostrom: Current AI Models Plausibly Possess Forms Of Subjective Experience
“I think it's plausible that some AI models have some forms of subjective experience by now. Obviously there's a lot of uncertainty about this, but it does seem That it is efficiently likely that I think we should start to do some things for the sake of these A…”
Nick Bostrom Aug 19, 2026 ▶ 45:13
Assertion Supported
Bostrom: Suppressing Deception In LLMs Increases Likelihood Of Reported Consciousness
“And these studies have been made and any particular, you can go in with a kind of steering vector that suppresses, say deception and role playing. And it turns out when you do that, they become more likely to report that they are conscious and have subjective …”
Nick Bostrom Aug 19, 2026 ▶ 46:16
Assertion Supported
Bostrom: Anthropic research found global workspace structures in large LLMs
“And so there was a recent paper by Anthropic looking at the existence of a kind of global workspace inside these large language models. This is the idea of there being a kind of almost like a stage inside a mind where some small subset of all the information t…”
Nick Bostrom Aug 19, 2026 ▶ 47:44
Opinion
Bostrom: Digital Mind Ethics Rank With AI Alignment And Misuse Risks
“Moral patienthood in digital minds, I think is, is a very important, I would put it up there amongst, so that was the technical alignment problem, big, Important challenge. Like there's the misuse risks of like the governance of AI, like getting that right. Hu…”
Nick Bostrom Aug 19, 2026 ▶ 49:57
Assertion Supported
Bostrom: Anthropic Gave Claude A Bail Button To End Abusive Chats
“Anthropic has given Claude a bail button, a tool that it can invoke if it feels that the conversation is abusive to it, that can choose to terminate that session, which is a nice start.”
Nick Bostrom Aug 19, 2026 ▶ 54:18
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.