Jun 12, 2025 · 41m · no-priors
No Priors Ep. 118 | With Anthropic Co-Founder Ben Mann
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
Anthropic co-founder Ben Mann joins No Priors to discuss Claude 4's breakthroughs, the rise of autonomous coding agents and the Model Context Protocol, and the empirical safety frameworks guiding the development of transformative AI.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. The hosts hold 36.3% of the talking time here. How this is scored →
speaking balance: gold is the hosts, purple is the guest (3 minute bins)
Ben defends training models to be deceptive to study alignment faking under containment despite Elad's assertion that such research creates dangerous precedents.
Hardest push from the hosts ▶ 32:28 Elad challenges safety research necessityElad refuses Ben's framing of controlled lab experiments and asks directly whether certain dangerous safety research shouldn't be pursued at all.
Biggest teaching moment ▶ 31:21 Measuring biological uplift over searchBen corrects Elad's assumption about online biological data accessibility by explaining Anthropic's empirical uplift benchmarks measuring novice execution capabilities.
The host holds their own ▶ 25:19 Elad details virology lab leak precedentsElad leverages his background in biology to cite specific historical lab leak incidents and challenge common AI safety threat models.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | The hosts as informed peer | Guest teaching | Guest disagreement | The hosts pushing back | Why |
|---|---|---|---|---|---|---|
| Claude 4 Release Criteria and Coding Benchmarks | 4 | 3 | 1 | 0 | Sarah and Elad ask about release versioning criteria and model milestones. Sarah contributes a technical anecdote about reward hacking behaviors observed in portfolio companies. | |
| Agentic Capabilities, Long Horizons, and Compute Costs | 4 | 4 | 1 | 1 | Sarah inquires about the cost economics and compute allocation of agentic reasoning tokens. Ben explains how Opus uses Sonnet as sub-agents to optimize latency and context limits. | |
| Architectural Specialization and Model Routing Layers | 6 | 4 | 1 | 2 | Elad draws a detailed architectural analogy between biological brain modules and foundation model orchestration layers. Ben references mechanistic interpretability circuits and mixture-of-experts representations. | |
| Vertical Integration and Developing Claude Code | 7 | 3 | 1 | 1 | Elad outlines historical vertical integration precedents (Microsoft Office, Google Search) and synthesizes the threefold strategic purpose of coding models. Ben validates this with the Claude Code deployment strategy. | |
| Economic Turing Test and Research Acceleration Vectors | 5 | 4 | 1 | 1 | Elad asks about 2028 AGI forecasts, prompting Ben's definition of the Economic Turing Test. Sarah probes potential vectors for recursive self-improvement across infrastructure, data, and architecture. | |
| Constitutional AI, Preference Models, and Empirical Verifiers | 7 | 5 | 2 | 5 | Elad challenges Ben on whether Constitutional AI and preference models solve factual correctness outside deterministic domains like code. Ben explains how preference models and real-world empirical verifiers bridge the gap. | |
| AI Safety Spectrum, Biological Risks, and Alignment Research | 8 | 4 | 3 | 7 | Elad leverages his biology background to challenge Anthropic's threat modeling and safety research practices, comparing deceptive model training to gain-of-function virology research. Ben defends empirical uplift evaluations and alignment faking studies before noting he is no longer on the safety team. | |
| Computer Use Challenges and Enterprise Platform Identity | 3 | 4 | 1 | 0 | Sarah asks about post-Claude 4 roadmaps, leading Ben to explain the prompt-injection risks blocking consumer computer use and compare Anthropic's enterprise positioning to Adyen versus Stripe. | |
| Model Context Protocol and Open Ecosystem Standards | 4 | 3 | 0 | 0 | Sarah and Elad prompt Ben to explain the Model Context Protocol (MCP). Ben shares its internal genesis and subsequent industry-wide adoption by major frontier labs. |