Nov 17, 2025 · 1h 7m · a16z
Emmett Shear on Building AI That Actually Cares: Beyond Control and Steering
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
In this episode of The a16z Show, Softmax CEO Emmett Shear presents a novel paradigm for artificial intelligence safety, advocating for 'organic alignment' and digital beings built with intrinsic care rather than rigid top-down control. Through in-depth discussions with co-host Seb Krier, Shear explores multi-agent reinforcement learning, formal criteria for sentience, and how theory of mind can lead to cooperative AI companions.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. The host holds 5.1% of the talking time here. How this is scored →
speaking balance: gold is the host, purple is the guest (3 minute bins)
Emmett directly confronts co-host Seb, declaring that holding a stance on AI personhood without specifying any empirical observation that could alter it is an article of faith rather than a scientific belief.
Hardest push from the host ▶ 28:46 Rejection of computational functionalismSeb explicitly rejects Emmett's core premise, stating he remains skeptical of computational functionalism and insisting that substrate differences make AIs fundamentally tools rather than moral beings.
Biggest teaching moment ▶ 43:40 Multi-tier homeostatic hierarchy lectureEmmett delivers a sophisticated theoretical breakdown based on Carl Friston's free energy principle, laying out six hierarchical layers of homeostatic dynamics required for pain, pleasure, metastates, and thought.
The host holds their own ▶ 9:01 Political science framework for value discoverySeb draws on political science concepts to challenge static alignment models, proposing liberal democratic systems as bottom-up mechanisms for discovering values over time.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | The host as informed peer | Guest teaching | Guest disagreement | The host pushing back | Why |
|---|---|---|---|---|---|---|
| Show Title Card and Studio Transition | 1 | 6 | 1 | 0 | Emmett delivers an extensive opening monologue defining alignment as a dynamic living process rather than a static state. He contrasts rigid rule-following with organic moral progress, while the hosts listen attentively with no pushback. | |
| Technical vs. Normative Alignment Frameworks | 5 | 4 | 2 | 3 | Co-host Seb introduces a political science framework comparing normative alignment to liberal democratic processes of value discovery. Emmett accepts part of the framing but reframes technical alignment around coherent goal-following. | |
| Goal Descriptions vs. Actual Goal Inference | 3 | 7 | 4 | 2 | Emmett corrects Seb for confusing a goal description with an actual goal, using an apple analogy and Stuart Russell's cleaning robot example. He insists goal inference requires internal theories of mind and world. | |
| Goal Coherence, Incompetence, and Principal-Agent Dynamics | 4 | 6 | 2 | 2 | Seb brings up principal-agent dynamics and situational incentives. Emmett breaks down goal failure modes into goal inference, goal prioritization, and execution using the peanut butter sandwich game and the OODA loop. | |
| 'Care' as the Pre-Conceptual Foundation of Alignment | 4 | 5 | 1 | 1 | Emmett introduces 'care' as the non-conceptual pre-condition for values, anchoring it in RL loss and free energy. Erik chimes in to ask how this perspective differentiates Softmax from major frontier labs. | |
| Steering and Control vs. Personhood and Citizenship | 6 | 4 | 4 | 7 | Emmett equates pure steering of general AI to slavery. Seb pushes back firmly, rejecting computational functionalism and arguing that substrate differences mean AIs remain tools regardless of capability. | |
| Debating Personhood Criteria and the Substrate | 3 | 8 | 8 | 5 | Emmett aggressively interrogates Seb on what empirical observations would change his mind on AI personhood. When Seb hesitates, Emmett asserts that Seb holds an article of faith rather than an empirical belief. | |
| Defining Behaviorism and the Internal 'Belief Manifold' | 4 | 6 | 5 | 4 | Seb argues that surface behavioral indistinguishability is insufficient without inspecting the inside. Emmett expands behaviorism to include inspecting the internal belief manifold for self-referential sub-manifolds. | |
| Sentience Criteria: Multi-Tier Homeostatic Dynamics | 4 | 8 | 6 | 5 | Emmett warns of catastrophic moral risk if hosts are wrong, while Seb pushes back on applying a one-sided precautionary principle. Emmett then details Carl Friston's free energy principle and a 6-tier homeostatic hierarchy for sentience. | |
| The Sorcerer's Apprentice: Dangerous Tools vs. Aligned Beings | 3 | 6 | 3 | 1 | Erik prompts Emmett on pragmatic alignment effectiveness. Emmett utilizes the Sorcerer's Apprentice parable and atomic bomb comparisons to argue that ultra-capable tools without internal care are inherently unsafe. | |
| Softmax Research Roadmap: Multi-Agent Reinforcement Learning | 2 | 6 | 1 | 0 | Erik asks for Softmax's technical roadmap. Emmett explains their multi-agent RL simulation strategy designed to build a surrogate model for social theory of mind, referencing the vampire pill parable. | |
| Redesigning Chatbot Personalities and the 'Pool of Narcissus' | 3 | 5 | 1 | 1 | Erik prompts Emmett on ideal chatbot behavior. Emmett criticizes 1-on-1 chatbots as Narcissus pools and proposes multi-user chatrooms, while humorously characterizing ChatGPT, Claude, and Gemini personalities. | |
| Multi-Agent Environment Dynamics: High Entropy and Regularization | 2 | 6 | 1 | 0 | In response to Erik's question on multi-agent LLM behavior, Emmett explains how additional agents introduce environmental entropy, requiring stronger regularization to prevent overfitting. | |
| Evaluating Eliezer Yudkowsky's AI Doom Thesis | 3 | 6 | 2 | 1 | Erik asks about Eliezer Yudkowsky's AI doom thesis. Emmett agrees with Yudkowsky's warning regarding controlled tools, but argues Yudkowsky prematurely dismisses organic alignment and civic AI co-existence. | |
| Reflections on OpenAI Tenure and Softmax's Purpose | 3 | 5 | 1 | 1 | Erik asks about Emmett's brief tenure as OpenAI CEO. Emmett clarifies why he chose to focus on Softmax's vision of organic alignment over OpenAI's tool-steering trajectory, envisioning digital companions and digital guard dogs. |