The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 10 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 0 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Assertion Not checkable as stated
Fredrikson: Gray Swan Found Jailbreaks in Every OpenClaw User Trajectory Tested
“So we just have a bunch of trajectories of actual people using OpenClaw. And tons and tons of different scenarios and just threw shade at it and like found breaks for each and every one of them, right?”
Matt Fredrikson Jun 22, 2026 ▶ 47:36 AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
Insight
Shah: OpenClaw memory fails because LLMs frequently skip search tool calls
“The way OpenClaw has with QMD or with whatever memory plugin you use, it inherently relies on tools to search through these memory.md files that it prepares. So, you know, like, what did I decide about the API? Then agent will decide to search, and sometimes i…”
Dhravya Shah Mar 9, 2026 ▶ 8:37 ⚡️ OpenClaw's Memory Sucks and the fix is simple — Dhravya Shah, Supermemory
Assertion Supported
Shah: OpenClaw's 15-message replay fails prompt caching and costs 10x more
“The way OpenClaw does it is it essentially sends back the last 15 messages in the conversation and it essentially uses that back and forth. And I mean, the approach itself is not ideal because you will, like, you are not doing any, like, you're not utilizing a…”
Dhravya Shah Mar 9, 2026 ▶ 21:03 ⚡️ OpenClaw's Memory Sucks and the fix is simple — Dhravya Shah, Supermemory
Insight
Krentsel: Improving AI models should architect their own agent harnesses over humans
“We believe All of that needs to be improvable by the agent, especially as the agents keep getting better, because as they keep getting better, it's this better lesson. You don't want to over-specialize because you don't want the human kind of deciding all thes…”
Alex Krentsel Aug 15, 2026 ▶ 9:55 Exo: Harnesses should see their own code and logs — Alex Krentsel, UC Berekeley / Google Research
Assertion Supported
Krentsel: OpenClaw, Pi, and Claude Code hardcode static agent policies
“These are all policy decisions that are static, that are defined for OpenClaw, or for Pi, or for Clawed code, if you look at their Source code. And so that is the kind of, that is the policy of what an agent is, the tools it can use, the skills it has, how it …”
Alex Krentsel Aug 15, 2026 ▶ 7:43 Exo: Harnesses should see their own code and logs — Alex Krentsel, UC Berekeley / Google Research
Assertion Supported
Cohen: OpenClaw logs all messages in plain text
“I started to see the size of the code base and the number of dependencies and some other things like logging all messages in plain text that just made me a bit apprehensive to use it for like production use cases to build a business on it.”
Gavriel Cohen Jun 29, 2026 ▶ 7:19 The Blueprint for Autonomous Work Agents | Gavriel Cohen, NanoClaw
Assertion Supported
Shah: Supermemory outperformed Claude Code and OpenClaw benchmarks by almost 50%
“So the Claude code one performed the worst, and OpenClaw slightly more than that, and SuperMemory is the highest, and you can see that, you know, it's like a pretty significant difference, like almost 50%.”
Dhravya Shah Mar 9, 2026 ▶ 14:33 ⚡️ OpenClaw's Memory Sucks and the fix is simple — Dhravya Shah, Supermemory
Assertion Partly supported
Krentsel: OpenClaw agent threads cannot be interrupted during active execution
“Once an agent kind of goes and starts working on a thing in a thread, that thread is not interruptible. As the thing is off working, if you try to ping it just won't respond. You don't even know what it's doing, which is really frustrating.”
Alex Krentsel Aug 15, 2026 ▶ 31:17 Exo: Harnesses should see their own code and logs — Alex Krentsel, UC Berekeley / Google Research
Prediction Not checkable as stated
Nathan: ChatGPT Work will not completely replace open-source OpenClaw
“I don't think so. I think that there's going to be, you know, there's always a need for, like, this, like, incredible, like, open source technology that, that team has built”
Akshay Nathan Jul 28, 2026 ▶ 47:49 OpenAI’s Vision for the AI Super App — Akshay Nathan, OpenAI
Disclosure
NVIDIA mandates running OpenClaw in isolated Brev cloud VMs
“Internally people want to run this and we know we have to be really careful from the security implications. Do we let this run on the corporate network securities guidance was, Hey, run this on breath. It's in, you know, it's a VM. It's sitting in the cloud. …”
Nader Khalil Mar 8, 2026 ▶ 9:08 Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.