Mar 9, 2026 · 27m · latent-space
⚡️ OpenClaw's Memory Sucks and the fix is simple — Dhravya Shah, Supermemory
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
Supermemory founder Dhravya Shah explains why naive RAG and file-based agent memory fail in modern AI systems, presenting dynamic knowledge graphs, deterministic context hooks, and open benchmarks as the foundation for scalable context infrastructure.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. How this is scored →
speaking balance: gold is the hosts, purple is the guest (3 minute bins)
Dhravya aggressively dismisses the Locomo benchmark as outdated and flawed, arguing you can easily game 100% scores simply by dumping raw context into newer models.
Hardest push from the hosts ▶ 14:47 Benchmark reproduction challengeAlessio interrupts the guest's celebratory benchmark results to insist on methodological clarity, forcing Dhravya to concede that the benchmark evaluated Supermemory's custom re-implementation rather than native Claude code.
Biggest teaching moment ▶ 16:43 Explaining extraction cost fallaciesWhen Alessio suggests that maximizing memory extraction to win benchmarks does not seem problematic, Dhravya breaks down the prohibitive real-world token extraction costs and lack of forgetfulness testing.
The host holds their own ▶ 11:11 Local vs cloud memory trade-off critiqueAlessio demonstrates technical depth by contrasting Toby Lutke's local-first QMD approach with Cloudflare architectures, illustrating the real-world flaw of relying on models to voluntarily trigger search tools.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | The hosts as informed peer | Guest teaching | Guest disagreement | The hosts pushing back | Why |
|---|---|---|---|---|---|---|
| From Buildspace Side Project to Viral Open Source Architecture | 4 | 2 | 1 | 1 | Host kicks off the interview with familiar rapport, joking about Buildspace's founder and noting he starred the Supermemory repository when it first went viral. | |
| Re-Architecting AI Memory Beyond Naive RAG and Triplets | 6 | 5 | 2 | 1 | Guest outlines the evolution from naive vector RAG to knowledge graphs, while the host demonstrates domain familiarity by instantly completing definitions around triplet entities and graph traversal trade-offs. | |
| Supermemory Today: Context Infrastructure and Company Profile | 3 | 2 | 1 | 1 | Standard founder profile segment where Dhravya explains Supermemory's VC funding, transition from consumer tool to context infrastructure API, and plugin traction. | |
| Diagnosing OpenClaw's Memory Flaws and the Hooks Solution | 6 | 5 | 3 | 3 | Guest diagnoses why tool-based search fails in OpenClaw compared to automated hooks. Host contributes sharp technical context by contrasting Toby Lutke's local QMD philosophy with Cloudflare cloud architectures and tool-calling failure modes in models like Grok. | |
| File System Memory, Token Incentives, and MemoryBench | 6 | 5 | 4 | 6 | Guest presents a hot take that AI coding assistants lack token reduction incentives and presents a benchmark favoring Supermemory. The host pushes back directly, forcing the guest to admit the benchmark tested their own reproduction rather than native tools. | |
| Critiquing AI Memory Benchmarks and the User Profile Dilemma | 5 | 7 | 4 | 4 | Guest deconstructs popular AI memory benchmarks, strongly criticizing Locomo and explaining why extracting everything on LongMemEval creates unrealistic production costs. When host doubts this flaw, guest educates him on extraction pricing and episodic user profiling. | |
| OpenClaw Edge Cases, Token Economics, and Hybrid Retrieval | 6 | 5 | 3 | 5 | Host presses on whether real developers actually care about memory token costs. Guest pushes back with real-world customer economic realities and introduces Supermemory's hybrid RAG fallback architecture. | |
| Supermemory Future Roadmap, Voice Agents, and Industry Outlook | 5 | 2 | 1 | 1 | Wrap-up segment covering voice agents and future roadmaps. Host lists several competing memory architectures and invites the guest to deliver an upcoming technical workshop. |