The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 93 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 0 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Assertion Contradicted
Hotz: Nobody has successfully trained models in INT8
“No one's gotten training to work with Indate yet. There's a few papers that vaguely show it, but if you're training, you're going to need BF-sixteen or float-sixteen.”
George Hotz Jun 20, 2023 ▶ 34:48 Ep 18: Petaflops to the People — with George Hotz of tinycorp
Assertion Contradicted
Anandkumar: Existing video and vision world models incorrectly assume fixed resolutions
“That immediately distinguishes us from other so-called world models, whether it's video models, vision models, they all assume during training and inference, it's a fixed resolution.”
Anima Anandkumar Sep 4, 2026 ▶ 10:15 Faster Chips That Don't Melt — Anima Anandkumar & Benedikt Jenik, Accelerated Understanding
Assertion Contradicted
All modern AI models requiring multi-GPU parallelization are Mixture-of-Experts
“Effectively, all models today are MOE models that are, you know, at least all models large enough that you would care to parallelize them across multiple GPUs.”
Philip Kiely Aug 3, 2026 ▶ 52:42 Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
Assertion Contradicted
He: Grok Imagine 0.9 was first large-scale joint audio-video model deployed
“So Grok Imagine, there were .9, I believe it's is a first first audio video trends model deployed at a large scale.”
Ethan He Jun 1, 2026 ▶ 42:45 Inside xAI: Building Grok Imagine in 3 Months, Videogen vs World Models, and Video Agents— Ethan He
Assertion Contradicted
Sanseviero: 31B is the largest quantized model fitting consumer GPUs
“The 31 is really like the largest model size that quantize would fit in a consumer GPU.”
Omar Sanseviero May 24, 2026 ▶ 17:17 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
Assertion Contradicted
No commercial products augmented GPT models with custom data pre-ChatGPT
“Like I saw some people doing demos, but like in like a CLI or something like that, but there was no product doing like this model, but with additional data on top of it.”
Yasser Elsaid May 2, 2026 ▶ 2:59 ⚡️ Competing with ChatGPT and Sierra, building a $10M ARR company — Yasser Elsaid, Founder, Chatbase
Assertion Contradicted
Welling: Keeping warming under 2°C requires century-long atmospheric carbon removal
“In order to get, you know, to stay within two degrees, let's say, we would not only have to reduce our emissions to zero by 2050, but then, you know, another half century or even a century, Of removing carbon dioxide from the atmosphere, not by reducing your e…”
Max Welling Feb 25, 2026 ▶ 14:48 🔬Max Welling: Materials Underlie Everything
Assertion Contradicted
O'Laughlin: Running Kimi agent swarms requires 16 Nvidia H100 nodes
“To just run the swarm, I think it's like a 16 node of H-one hundreds.”
Doug O'Laughlin Feb 24, 2026 ▶ 35:31 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Prediction Didn’t hold up
Swyx: OpenAI will always release both general and Codex model variants
“I'm pretty, like, have pretty high confidence that basically OpenAI will always release a GPT-V and a GPT-V codex.”
Shawn Wang Feb 19, 2026 ▶ 36:15 Inside AI’s $10B+ Capital Flywheel — Martin Casado & Sarah Wang of a16z
Assertion Contradicted
Yegge: Developers need 2,000 hours with AI before trusting it
“Jean just pulled up a study that showed that you actually have to spend a year or 2000 hours with AI before you trust it. And what does trust mean? Trust in this case specifically means before you as a user can predict what it's going to do.”
Steve Yegge Dec 26, 2025 ▶ 4:30 Steve Yegge's Vibe Coding Manifesto: Why Claude Code Isn't It & What Comes After the IDE
Assertion Contradicted
Wagner claims Flux shipped embedded AI chat before GPT-4 released
“So I think I'm going to claim here, I think we were the first engineering tool or design tool that had an AI chat in it. We shipped that I think a month or two months before GPT-IV became publicly available.”
Matthias Wagner Nov 22, 2025 ▶ 4:18 ⚡️ Building the AI Hardware Engineer with Matthias Wagner, Co-founder of Flux
Assertion Contradicted
Merrill: 60% to 70% of SWE-bench Verified tasks come from Django
“If you go look at sweet bench verified, I think like 60, 70% of the tasks in there are from Django.”
Mike Merrill Oct 18, 2025 ▶ 19:06 Terminal-Bench: Pushing Claude Code, OpenAI Codex, Factory Droid, et al to the limits
Assertion Contradicted
Lenz: AI21's Jamba is the first hybrid model architecture
“Since then, we've released several models, recent model lines in called Jamba, which I think the fascinating part about it is, is the first hybrid model. It's not just attention.”
Barak Lenz Oct 11, 2025 ▶ 2:20 Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21
Assertion Contradicted
Sohmers: Cray-2 was the last major system with balanced memory-to-compute ratio
“And if you look at sort of traditional big iron compute systems, the last like major compute, compute platform that had that balance of memory to compute ratio was the Cray two supercomputer.”
Thomas Sohmers Aug 18, 2025 ▶ 10:39 ⚡️Accelerators @ 3x NVIDIA H200 perf, Made in the USA - Thomas Sohmers + Mitesh Agrawal, Positron AI
Assertion Contradicted
Swix: Copilot report estimates 60-70% of AI-generated code is checked in
“There's a report this morning from Copilot where they were estimating the key tabs on amount of code generated by a Copilot that is then left in code repos and checked in. And it's something like 60 to 70%.”
Shawn Wang Jul 28, 2025 ▶ 12:17 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Prediction Didn’t hold up
Mohan: Automated PR generation will require specialized models trained on diffs
“A lot of things people are excited about right now are I write a comment and it generates a PR for me. And that's like really awesome in theory. I think that's like a really cool thing. And I'm sure at some point we will be able to get there. That will probabl…”
Varun Mohan Jul 28, 2025 ▶ 20:20 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Assertion Contradicted
Mohan: Over 80% of software developers are on Windows
“A lot of people, once again, over 80% of developers are on Windows.”
Varun Mohan Jul 28, 2025 ▶ 2:15:07 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Assertion Contradicted
The Lean theorem proving language has only about one million training tokens
“For Lean there are not enough data for Lean, right? There are, like, probably one million tokens in about Lean. Right now, and I know a lot of, like, professors and PhD researchers are trying to build the biggest lean data set in the future.”
Dr. Jasper Zhang Jul 24, 2025 ▶ 24:50 ⚡️Math Olympiad gold medalist explains OpenAI and Google DeepMind IMO Gold Performances
Assertion Contradicted
Morris: Top AI graduate programs do not teach multi-node distributed training
“Oh, to be clear, they don't teach you anything, like anything, like if you see a paper coming out from even, you know, Stanford, they're probably the best school in AI if you had to choose. And it's not like they're learning how to do like multi-node distribut…”
Jack Morris Jul 2, 2025 ▶ 9:38 Information Theory for Language Models: Jack Morris
Assertion Contradicted
Mohan: Over 80% of software developers are on Windows
“A lot of people, once again, over 80% of developers are on Windows.”
Varun Mohan Dec 13, 2024 ▶ 38:08 Windsurf: The Enterprise AI IDE
Assertion Contradicted
Ramachandran: Codeium is the only AI code assistant supporting Eclipse
“Like, we're still the only code assistant that has an extension of Eclipse. That's still true years in, right?”
Anshul Ramachandran Dec 13, 2024 ▶ 39:43 Windsurf: The Enterprise AI IDE
Assertion Contradicted
Vibhu: Synthetic data research shows LLMs verify better than they generate
“A lot of the synthetic datagen papers, like orca-three, wizard-lm, they show that models are better at verifying output than generating output.”
Vibhu (Veebu) Nov 29, 2024 ▶ 27:14 [Paper Club] DocETL: Agentic Query Rewriting + Eval for Complex Document Processing w Shreya Shankar
Assertion Contradicted
Pullen: Most scraped open-source data consists of README and documentation updates
“When you scrape enough of it, most of open source is updating readmes and docs.”
Alistair Pullen Aug 22, 2024 ▶ 19:17 Is finetuning GPT4o worth it?
Prediction Didn’t hold up
Cheah: Cloud providers will slash model inference prices before raising them
“One thing to warn about pricing is that you're going to see a lot of providers jumping in, and everyone's just trying to get the piece of the pie. So, so, so like with some of the previous model launches, you see some people coming in at lower and lower price,…”
Eugene Cheah Jul 29, 2024 ▶ 15:06 [LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models
Assertion Contradicted
Shulman: Bark was the first open-source transformer-based TTS model
“As far as I know there was no other certainly not in the open source text to speech that was kind of transformer based.”
Mikey Shulman Mar 14, 2024 ▶ 15:48 Making Transformers Sing - with Mikey Shulman of Suno
Assertion Contradicted
Doshi Claims Playground AI Was First to Ship LCM Preview Feature
“I think we were the first company to integrate it... I think we were the first company to actually ship a quick LCM thing.”
Suhail Doshi Jan 2, 2024 ▶ 54:53 The AI-First Graphics Editor - with Suhail Doshi of Playground AI
Assertion Contradicted
Hotz: Consumer AMD GPUs lack peer-to-peer support
“If you have a consumer AMD GPU, they don't support peer to peer.”
George Hotz Jun 20, 2023 ▶ 27:01 Ep 18: Petaflops to the People — with George Hotz of tinycorp
Assertion Contradicted
Midha: Stanford Holds America's Second-Largest Longitudinal Patient Dataset Behind the VA
“Stanford is one of the only research facilities in America that has a longitudinal patient data set that's Larger at scale, I think it's at least twelve million patient lives. The only larger data set is the VA, the Veterans Affairs, you know, of America.”
Anjney Midha Jun 18, 2026 ▶ 17:27 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Assertion Contradicted
Haneke: ASIC verification takes 3 to 4 times more resources than design
“My understanding is that the industry standard for design to verification in ASIC ASIC project is like one to three, one to four.”
RJ Haneke Jun 3, 2026 ▶ 47:52 Scaling Past Informal AI - Carina Hong, Axiom Math
Assertion Contradicted
Eifrem: Novo Nordisk graph deployment spans over 60 million documents
“Nowhere, nor this is one of the Like public case studies we have here over sixty million documents, you know, billions of notes and relationships use lots of kind of savvy NER and ER.”
Emil Eifrem Apr 18, 2026 ▶ 12:22 ⚡️ How to turn Documents into Knowledge: Graphs in Modern AI — Emil Eifrem, CEO Neo4J
Assertion Contradicted
Andreessen: AI Dungeon was the only public way to access GPT-3 for a year
“There was like a year where like the only way for a normal person to use GPT-III was in an AI dungeon.”
Marc Andreessen Apr 3, 2026 ▶ 5:15 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Assertion Contradicted
Swix: Sam Altman's Top AI Wish Is a Self-Completing To-Do List
“Do you know this is Sam Altman's number one ask for an AI app? It's the self-completing to-do list.”
Shawn Wang Mar 20, 2026 ▶ 21:10 Dreamer: the Agent OS for Everyone — David Singleton
Assertion Contradicted
Mirzadegan: Glean is the last incubation Kleiner Perkins has done
“I was working really closely with Arvind at Glean. It's actually the last incubation that we've done here.”
Joubin Mirzadegan Dec 12, 2025 ▶ 6:56 AI to AE's: Grit, Glean, and Kleiner Perkins' next Enterprise AI hit — Joubin Mirzadegan, Roadrunner
Assertion Contradicted
Ubl: Cognition and Cursor shipped RL fine-tunes of open-source models
“Just yesterday, I think we saw both Cognition, congrats, SWIX, and Cursor to ship RL fine tunes of unnamed open source models.”
Malte Ubl Dec 7, 2025 ▶ 5:22 The Great Evals Debate — Ankur Goyal & Malte Ubl
Assertion Contradicted
AWS has about 20,000 private equity-backed customers
“And what's interesting is Amazon has about AWS has about 20,000 customers that are backed by PE firms.”
Brendan Falk Sep 18, 2025 ▶ 2:23 ⚡️No, Don't Do Palantir for AI - Brendan Falk, Hercules (AUDIO FIXED)
Assertion Contradicted
Brockman: OpenAI's Dota AI used only 300 million parameters
“And by the way, Dota was like a three hundred million parameter neural net. Tiny, tiny little insect brain, right?”
Greg Brockman Aug 15, 2025 ▶ 15:20 Greg Brockman on OpenAI's Road to AGI
Assertion Contradicted
Google Gemini Flash-Lite ranks in top five on Galileo Agent Leaderboard
“In the, on the original leaderboard, we also have flashlight, which is, nobody talks about, I think, even now, but I think it's got a very decent score, which is, like, in the top five models. It's extremely cheap model, like, it's so dirt cheap that it makes …”
Pratik Bhavsar Jul 14, 2025 ▶ 7:50 ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
Assertion Contradicted
Ameisen: Golden Gate Claude feature specifically encoded awe of bridge's beauty
“We realized later on that it wasn't really like a Golden Gate Bridge feature. It was like being in awe at the beauty of the majestic Golden Gate Bridge, right?”
Emmanuel Ameisen Jun 6, 2025 ▶ 41:20 The Utility of Interpretability — Emmanuel Amiesen
Assertion Contradicted
Swyx: Gemini Flash accounts for 50% of OpenRouter requests
“Gemini Flash, according to Open Router, is now 50% of their Open Router requests.”
Shawn Wang Jan 1, 2025 ▶ 12:04 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
Assertion Contradicted
Distilling sCM Requires Roughly Twice the Compute of Teacher Training
“One thing that they said in the paper, it's not here, but that that it, they, it takes about two X to compute to train the This consistency model from as a as a distillation of whatever they distilled from. So approximately twice the compute.”
RJ Honicky Nov 2, 2024 ▶ 44:50 [Paper Club] Intro to Diffusion Models and OpenAI sCM: Simple, Stable, Scalable Consistency Models
Assertion Contradicted
Julien: BFCL v1 was the first dedicated LLM function calling benchmark
“So the first one that came out, like I said, in March was really the first of its kind to Make a function calling leaderboard.”
Sam Julien Oct 5, 2024 ▶ 2:10 [Paper Club] Berkeley Function Calling Paper Club! — Sam Julien, Writer
Assertion Contradicted
Prior video object segmentation models lacked error recovery mechanisms
“That actually is a big limitation of current models, current video object segmentation models. Don't allow any way to recover if the model makes a mistake.”
Nikhila Ravi Aug 7, 2024 ▶ 43:46 Segment Anything 2: Memory + Vision = Object Permanence — with Nikhila Ravi and Joseph Nelson
Assertion Contradicted
Tay: Most first-author papers by Jason Wei reach 1,000 annual citations
“Like, every single, so every single first author paper that, that, like, Jason writes in the last, has like, 1000 citations in one year. Like, no, I mean, not every, but like, most of it that he leads.”
Yi Tay Jul 5, 2024 ▶ 30:34 The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
Assertion Contradicted
Haisfield: Dylan Field tested generating Figma inside WebSim
“Dylan field actually posted this recently, like trying Figma in Figma or in web sim.”
Rob Haisfield Apr 27, 2024 ▶ 32:18 This World Does Not Exist — Joscha Bach, Karan Malhotra, Rob Haisfield (WorldSim, WebSim, Liquid AI)
Assertion Contradicted
Chintala: PyTorch is around 190,000 lines of code
“PyTorch is like a 190,000 lines of code or something at this point.”
Soumith Chintala Mar 6, 2024 ▶ 6:37 Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.