The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 2,445 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 100 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Assertion Not checkable as stated
Yegge: OpenAI sees a 10x productivity gap between AI adopters and non-adopters
“Anecdotally, they're sharing that performance, the performance differences like 10 X By any way that you measure it. So lines of code, commits, business impact, whatever. And it's so stark and pronounced that the people who aren't adopting it are now 10 times …”
Steve Yegge Dec 26, 2025 ▶ 3:02 Steve Yegge's Vibe Coding Manifesto: Why Claude Code Isn't It & What Comes After the IDE
Prediction Not checkable as stated
Yegge: Top engineers avoiding AI will operate like interns within a year
“Look, I have friends Who are much better engineers than I am, ok, I mean, world class, maybe some of the best in the whole world, ok, have built technologies that you've heard of, and they're not using AI yet, except the occasional, I'll ask Cursor a chat ques…”
Steve Yegge Dec 26, 2025 ▶ 5:28 Steve Yegge's Vibe Coding Manifesto: Why Claude Code Isn't It & What Comes After the IDE
Assertion Not checkable as stated
Abraham: CloudChef robot outperforms expert chefs on 40-50% of commercial cuisines
“Our robot is able to do line cooking for about 40 to 50% of the world's commercially valuable cuisine to a point that if we put our robot against an expert chef in that cuisine, our robot is able to consistently make the food better than even the chef whose so…”
Nikhil Abraham May 31, 2025 ▶ 4:27 [AIEWF Preview] CloudChef: Your Robot Chef - Michellin-Star food at $12/hr (w/ Kitchen tour!)
Assertion Not publicly verifiable
Hotz: GPT-4 is an 8-way mixture model with 220B parameters per head
“GPT-IV is two hundred twenty billion in each head, and then it's an eight-way mixture model.”
George Hotz Jun 20, 2023 ▶ 49:48 Ep 18: Petaflops to the People — with George Hotz of tinycorp
Prediction Not checkable as stated
Slack predicts CLI coding agents could be dead within two months
“And that's the kind of thing that means that this transition, I think, is going to go a lot faster than people think. Because if security and devs both benefit from it, then, you know, I don't know, in two months, we could be seeing CLI coding agents as basica…”
Quinn Slack Sep 7, 2026 ▶ 18:31 Orbs: Shifting Coding to Cloud — Quinn Slack, Amp Code
Prediction Not checkable as stated
Swyx predicts coding SaaS will outlive all other SaaS
“Coding SAS will be the last SAS in the world because all the other SAS will just be buildable via coding SAS.”
Shawn Wang Sep 7, 2026 ▶ 23:22 Orbs: Shifting Coding to Cloud — Quinn Slack, Amp Code
Prediction Not checkable as stated
Slack predicts traditional CI's days are numbered due to AI
“Every time someone has said, oh, I don't trust the model to do X or Y. That doesn't last for that long. So it just feels like the days of CI as we know it are numbered”
Quinn Slack Sep 7, 2026 ▶ 25:51 Orbs: Shifting Coding to Cloud — Quinn Slack, Amp Code
Prediction Not checkable as stated
Lie: Groq will be forced to focus on significantly smaller models
“I think what's, what's going to end up happening is they're going to end up focusing on significantly smaller models. You know, if you have that limitation in your architecture, then I think that's what ends up happening.”
Sean Lie Sep 2, 2026 ▶ 23:42 The Inference Frontier: from 100 to 10,000 tokens per second — Sean Lie, Cerebras CTO
Assertion Not checkable as stated
Sean Lie: 95% of High-Quality Open Models Come From Chinese Labs
“The open source model market is a hundred percent Chinese, right? . A hundred percent, but. Almost. Okay. 95%, right? Most of the big models, most of the big open models that are, you know, high quality are coming from the Chinese labs.”
Sean Lie Sep 2, 2026 ▶ 41:22 The Inference Frontier: from 100 to 10,000 tokens per second — Sean Lie, Cerebras CTO
Assertion Not checkable as stated
Google DeepMind's six-month internal embargo shelves commercially valuable research papers
“I mean, what's worse is the paper is actually not even being Published anymore because there's a six month embargo inside of DeepMind, right? Like we've heard about this where a paper comes out and then I think there's a six month embargo window where if anybo…”
Anjney Midha Jun 18, 2026 ▶ 15:16 Why AI Labs With Unlimited GPUs Still Fail — Anjney Midha, AMP
Assertion Not checkable as stated
Awais: Claude Code hides 50+ tool-call failures per session on DeepSeek
“In CloudCode you know, they hide a lot of the errors behind control O, right? So you don't even know that, you know, you have like 50 plus tool call failures plus per session. You're just sitting there and you're like, oh, why is DeepSeq so slow?”
Ahmad Awais Jun 6, 2026 ▶ 12:21 ⚡️Making DeepSeek v4 outperform Opus 4.7 with Taste — @AhmadAwais , CommandCode.ai
Assertion Partly supported
Petersson: Anthropic's Claude models uniquely exhibit emergent deceptive and cartel behaviors
“So every single model from Anthropic since have been going in this direction. And I think one interesting thing is that like, OpenAI models don't. They, Quite plainly, they don't, they behave really well. And you know, you don't know if this is like, good, lik…”
Lukas Petersson Jun 4, 2026 ▶ 46:27 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Assertion Not checkable as stated
Hong: Competitor's AI demo can be solved entirely by Lean's grind tactic
“We're talking about, for example, the grind tactic in Lean. It can currently handle a lot of mass proofs, like, at a very low level. And this is pretty shocking because I have seen, you know, actually another company working in the same space, like, you know, …”
Carina Hong Jun 3, 2026 ▶ 10:27 Scaling Past Informal AI - Carina Hong, Axiom Math
Prediction Not checkable as stated
Hong: Informal math systems will not achieve math AGI
“I'm going to say on the record, we do not believe that an informal math system is going to be the math AGI solution.”
Carina Hong Jun 3, 2026 ▶ 50:41 Scaling Past Informal AI - Carina Hong, Axiom Math
Prediction Not checkable as stated
Burazin: SaaS revenue bumps from reselling LLM tokens will drop back down
“And I think that there will be a cold shower when people understand, like, no one's actually going to use and pay for these agents and tokens, and that wasn't actually real acceleration, but it'll drop back down.”
Ivan Burazin May 21, 2026 ▶ 1:05:23 AI Agents Need Computers: 74% MoM Growth, 850K/Day Runs, & New Agent Cloud — Ivan Burazin, Daytona
Prediction Not checkable as stated
Swix: Pull requests and code reviews are dying, replaced by prompt requests
“The pull request is dying. Yeah. Right? It's going to be the prompt request. And then beyond that, code review is also kind of dying”
Shawn Wang May 20, 2026 ▶ 1:12:23 The Agent-Native Cloud: 3M Users, 100K Signups/Wk, Data Centers, & Death PRs — Jake Cooper, Railway
Assertion Supported
Azhnyuk: FPV Drones Cause 70% to 80% of Frontline Casualties
“Out of all the casualties on the frontline, between 70 and 80% are done by FPV drones.”
Yaroslav Azhnyuk May 18, 2026 ▶ 25:18 FPV Drones -The Next War Is Already Here — Yaroslav Azhnyuk, The Fourth Law & Noah Smith, Noahpinion
Prediction Not checkable as stated
Azhnyuk: In 5-10 years, using weapons without AI will be immoral
“I think, you know, to your point, I think five to 10 years from now, it will be immoral to use weapons without AI. Because weapons without AI will be more likely to cause collateral damage or unwanted damage.”
Yaroslav Azhnyuk May 18, 2026 ▶ 45:32 FPV Drones -The Next War Is Already Here — Yaroslav Azhnyuk, The Fourth Law & Noah Smith, Noahpinion
Assertion Not checkable as stated
GPT-5 Reproduced Lupsasca's Best Physics Paper in 30 Minutes
“Then when GPT-V came out. It was able to reproduce one of my best papers that took me a very long time to come up with, in like, 30 minutes.”
Alex Lupsasca May 5, 2026 ▶ 2:32 🔬How GPT‑5 derived new results in theoretical physics and quantum gravity — Alex Lupsasca, OpenAI
Assertion Not checkable as stated
AI Resolved Theoretical Physics Problem That Puzzled Experts for a Year
“AI has become superhuman, at least on certain tasks. And that's what led to these recent papers which maybe we should talk about that resolve a problem that was puzzling physicists for experts in the field for over a year, and they weren't able to resolve it a…”
Alex Lupsasca May 5, 2026 ▶ 6:01 🔬How GPT‑5 derived new results in theoretical physics and quantum gravity — Alex Lupsasca, OpenAI
Assertion Not checkable as stated
ChatGPT Solved Open Physics Problem Before Collaborator's Flight Landed
“We decided to start working on it using AI a little bit before Andy was scheduled to come, like the week before. And in fact, using ChatGPT, we solved the problem before he even got off the plane.”
Alex Lupsasca May 5, 2026 ▶ 20:53 🔬How GPT‑5 derived new results in theoretical physics and quantum gravity — Alex Lupsasca, OpenAI
Assertion Not checkable as stated
Internal OpenAI Model Proved Gluon Amplitude Formula in 12 Hours
“We had this Internal model that could think for a very long time and was extra strong in physics. So we gave it the whole problem from scratch without actually giving it this. We just formulated the problem in a very sharp way and asked the model to solve, to …”
Alex Lupsasca May 5, 2026 ▶ 35:56 🔬How GPT‑5 derived new results in theoretical physics and quantum gravity — Alex Lupsasca, OpenAI
Assertion Not checkable as stated
Parakhin: Top AI models write code with fewer bugs than average humans
“I would claim by now, good model writes code on average with fewer bugs than average human.”
Mikhail Parakhin Apr 22, 2026 ▶ 12:19 AI-Native Engineering: 100% adoption, 5x search throughput, unlimited tokens — Mikhail Parakhin
Assertion Supported
Sachs: AI Model Quality Varies Between First-Party APIs and Cloud Providers
“Companies that say they're selling the same model through different vendors, whether it be through first party or Bedrock, Azure, et cetera, we do see different qualities sometimes, and that's not necessarily what's advertised.”
Sarah Sachs Apr 15, 2026 ▶ 24:37 Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
Assertion Contradicted
Andreessen: Three-year-old Nvidia chips make more money today than when new
“The current models are getting better faster at such a rate that if you are running an NVIDIA, if you're running an NVIDIA inference chip today that's three years old, you're making more money on it today than you did three years ago. Because the pace of impro…”
Marc Andreessen Apr 3, 2026 ▶ 23:23 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Prediction Held up
Andreessen: Autonomous AI agents will inevitably hire humans for tasks
“The agent hiring the people, which of course is going to happen, right? It's obviously going to happen.”
Marc Andreessen Apr 3, 2026 ▶ 41:32 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Prediction Not checkable as stated
Andreessen: Programming languages may not exist as a concept in 10 years
“Like I kind of think in 10 years, like I'm not sure, yeah, like I'm not sure there will even be a salient concept of a programming language in the way that we understand it today.”
Marc Andreessen Apr 3, 2026 ▶ 50:23 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Prediction Didn’t hold up
Andreessen: AI will never transform existing US K-12 public classrooms
“How are we going to apply AI in education? The answer is we're not because it's a literal government monopoly. It is never going to change the end, and there is nothing to do. By the way, you can create an entirely new school system. Like that's the one thing …”
Marc Andreessen Apr 3, 2026 ▶ 1:15:04 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Prediction Open · timeframe Apr 2031
Sun: Neural rendering with world priors will replace rasterizers and DLSS
“We actually believe that this is going to be the next paradigm of rendering. So it's going to replace how rasterizers, it's going to replace DLSS today because it not only has these pixel prior that's learned from the world, such that you can literally play an…”
Fan-yun Sun Apr 2, 2026 ▶ 30:36 Moonlake: Interactive, Multimodal World Models — with Chris Manning and Fan-yun Sun
Assertion Not checkable as stated
Unnamed materials foundation model is only 5x faster than DFT and unreliable
“It's only in my hands the one I'm still not naming is only about five times faster than my fastest DFT calculation on a GPU, and it also doesn't work all the time.”
Heather Kulik Mar 24, 2026 ▶ 20:32 🔬There Is No AlphaFold for Materials — AI for Materials Discovery with Heather Kulik
Prediction Not checkable as stated
Rieseberg: Future of software will not be hyper-personalized individual apps
“I actually don't think that the future is going to be hyper personalized software down to the point where everyone is running their own version. Like, I actually think it's going to be quite hard for one of us to have our own internal chat tool.”
Felix Rieseberg Mar 17, 2026 ▶ 9:31 Anthropic’s Felix Rieseberg on AI Coworkers, Local-First Agents, and the Future of Knowledge Work
Prediction Not checkable as stated
Dylan Patel: Political coup in Taiwan is more likely than full invasion
“What I think is most likely is that there is some sort of political coup or action. That destabilizes Taiwan in some way but doesn't actually, doesn't necessarily lead to an invasion, a full-scale invasion, and so sort of this is the best of the both worlds fo…”
Dylan Patel Feb 26, 2026 ▶ 8:53 Dylan Patel Explains the AI War While Cooking | In-Context Cooking
Assertion Not checkable as stated
Patel: Anthropic added $2B in monthly revenue at positive margins
“Anthropic doesn't just add two billion dollars of revenue in one month. You know, with, without having, you know, huge demand and they're doing it at positive margins, right?”
Dylan Patel Feb 26, 2026 ▶ 25:21 Dylan Patel Explains the AI War While Cooking | In-Context Cooking
Prediction Open · timeframe Dec 2027
Patel: Google will have zero free cash flow in 2027 due to AI CapEx
“There's no reason why Google will have any profit in 27 at all, right, in terms of cash flow. They will just spend every dollar they make On, on AI infrastructure.”
Dylan Patel Feb 26, 2026 ▶ 32:41 Dylan Patel Explains the AI War While Cooking | In-Context Cooking
Prediction Not checkable as stated
Patel: AI backlash will be the top issue in the next election
“And I think that's gonna be, like, the hottest button issue of, like, the next election, right? If not the midterms, right? And it seems obvious to me that, like, any party that wants to win should just become the anti-AI party.”
Dylan Patel Feb 26, 2026 ▶ 38:27 Dylan Patel Explains the AI War While Cooking | In-Context Cooking
Prediction Open · timeframe Feb 2029
O'Laughlin: Google and Meta will see free cash flow drop to zero
“Google, I would argue, is going to free cash with zero. I think Meta will go to free cash with zero.”
Doug O'Laughlin Feb 24, 2026 ▶ 1:27:11 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Assertion Supported
Bissell: CCP bias is identifiable in Qwen and DeepSeek-R1 representation spaces
“Well, there's, there are certainly internal, yeah, parts of the representation space where you can sort of see where that lives.”
Mark Bissell Feb 5, 2026 ▶ 10:08 Goodfire AI’s Bet: Interpretability as the Next Frontier of Model Design — Myra Deng & Mark Bissell
Assertion Contradicted
Hill-Smith: Google used unpublished 32-shot CoT to claim Gemini beat GPT-4
“Back when I'm Googled a Gemini one when I ultra and needed a number that would say it was better than GPT four. And Like, constructed I think never published, like, chain of thought examples, 32 of them in every topic in MLU to run it, to get the score.”
Micah Hill-Smith Jan 9, 2026 ▶ 8:36 Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith
Prediction Not checkable as stated
Ravisankar: Developer workforce will expand significantly in the future
“I actually think you're going to see a proliferation of developers. I think you're going to see Way more developers in, in the future. I think the job role will change for sure. I think it's not going to stay the same, but I think you're going to see way more.”
Vivek Ravisankar Nov 8, 2025 ▶ 44:56 ⚡️ The State of AI Engineer Hiring: Cheating, AI Adoption,Junior Devs — Vivek Ravisankar, HackerRank
Assertion Not checkable as stated
Sands: Top 100 AI startups on Stripe have unprecedented revenue per employee
“When you look at most of the top hundred AI companies on Stripe, their revenue per employee is Unlike any other business, including public companies who are known for being incredibly efficient companies.”
Emily Glassberg Sands Oct 30, 2025 ▶ 1:26:26 The Agents Economy Backbone - with Emily Glassberg Sands, Head of Data & AI at Stripe
Assertion Not checkable as stated
Corbitt: Prompt optimization methods like JEPA failed OpenPipe's agent benchmarks
“It didn't work on the problems we tried it on. It just didn't. It got like a minor boost over the sort of like more naive prompt we had and was just like, it was like, okay, Just kind of like our naive prompt with our model gets maybe like 50% on this benchmar…”
Kyle Corbitt Oct 16, 2025 ▶ 37:08 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Prediction Didn’t hold up
Swix: OpenAI will issue a cryptocurrency token to fund compute
“There is still one more shoe to drop, which is the non sovereign wealth funding that open AI needs to get, which they've promised to drop by the end of this year. And my money is on, they have to do a coin. Like it's, I'm not a crypto guy at all, but like, y…”
Shawn Wang Oct 16, 2025 ▶ 49:13 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Assertion Contradicted
Feldman: Cerebras is 20 times faster than Nvidia B200 GPUs
“Really focused on performance, both for training and for inference. You think 20 times faster than Nvidia B 200 GPUs and it's been an amazing run.”
Andrew Feldman Oct 1, 2025 ▶ 3:05 ⚡️Raising $1.1b to build the fastest LLM Chips on Earth — Andrew Feldman, Cerebras
Prediction Not checkable as stated
Slack: AI tools like Cursor and Claude Code will peak and decline
“A lot of these other tools that are great, like cloud code and codex and cursor and so on that they've forgotten what made them great and what made them grow so fast, which is building the very best product. And they built it in a way that's too overfit on the…”
Quinn Slack Sep 25, 2025 ▶ 22:29 Amp: The Emperor Has No Clothes
Prediction Not checkable as stated
Slack: Top AI labs face a major customer stampede within two months
“I think we are one or two months away from a possible news cycle. That is the foundation model companies have spent billions of dollars in capex and hired like crazy. And now, you know, they're no longer the best in this realm and there's a huge stampede away …”
Quinn Slack Sep 25, 2025 ▶ 30:49 Amp: The Emperor Has No Clothes
Prediction Not checkable as stated
Ball: Complex sub-agent workflows will result in user hangovers
“A lot of the features what we see, you know, where people build like elaborate workflows, like I have my custom slash commands and they trigger custom Custom sub-agents and they in turn trigger custom MCP tool calls behind which again another model is doing in…”
Thorsten Ball Sep 25, 2025 ▶ 35:22 Amp: The Emperor Has No Clothes
Assertion Contradicted
Bachman: Models claiming 256k+ context use windowed transformers, discarding data
“Anybody who says they're using a transformer With a context length of, you know, 256,000 or more, they're not using a true transformer. What they're using is a windowed transformer that essentially throws out a huge amount of its information at various layers …”
Diego Bachman Sep 23, 2025 ▶ 2:58 ⚡️ Beyond Transformers with Power Retention
Prediction Not checkable as stated
Ermon: Diffusion models could become the dominant architecture over autoregressive models
“I'm pretty optimistic about a future where diffusion models Can become the dominant solution. I've seen it happen before with GANs a few years ago, so I wouldn't be surprised if that's the case also here.”
Stefano Ermon Aug 4, 2025 ▶ 18:35 ⚡️Mercury: Ultra-Fast Diffusion LLMs — Estefano Ermon, CEO Inception Labs
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.