The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

why aren't all 6,166 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 36 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Insight
Douglas: Independent technical blogs are the highest AI hiring signals
“The fastest route, or like, the most immediate one is whenever we see a really good blog post where people have, like, done incredible amount of work in an independent fashion, it's one of the highest signal things there is.”
Sholto Douglas Oct 2, 2025 ▶ 10:54 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas predicts DeepMind will lead the world in AI science discoveries
“DeepMind, if you wanted to solve science, is the best place in the world. Like, I think that DeepMind will directly contribute to more scientific discoveries from AI than anything else, right?”
Sholto Douglas Oct 2, 2025 ▶ 16:54 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Disclosure
Douglas: Anthropic's ethos is that scaling current techniques achieves AGI
“Like really for the last five or six years, Anthropics ethos has been scaling compute with broadly the current set of techniques is like AGI is tractable within those bounds.”
Sholto Douglas Oct 2, 2025 ▶ 27:52 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: AI takeoff speed depends on AI assisting AI research
“We think that one of the most important signals of whether or not we are basically the speed of takeoff, the speed of progress is driven by how much AI is able to assist AI research.”
Sholto Douglas Oct 2, 2025 ▶ 29:18 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Opinion
Douglas: Claude 3.5 Sonnet drove product-market fit for code editor Cursor
“In many ways, this model is what caused PMF for Cursor. Cursor took off like a rocket, right, with because they were in the right place, and they were able to capitalize on that model as offering a coding experience that didn't previously exist.”
Sholto Douglas Oct 2, 2025 ▶ 34:37 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: An Anthropic AI agent operated autonomously for 30 hours building apps
“We asked it to build something that looks roughly like a chat app, you know, something like Slack or, you know. And it was, it, the model just worked for 30 hours. Like, it was just spinning there on a computer for 30 hours, and came out with a really good wor…”
Sholto Douglas Oct 2, 2025 ▶ 36:03 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Supported
Douglas: AI autonomous task execution time horizons double every six months
“And so I think it's like every couple of months, the time horizon that the AIs are capable of doing is doubling or something, something crazy. Maybe, maybe every six months the time horizon doubles”
Sholto Douglas Oct 2, 2025 ▶ 40:13 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas: AI application development will see another massive leap next year
“Over the next six months, over the next year, expect dramatic progress here. And like look at where we are now versus where we were a year ago. And the difference is I expect the same jump basically.”
Sholto Douglas Oct 2, 2025 ▶ 42:57 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: AI coding interventions stem from taste, not raw programming capability
“Right now you need to intervene quite frequently, but it's usually on questions of taste rather than it is questions of, like, raw programming ability.”
Sholto Douglas Oct 2, 2025 ▶ 43:54 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: Solving AI hallucinations intrinsically requires reinforcement learning
“Saying, I don't know, or solving, you know, hallucinations is, ah, intrinsically requires reinforcement learning in many ways.”
Sholto Douglas Oct 2, 2025 ▶ 49:35 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Insight
Douglas: AI reasoning strategies emerge naturally with enough compute and RL feedback
“Give it math questions, tell it whether it got them right or wrong, and the model will learn. This is, it comes down to a bit of lesson in scale and search, is just allow the model to search, have enough compute to run the experiments, and the model actually e…”
Sholto Douglas Oct 2, 2025 ▶ 55:19 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Opinion
Douglas: LLM pipelines are two and a half years of desperate effort
“When I look at an LLM training pipeline, it is two and a half years of best effort, last minute, desperate effort. And there's just so much room to go on every part of it.”
Sholto Douglas Oct 2, 2025 ▶ 1:03:26 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas: AI beating GDP benchmarks won't immediately alter the broader economy
“We'll probably reach like better than human on the GDP eval, and it won't change anything economically because It'll be all the connective tissue, and all the, like, you know, the context, and actually, like, the task won't be representative.”
Sholto Douglas Oct 2, 2025 ▶ 1:04:53 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Not checkable as stated
Douglas: Individuals will manage 24/7 AI agent teams within two years
“If coding agents progress in the way I've been saying, in a year or two, you'll be able to manage a team, basically, that works 24 seven for you doing work.”
Sholto Douglas Oct 2, 2025 ▶ 1:06:11 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Assertion Not checkable as stated
Douglas: Robotic locomotion is essentially solved using basic reinforcement learning
“Locomotion's kind of solved, to be honest, with basic RL.”
Sholto Douglas Oct 2, 2025 ▶ 1:08:09 Sonnet 4.5 & the AI Plateau Myth — Sholto Douglas (Anthropic)
Prediction Open · timeframe Sep 2035
Crespo: Microsoft Excel will still be actively used in ten years
“I have a prediction that Excel will still be here in five years. And there is a good reason. You can keep that. In five years, Excel will still be there. And in 10 years, Excel will still be there.”
Eleonore Crespo Sep 11, 2025 ▶ 36:31 Goodbye Excel? AI Agents for Self-Driving Finance – Pigment CEO
Disclosure
Crespo: Pigment aims to build a $100B+ to $200B+ business
“We have a very, very large ambition with Romain. We want to build a, You know, a hundred billion dollar business, if not more two hundred billion, if we can, or even more.”
Eleonore Crespo Sep 11, 2025 ▶ 1:04:07 Goodbye Excel? AI Agents for Self-Driving Finance – Pigment CEO
Prediction Not checkable as stated
Valenzuela: Unified AI models will render specialized video task models completely obsolete
“Our thesis has always been that, like, you don't, that's not gonna matter. Like, it's just, none of those things will matter the moment you have a model that can learn how to do all of those things at once.”
Cris Valenzuela Sep 4, 2025 ▶ 25:25 AI Video’s Wild Year – Runway CEO on What’s Next
Insight
Valenzuela: AI software scales through core principles rather than vertical specialization
“It used to be the case that you pick verticals. I think now you pick principles. And the principles allow you to scale much better than any other previous generation of software.”
Cris Valenzuela Sep 4, 2025 ▶ 28:57 AI Video’s Wild Year – Runway CEO on What’s Next
Opinion
Valenzuela: Runway's true moat is its shipping infrastructure, not any individual model
“And I think that for me is the most valuable part of runway. It's not a model that we put out because models will like Completely like change every single, every now and then is the organizational knowledge and the infrastructure that it takes to ship a model …”
Cris Valenzuela Sep 4, 2025 ▶ 42:35 AI Video’s Wild Year – Runway CEO on What’s Next
Prediction Not checkable as stated
Valenzuela: Full-stack AI companies will eventually reach 80% to 90% SaaS margins
“I think eventually you'll get to, like, best in SAS margins, like, as 80, 90% over time.”
Cris Valenzuela Sep 4, 2025 ▶ 56:57 AI Video’s Wild Year – Runway CEO on What’s Next
Disclosure
Pedregal: Granola is a Trojan horse to aggregate user context
“And in a way the product we have today is, is a Trojan horse to collect a lot of your context so that you can then use all the information in that context to do future work.”
Chris Pedregal Aug 21, 2025 ▶ 37:19 How to Build a Beloved AI Product - Granola CEO Chris Pedregal
Disclosure
Cherny: Anthropic builds minimal product interfaces to keep up with model evolution
“The way we think about it is the model is evolving so quickly that we build a minimal possible product to keep up with it.”
Boris Cherny Aug 7, 2025 ▶ 11:55 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Insight
Cherny: Code is the primary path for AI to reach AGI
“Maybe coding is the way that we get to the next level of intelligence. If you call it like AGI or ASI or whatever, the model needs some way to interact with the world and for a model, the natural way is code.”
Boris Cherny Aug 7, 2025 ▶ 14:40 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Prediction Not checkable as stated
Cherny: AI models won't need rigid sub-agent roles in 6-12 months
“But I think that the models six or 12 months from now, they probably won't need this anymore because they're all going to be pretty good and you won't have to define very rigidly what each one's responsibilities are anymore.”
Boris Cherny Aug 7, 2025 ▶ 27:02 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Assertion Supported
Cherny: Claude Code does not use RAG for codebase memory
“And so quad code actually doesn't use this technique called rag. Instead, what it does is it just searches files the same way that a human would.”
Boris Cherny Aug 7, 2025 ▶ 30:34 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Opinion
Cherny rates Claude Code 10/10 for codebase research, 6/10 for writing
“I feel like if you look at answering questions about the code base and kind of code base research, I think it's like 10 out of 10 good. It's as good as it can get. When it comes to writing code, it's maybe like a six out of 10. It's pretty good. It won't get e…”
Boris Cherny Aug 7, 2025 ▶ 45:28 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Disclosure
Cherny: I use both Claude Code and Cursor every day
“I personally use a lot of these products and you know, I use quad code every day, but I also use cursor every day and I use other products every day.”
Boris Cherny Aug 7, 2025 ▶ 50:58 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Prediction Not checkable as stated
Boris Cherny: Programming will shift from text manipulation to working with agents
“I think one way it will definitely play out is it's going to change programming where programming is no longer direct text manipulation, but it's more working with agents to get the work done.”
Boris Cherny Aug 7, 2025 ▶ 56:57 Anthropic's Surprise Hit: How Claude Code Became an AI Coding Powerhouse
Prediction Not checkable as stated
Laskin: AI models will interact with enterprise software primarily via APIs
“And so the way these language models are going to interact with any piece of software, not just Software engineering software, like Salesforce and other CRMs and creative tools and so forth. The majority of those interactions are going to be through function c…”
Misha Laskin Jul 17, 2025 ▶ 7:22 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Prediction Not checkable as stated
Laskin: Advanced AI without comprehension are 'L9 engineers with amnesia'
“What we're going to get to if we don't solve the comprehension piece is basically L-nine engineers with amnesia.”
Misha Laskin Jul 17, 2025 ▶ 10:41 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Insight
Laskin: Solving organizational code context yields all capabilities for superintelligence
“Like if you really solve this oracle for organizations just for coding, you've basically built all the capabilities you need to have super intelligence.”
Misha Laskin Jul 17, 2025 ▶ 15:53 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Insight
Laskin: Reinforcement learning makes LLM capabilities jagged, not broadly general
“When you train large language models with reinforcement learning, they become jagged in the sense that they become good at what you wanted them to be good at. And there are some generalization capabilities, but they're much weaker than people think.”
Misha Laskin Jul 17, 2025 ▶ 49:32 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Prediction Not checkable as stated
Laskin: Principal-level AI engineers are a couple of years away
“And that the combination of this you know, L-Nine with Amnesia and the L-Nine's context core will together, you know, that will become the principal level engineer, the AI engineer. And so I actually think that that's not too far away. That's I would say in, y…”
Misha Laskin Jul 17, 2025 ▶ 56:41 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Assertion Not checkable as stated
Laskin: Reflection AI regularly beats OpenAI, Anthropic, and DeepMind for talent
“We win over candidates over OpenAI and Anthropic Meta, DeepMind regularly.”
Misha Laskin Jul 17, 2025 ▶ 1:01:22 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Prediction Not checkable as stated
Laskin: Reflection AI will ship research requiring 100k GPU equivalence in 2025
“Later this year we'll be shipping things that I don't think anyone ever thought a startup could do. Like, I think that we're going to be shipping some things on the research side that I think everyone thinks you need to be a giant lab with a 100,000 GPUs to do…”
Misha Laskin Jul 17, 2025 ▶ 1:02:06 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Insight
Laskin: Focused AI startups can operate with 10x less capital than frontier labs
“You can't operate at a hundred X less capital than a frontier lab, but you can operate at, say, 10 X, like an order of magnitude less capital when you're really focused.”
Misha Laskin Jul 17, 2025 ▶ 1:05:03 Ex‑DeepMind Researcher Misha Laskin on Enterprise Super‑Intelligence | Reflection AI
Prediction Not checkable as stated
Merchant APIs will require canonical intent endpoints for autonomous AI agents
“And so every merchant API is probably going to need one canonical kind of intent endpoint that accepts those structured desires instead of sort of this UI click world that we live in today.”
Emily Glassberg Sands Jul 10, 2025 ▶ 52:34 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Prediction Not checkable as stated
Merchants must expose machine-readable product schemas to sell through AI agents
“And so I think early adopters who want to sell through agentic channels are going to need to expose kind of an open product schema, like the SKU and the inventory and the price and the constraints and, you know, maybe even the wedge that you're willing to give…”
Emily Glassberg Sands Jul 10, 2025 ▶ 53:06 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Prediction Not checkable as stated
AI agents will shrink e-commerce latency budgets to a few hundred milliseconds
“I think latency budgets are gonna shrink to machine time. We talked about latency budgets in the context of the charge path, but, like, you know, people will wait three seconds for a spinner. I think an agent's just gonna retry somewhere else after a couple hu…”
Emily Glassberg Sands Jul 10, 2025 ▶ 53:40 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Prediction Not checkable as stated
Model Context Protocol is becoming the default standard for LLM integration
“I mean, it's pretty clear that MCP is becoming the default way that any single service Stripe or GitHub or Notion talks to an LLM.”
Emily Glassberg Sands Jul 10, 2025 ▶ 59:05 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Assertion Not checkable as stated
AI startups reach $30M ARR three times faster than fast-growing SaaS startups
“Those that already hit thirty million in annualized revenue got there in about a year and a half. For comparison, you know, many of us were around five years ago, like the fastest growing SaaS startups on Stripe took, you know, five and a half years to hit tha…”
Emily Glassberg Sands Jul 10, 2025 ▶ 1:01:34 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Assertion Supported
Stockholm-based AI startup Lovable reached $50 million ARR in six months
“European breakouts lovable out of Stockholm hit fifty million ARR in six months and is now for sure the fastest growing startup in Europe.”
Emily Glassberg Sands Jul 10, 2025 ▶ 1:02:20 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Prediction Not checkable as stated
AI software pricing will shift to outcome-based models within five years
“I think it's where actually like the market equilibrium, like where clearing will actually happen, you know, two, three, five years from now is experimenting with new pricing models, like outcome based pricing and actually increasingly using outcome based pric…”
Emily Glassberg Sands Jul 10, 2025 ▶ 1:08:57 The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)
Assertion Not checkable as stated
Walsher: Cursor announced reaching $500M in ARR
“Cursor, maybe three weeks ago, announced they're at five hundred million dollars of ARR.”
David Walsher Jul 3, 2025 ▶ 7:51 AI Engineering Revolution: Winners, Chaos & What’s Next | FirstMark
Insight
Vercel CTO says top engineers are becoming primarily code reviewers
“Malte, among many interesting things, I think one of the most fascinating things that he said was that most of his great engineers who have been with the company for a while, as they've dogfed you know, vZero, and they've used things like Cursor and Windsurf i…”
David Walsher Jul 3, 2025 ▶ 20:10 AI Engineering Revolution: Winners, Chaos & What’s Next | FirstMark
Assertion Partly supported
Walsher: Computer science graduates face top-tier college unemployment rates
“Computer science grads are actually among the top five or six majors graduating from college right now with the highest unemployment rate.”
David Walsher Jul 3, 2025 ▶ 29:46 AI Engineering Revolution: Winners, Chaos & What’s Next | FirstMark
Insight
Walsher: Tech buyers value Cursor's opinion over $80B legacy software incumbents
“Most people will care, at least in our world, What the smartest buyer at Cursor thought about a given tool than what the smartest person at the, you know, eighty billion software business that IPO'd in, in 2012 might think.”
David Walsher Jul 3, 2025 ▶ 45:25 AI Engineering Revolution: Winners, Chaos & What’s Next | FirstMark
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.