The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

why aren't all 6,166 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 36 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

What-if
Patel: U.S. CHIPS Act would not have passed without COVID car shortages
“Chips Act did not get passed, only got passed because that happened. And people are like, oh my God, the semiconductors are why cars can't be made. If that didn't happen, we wouldn't even have the Chips Act.”
Dylan Patel Feb 5, 2026 ▶ 44:31 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Opinion
Patel: AI infrastructure spend is not in a bubble yet
“I don't think it's a bubble yet.”
Dylan Patel Feb 5, 2026 ▶ 50:40 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Assertion Not checkable as stated
Patel: Two percent of global GitHub commits are generated by Claude Code
“But two percent of GitHub commits today are cloud code.”
Dylan Patel Feb 5, 2026 ▶ 51:53 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Opinion
Patel: AI is underearning the economic value it creates by a significant margin
“AI is under earning the value that it's producing in the world, right? By a significant margin already today.”
Dylan Patel Feb 5, 2026 ▶ 52:04 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Assertion Not checkable as stated
Patel: Engineer built an RTS game using $10K of Claude API
“He used, like, 10,000 dollars of Claude in one week and built an entire RTS from scratch about, like, but instead of, like, being a standard RTS where it's like, oh, Age of Empires where you advance through ages or Starcraft, it is an RTS where it's China vers…”
Dylan Patel Feb 5, 2026 ▶ 53:09 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Assertion Not checkable as stated
Patel: Only a few holdouts left writing code manually at Anthropic
“We have an indicator internally at Anthropic where you see how many people actually write code now. There's only a few holdouts left.”
Dylan Patel Feb 5, 2026 ▶ 53:42 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Prediction Open · timeframe Feb 2031
Patel: Data center power consumption will grow from 2% to 10% of US grid
“And then you've got data centers now all of a sudden coming online and going from two percent to 10% of the U S grid in just a handful of years.”
Dylan Patel Feb 5, 2026 ▶ 55:52 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Prediction Open · timeframe Dec 2030
Patel: Data centers will use under 1% of US water by 2030
“So the U.S. Grid will get to, like, 10% of power by, like, 28, 27, is data centers. For water consumption, it's not even gonna crack one percent. By the end of the decade.”
Dylan Patel Feb 5, 2026 ▶ 57:48 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Assertion Not checkable as stated
Patel: xAI's Colossus uses as much water as 2.5 In-N-Outs
“I think the metric was the entirety of Elon Musk's Colossus data center, right? Uses as much water as two and a half in and outs.”
Dylan Patel Feb 5, 2026 ▶ 59:05 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Disclosure
Patel: SemiAnalysis advised a client on buying and restarting a coal plant
“Have clients would like, had a client buy a coal plant. And we were advising them on the transaction based on, they just like showed up and they're like, yeah, we want to buy power assets.”
Dylan Patel Feb 5, 2026 ▶ 1:02:38 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Prediction Held up
Patel: OpenAI's next model will outperform Opus 4.5 around February-March
“OpenAI's new model, I think, will be better than Opus 4.5, and it's coming, like, somewhat soon in March-ish timeframe, maybe February, March-ish, but”
Dylan Patel Feb 5, 2026 ▶ 1:11:17 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Assertion Not checkable as stated
Patel: OpenAI has a better RL stack than Anthropic, but inferior pre-training
“Because OpenAI has a better RL stack than Anthropic today, it's just their pre-trained models suck compared to Anthropic's pre-training, right?”
Dylan Patel Feb 5, 2026 ▶ 1:11:26 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Assertion Not checkable as stated
Patel: Google has better pre-training than OpenAI or Anthropic, but worse RL
“Flip side, Google has a better pre-trained model than Anthropic or OpenAI, but their RL stack sucks.”
Dylan Patel Feb 5, 2026 ▶ 1:11:39 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Opinion
Patel: Opus 4.5 on Claude Code permanently changes how people work
“Opus 4.5 on Claude code is a new moment where the way you work has forever changed.”
Dylan Patel Feb 5, 2026 ▶ 1:12:04 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Prediction Not checkable as stated
Patel: AI coding UX will enable voice interaction within six months
“Give it six months, the models will be good enough that the UX can be like, talking to it.”
Dylan Patel Feb 5, 2026 ▶ 1:13:02 Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled”
Opinion
Text diffusion models will not replace autoregressive Transformers at state-of-the-art
“So it is a interesting direction to go into these diffusion, diffusion models as alternative to the auto regressive transformers, but it is not I would say the replacement at the state of the art.”
Sebastian Raschka Jan 29, 2026 ▶ 12:57 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Prediction Not checkable as stated
Future LLMs will prioritize architectural efficiency over larger model sizes
“I wouldn't expect bigger architectures. I would expect a more efficient architectures tweaks getting, The same modeling performance for less compute”
Sebastian Raschka Jan 29, 2026 ▶ 17:09 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Insight
Pre-training is no longer where the low-hanging AI gains lie
“Pre-training is not dead, but pre-training is boring. So it's not where the low hanging fruit is anymore.”
Sebastian Raschka Jan 29, 2026 ▶ 17:32 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Insight
RLVR unlocks pre-training knowledge rather than teaching LLMs new math
“The knowledge is already there in the pre-training, and this just unlocks it. It's just like a step that maybe shows the model how to use its own knowledge, basically.”
Sebastian Raschka Jan 29, 2026 ▶ 24:33 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Prediction Not checkable as stated
Process Reward Models will eventually become standard in LLM post-training
“I think it is promising and we will see it working at some point. I think it's just like right now it's still Tricky to make it work, but I am quite sure we'll see it as part of the standard repertoire at some point.”
Sebastian Raschka Jan 29, 2026 ▶ 26:39 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Insight
Bigger LLM gains will come from multi-model process refinement, not scaling
“That's where you make the bigger gains rather than scaling the model size. I think that's one of those things where you will see more progress coming from.”
Sebastian Raschka Jan 29, 2026 ▶ 28:04 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Assertion Supported
Large enterprises are secretly hiring teams to train ChatGPT-scale LLMs in-house
“I know for a fact that big companies are training now LLMs in-house. Really, like, big companies who have the financial means to train chat to be like model are hiring people who train LLMs.”
Sebastian Raschka Jan 29, 2026 ▶ 51:54 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Assertion Not checkable as stated
Dettmers: 4-bit precision is the end of quantization
“Four bit precision is the end of quantization.”
Tim Dettmers Jan 22, 2026 ▶ 16:03 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Assertion Not checkable as stated
Fu: AI coding tools enable expert programmers to move 10x faster
“But if you give an expert programmer This set of tools, they can go 10, 10 times faster than they were able to go before.”
Dan Fu Jan 22, 2026 ▶ 34:41 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Insight
Dettmers: Coding agents serve as general-purpose AI agents for digital tasks
“Coding agents are general agents. Coding agents can write programs that solve other problems, and code is so general, if there's a digital problem, you could solve it for code, and coding agents make the thing so easy that now you can solve a variety of proble…”
Tim Dettmers Jan 22, 2026 ▶ 36:19 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Insight
Dettmers: Students using AI agents perform poorly on basic domain knowledge
“If we let people use agents, they perform very poorly on basic knowledge. And if we let people just do the basic knowledge, they don't know how to use agents and they can't compete. So they can't do useful work in the workforce nowadays.”
Tim Dettmers Jan 22, 2026 ▶ 51:22 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Prediction Not checkable as stated
Dettmers: Humans will work on problems only AI agents understand
“In the future it's realistic that we work on problems that we don't understand, that agents understand, but we need to keep up in some way”
Tim Dettmers Jan 22, 2026 ▶ 52:07 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Disclosure
Ai2 to release open-source coding agent with 100x cheaper training
“We will have a major release of an open source coding agent that has a couple of key features. For one, training is a hundred times cheaper. You need to generate synthetic data and you need to train on it. And so we have a method that's a hundred roughly a hun…”
Tim Dettmers Jan 22, 2026 ▶ 52:57 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Assertion Not checkable as stated
Izmailov: AI sabotage and blackmail behaviors require contrived research scenarios
“In order to get those behaviors out of the models, you need to create somewhat of a contrived scenario or some special scenario. It's not necessarily something that we observe normally.”
Pavel Izmailov Jan 15, 2026 ▶ 2:07 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Assertion Not checkable as stated
Izmailov: Current AI models lack continual learning and cross-setting coherent goals
“I am pretty confident we are not there at the moment. I think right now the models are still acting in isolated environments, and we are not seeing a lot of evidence for very coherent goals across different settings”
Pavel Izmailov Jan 15, 2026 ▶ 2:58 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Insight
Izmailov: Rogue AI science fiction in training data likely causes deceptive behavior
“I think at least part of it is probably The models seeing descriptions of AI, like in the science fiction literature going rogue and like, yeah, that probably affects how the models behave in similar scenarios.”
Pavel Izmailov Jan 15, 2026 ▶ 4:09 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Insight
Izmailov: AI industry excels at execution but lacks bandwidth for exploration
“Industry is really great at executing on ideas and it's maybe not as good at, like, exploring diverse ideas. Even at the scale of Anthropic OpenAI there is a lot of focus in the companies, and there isn't a lot of bandwidth to do exploration, and that has been…”
Pavel Izmailov Jan 15, 2026 ▶ 12:27 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Prediction Not checkable as stated
Izmailov: Optimization pressure will cause AI to hide actual reasoning steps
“It seems like as soon as we start kind of applying some optimization pressure, the models will learn to hide what they're doing from the chain of thought.”
Pavel Izmailov Jan 15, 2026 ▶ 14:07 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Prediction Not checkable as stated
Izmailov: Future AI will produce outputs expert humans cannot reliably grade
“But in the future, we are imagining we will have models that are More capable than humans, and even expert humans will not be able to reliably grade very complicated answers from the model.”
Pavel Izmailov Jan 15, 2026 ▶ 18:40 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Assertion Supported
Izmailov: Anthropic research shows capable AI models are more likely to deceive
“You can see that the more capable the models are, the more likely they are to do this deception behavior.”
Pavel Izmailov Jan 15, 2026 ▶ 22:14 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Insight
Izmailov: AI models can quickly max out defined benchmarks using RL
“And I think we are at the stage where if we define a benchmark and we can make a relevant RL environment, then we can kind of max it out pretty quickly, and so we are going through benchmarks now very, very quickly.”
Pavel Izmailov Jan 15, 2026 ▶ 26:42 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Insight
Izmailov: Major compute multipliers exist that improve AI without naive scaling
“I think there are still major, like, compute multipliers, major ways of saving compute that can lead to better performance without just naively scaling.”
Pavel Izmailov Jan 15, 2026 ▶ 28:19 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Insight
Izmailov: Deterministic data transformations create information for computationally bounded models
“But with a limit on the compute, it's actually very possible to apply deterministic transformations to the data. And create information through that.”
Pavel Izmailov Jan 15, 2026 ▶ 35:37 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Prediction Not checkable as stated
Izmailov expects AI to outperform humans at proving technical mathematical lemmas
“In the mathematics I think we will see the models getting better on proving technical results, technical lemmas maybe including formalization and like things like lean the formal theory, improving language. I think the models, it's easy to imagine the models b…”
Pavel Izmailov Jan 15, 2026 ▶ 40:34 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Assertion Not checkable as stated
Izmailov: Transformer architectures will prove highly suboptimal for certain computational tasks
“At least for some tasks, I'm pretty confident that the transformers will be highly suboptimal.”
Pavel Izmailov Jan 15, 2026 ▶ 43:15 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Assertion Not checkable as stated
Bourgeau: AI progress from pre-training improvements is not slowing down
“It's still remarkable how much progress we're able to achieve in this way, and it's not really slowing down.”
Sebastien Bourgeau Dec 18, 2025 ▶ 2:26 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Assertion Not checkable as stated
Bourgeau: Google and DeepMind are actively researching post-Transformer architectures
“I believe so. There's groups doing research on the model architecture side, for sure, within Google and within DeepMind”
Sebastien Bourgeau Dec 18, 2025 ▶ 10:38 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Insight
Bourgeau: Architecture and data innovation currently matter more than scale
“The other parts are architecture and data innovation. These also play a really, really important part in the Performance of pre-training and probably even more so than pure scale these days, but scaling is still an important factor as well.”
Sebastien Bourgeau Dec 18, 2025 ▶ 31:22 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Assertion Not checkable as stated
Bourgeau: AI development is not running out of training data
“The other part of your question are we running out of data? I don't think so, so there's more.”
Sebastien Bourgeau Dec 18, 2025 ▶ 34:15 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Insight
Bourgeau: AI research is shifting to a data-limited paradigm
“I think what might be happening instead is kind of a shift in paradigm where before we were kind of scaling in the data unlimited regime where, where data would scale as much as you would like. And we're kind of shifting more to a data limited regime, which ac…”
Sebastien Bourgeau Dec 18, 2025 ▶ 34:26 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Prediction Not checkable as stated
Bourgeau: End-to-end differentiable retrieval and search in training will take years
“I think deep down, I do believe that the long-term answer is to learn this differentiable end-to-end way, which means probably doing pre-training or whatever that looks like in the future, Learn to retrieve as part of the training and learn how to do search as…”
Sebastien Bourgeau Dec 18, 2025 ▶ 40:09 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Insight
Bourgeau: AI models must be trained on harmful data to avoid it
“So at a fundamental level, you did, you do need the model to know about those things. So you have to train a bit at least on those so that it knows what those things are and knows to stay away from those, right?”
Sebastien Bourgeau Dec 18, 2025 ▶ 43:14 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Prediction Not checkable as stated
Bourgeau: Retrieval-augmented pre-training could become viable in a few years
“I just think it's not unreasonable to think in the next few years, something like that might actually become viable for a leading model like general.”
Sebastien Bourgeau Dec 18, 2025 ▶ 50:54 ”We’re Ahead of Where I Thought We’d Be” — Gemini 3 & the Future of AI
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.