The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 11 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 1 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Assertion Not checkable as stated
Jack Morris: Fundamental AI science shifted to companies due to academic compute limits
“That's when I think things really started to change in terms of the types of questions you wanted to ask can't always be answered with academic resources. So a lot of the like fundamental kind of like boundary pushing and AI science moved into companies.”
Jack Morris Jul 2, 2025 ▶ 4:22 Information Theory for Language Models: Jack Morris
Assertion Not checkable as stated
Jack Morris: Most AI research was previously open, but is now closed
“Most stuff was open. Now most stuff is not open.”
Jack Morris Jul 2, 2025 ▶ 3:39 Information Theory for Language Models: Jack Morris
Assertion Supported
Morris: New embedding inversion model exactly recovers 90% of source text
“Like we ended up building a system that can do this quite well, like taking an embedding and I think our highlight number is like at a certain length, like a long sentence length, we can get 90% of the text back exactly.”
Jack Morris Jul 2, 2025 ▶ 26:53 Information Theory for Language Models: Jack Morris
Assertion Supported
Morris: Language models hit a hard memorization plateau regardless of dataset scaling
“Like, no matter how you scale the training size, you hit this like perfect, perfect ish plateau in auto memorization, which we call the model capacity.”
Jack Morris Jul 2, 2025 ▶ 38:32 Information Theory for Language Models: Jack Morris
Assertion Supported
Morris: 32-bit transformer models store only 3.6 to 3.9 bits per parameter
“Transformers that are trained in 32 bit precision, we approximate can store about 3.6 bits of information to maybe 3.9 bits somewhere in there per parameter.”
Jack Morris Jul 2, 2025 ▶ 56:00 Information Theory for Language Models: Jack Morris
Prediction Not checkable as stated
Morris: The next AI paradigm shift will stem from an unused data source
“And so whatever the fifth thing is, whether it's Video or embodied AI or some kind of crazy innovation on reasoning models. Whatever comes next will probably be some type of new data source that we're not using yet.”
Jack Morris Jul 2, 2025 ▶ 1:12:04 Information Theory for Language Models: Jack Morris
Assertion Contradicted
Morris: Top AI graduate programs do not teach multi-node distributed training
“Oh, to be clear, they don't teach you anything, like anything, like if you see a paper coming out from even, you know, Stanford, they're probably the best school in AI if you had to choose. And it's not like they're learning how to do like multi-node distribut…”
Jack Morris Jul 2, 2025 ▶ 9:38 Information Theory for Language Models: Jack Morris
Prediction Not checkable as stated
Morris: vLLM and SGLang are here to stay and will grow more complex
“I also think, ah, VLLM and SGLang seem, like, really good and important and here to stay. Like, they'll probably just get larger and more complex to accommodate future systems”
Jack Morris Jul 2, 2025 ▶ 14:10 Information Theory for Language Models: Jack Morris
Assertion Supported
Morris: Embedding inversion requires access to and repeated queries of the encoder
“Like none of the vector to text stuff works unless you have this assumption of like knowing the encoder and also being able to make a lot of queries to it.”
Jack Morris Jul 2, 2025 ▶ 32:44 Information Theory for Language Models: Jack Morris
Assertion Supported
Morris: CycleGAN mapping aligns disparate model embeddings without paired data
“We took it and we applied it to model embeddings where instead of zebras and horses, we have like BERT embeddings and GPT embeddings, or like two completely different models with different architectures. So I think these are GTR, which is a T five based retrie…”
Jack Morris Jul 2, 2025 ▶ 46:25 Information Theory for Language Models: Jack Morris
Prediction Open · timeframe Jul 2030
Morris: LLaMA architecture will likely store more information per parameter than GPT
“Maybe even if we tested this with LALAMA architecture, like, there's sort of like a GPT++ architecture, like, I would guess that can store better data just because the kind of numerical flow is a little bit better, the nonlinearities are maybe, like, A little …”
Jack Morris Jul 2, 2025 ▶ 57:30 Information Theory for Language Models: Jack Morris
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.