The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 9 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 0 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Assertion Not checkable as stated
Chen: GPT-4.5 performance jump matches leap from GPT-3.5 to GPT-4
“It signifies an order of magnitude improvement over the last models, kind of commensurate with the jump from 3.5 to four.”
Mark Chen Feb 27, 2025 ▶ 1:03 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Not checkable as stated
Chen: GPT-4.5 scaling returns remain consistent with OpenAI's prior projections
“You know, we are seeing the same returns, and I do want to stress that GPT-D 4.5 is that next point on this unsupervised learning paradigm, and, you know, we're very rigorous about how we do this. We make projections based on all the models we've trained befor…”
Mark Chen Feb 27, 2025 ▶ 7:55 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Not checkable as stated
Chen: GPT-4.5 hits expected benchmark progression consistent with OpenAI's trajectory
“Well, I really don't think that the accurate characterization is that it doesn't hit the benchmarks that, that we expect it to. So when you look at kind of the development of three to 3.5 to four to 4.5 this does hit the benchmarks that we expect.”
Mark Chen Feb 27, 2025 ▶ 21:08 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Prediction Not checkable as stated
Chen: GPT-5 could combine unsupervised scaling with reasoning paradigms
“And so I think, like GPT-V really could be the culmination of a lot of these things coming together.”
Mark Chen Feb 27, 2025 ▶ 3:24 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Partly supported
Chen: Users prefer GPT-4.5 over GPT-4o by 60% to 70% margins
“When we look at, kind of, comparisons against GPT-FORO you'll see that everyday use cases, people prefer, you know, by a margin of 60% for actually productivity and knowledge work against GPT-FORO, there's almost like a 70% preference rate.”
Mark Chen Feb 27, 2025 ▶ 5:14 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Not checkable as stated
Chen: Nearly all large language models today utilize mixture of experts
“I think pretty much all large language models today use, utilize mixture of experts.”
Mark Chen Feb 27, 2025 ▶ 11:38 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Not checkable as stated
Chen: Pausing and restarting training runs is standard across OpenAI models
“Actually, so I think it's interesting that this gets is a point that's attributed to this model because actually in, in, in developing all of our foundation models, right they're all experiments, right? I think you know, running all of the foundation models of…”
Mark Chen Feb 27, 2025 ▶ 8:51 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Not checkable as stated
Chen: GPT-4.5 creates ASCII art almost flawlessly, unlike previous models
“If you ask any of the previous models to create ASCII art for you, right? Actually, they mostly just fall down. This one can do it Almost flawless.”
Mark Chen Feb 27, 2025 ▶ 19:54 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Assertion Not checkable as stated
Chen: OpenAI inference costs dropped orders of magnitude since GPT-4
“The costs have dropped, you know, many orders of magnitude since we first launched GPT-IV.”
Mark Chen Feb 27, 2025 ▶ 11:21 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.