The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 15 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 1 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Opinion
Chen: Big Tech Could Fire 90% of Staff and Move Faster
“Like I used to work at a bunch of the big tech companies, and I always felt that we could fire 90% of people and we would move faster because the best people wouldn't have all these distractions.”
Edwin Chen Dec 7, 2025 ▶ 5:57 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Assertion Supported
Surge AI Surpassed $1B in Revenue With Under 100 Employees
“Yeah, so we hit over a billion of revenue last year with under a hundred people.”
Edwin Chen Dec 7, 2025 ▶ 5:40 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Prediction Open · timeframe Dec 2028
Chen: AI Efficiency Will Enable $100 Billion Revenue-Per-Employee Ratios
“And I think we're going to see companies with even crazier ratios, like a hundred billion per employee in the next few years. AI is just going to get better and better and make things more efficient. So that ratio just becomes inevitable.”
Edwin Chen Dec 7, 2025 ▶ 5:44 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Assertion Not checkable as stated
Chen: Frontier AI Labs Game Benchmarks via Prompt Tweaking and Test Leaks
“Sometimes, yeah, these benchmarks, they accidentally leak in certain ways, or the frontier labs will tweak the way they evaluate their models on these benchmarks. Like they'll tweak their system prompt. Or they'll tweak the number of times they run their model…”
Edwin Chen Dec 7, 2025 ▶ 19:32 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Assertion Not checkable as stated
Chen: LMSYS Chatbot Arena rewards bolding, emojis, and length over accuracy
“The easiest way to climb Alamarina, it's adding crazy boating. It's doubling the number of emojis. It's tripling the length of your model responses. Even if your model starts hallucinating and getting the answer completely wrong.”
Edwin Chen Dec 7, 2025 ▶ 24:15 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Opinion
Chen: Vibe coding is overhyped and will make codebases unmaintainable
“I definitely think that Vibe coding is overhyped. I think people don't realize, How much it's going to make your systems unmaintainable in the long term and decently dump this code into your code bases.”
Edwin Chen Dec 7, 2025 ▶ 51:54 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Insight
Chen: AI post-training is an art driven by taste, not pure science
“One of the things I often think about is that there's a, it's almost like there's an art to post training. It's not purely a science. Like when you were deciding what kind of model you're trying to create and what it's good at. There's this notion of taste and…”
Edwin Chen Dec 7, 2025 ▶ 15:35 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Opinion
Chen: Public AI Benchmarks Are Unreliable and Often Contain Wrong Answers
“I don't trust the benchmarks at all. And I think that's for two reasons. So one is, I think a lot of people don't realize, even researchers within the community, they don't realize that the benchmarks themselves are often honestly just wrong. Like they have wr…”
Edwin Chen Dec 7, 2025 ▶ 18:01 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Prediction Not checkable as stated
Chen: AI Will Automate 80% of L6 Engineer Tasks Within 2 Years
“In my head, I probably bet that within the next one or two years, yeah, the models are going to automate 80% of, you know, the average L six software engineer's job. But it's going to take another few years, do you move to 90%, and another few years to 99%, an…”
Edwin Chen Dec 7, 2025 ▶ 22:28 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Opinion
Chen: Anthropic stands out among frontier labs for principled model development
“I would say I've always been very, very impressed by Anthropic. Like, I think Anthropic takes a very principled view about what they do and don't care about. And how they want their models to behave in a way that feels a lot more principle to me.”
Edwin Chen Dec 7, 2025 ▶ 26:21 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Prediction Not checkable as stated
Chen: Achieving AGI will require breakthroughs beyond standard LLMs
“I'm in a camp where I do believe that something new will be needed.”
Edwin Chen Dec 7, 2025 ▶ 33:43 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Prediction Not checkable as stated
Chen predicts AI models will become increasingly differentiated across creator labs
“I think one of the things that's going to happen in the next few years is that the models are actually going to become increasingly differentiated because of the personalities and behaviors That the different labs have and the kind of objective functions that …”
Edwin Chen Dec 7, 2025 ▶ 48:20 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Opinion
Chen: Throwing bodies at data labeling fails to create quality AI data
“I think most people don't understand what quality even means in this space. They think you can just throw bodies at a problem and get good data, and that's completely wrong.”
Edwin Chen Dec 7, 2025 ▶ 9:48 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Disclosure
Surge AI tracks keystrokes and model gains to evaluate data annotators
“The way it works is we essentially gather thousands of signals about everything that you're doing when you're working on a platform. So we are looking at your keyboard strokes. We are looking how fast you answer things. We are using reviews. We are using code …”
Edwin Chen Dec 7, 2025 ▶ 11:57 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Opinion
Chen: Silicon Valley is just as money-obsessed as Wall Street
“Silicon Valley loves to score on Wall Street for focusing on money. But honestly, most of the Silicon Valley is chasing the same thing.”
Edwin Chen Dec 7, 2025 ▶ 29:47 The $1B Al company training ChatGPT, Claude & Gemini on the path to responsible AGI | Edwin Chen
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.