The Ledger, every show

Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.

shows every show 44 of 44
every show
clear all ✕
20VC Assertion Supported
Angelopoulos: Kimi K3 Beat Top US Closed Models in Web Development
“And for the first time ever, we saw a couple of weeks ago that Kimi K three actually beat the best closed source American models on a, you know, pretty important subset of tasks, for example, front end coding, like web development, which a huge fraction of dev…”
Anastasios Angelopoulos Aug 2, 2026 ▶ 3:00 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
20VC Assertion Supported
Angelopoulos: Most global AI inference spend remains on proprietary first-party APIs
“If you look at the whole space of all inference, most of it is still being consumed on first party APIs and on proprietary models. That's why anthropic revenue has been just a total hockey stick. It's not, you know, it's not like they're being completely canni…”
Anastasios Angelopoulos Aug 2, 2026 ▶ 5:44 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
20VC Assertion Supported
Angelopoulos: Arena has 30M+ monthly visitors, bigger than Hugging Face and xAI
“People don't know this, but Arena's one of the largest consumer AI apps in the world. We're bigger than, like, XAI. We're bigger than, like, Hugging Face, and Manus, and GenSpark, where it's so massive, like if you, like outside in, it's like 30 plus million m…”
Anastasios Angelopoulos Aug 2, 2026 ▶ 49:46 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
LATENT SPACE Assertion Supported
Arena sampled open-source models at 60/40, debunking Leaderboard Illusion paper
“But, you know, there, for example said that we were, that we only sampled, like, nine percent open source models and, like, you know, 60%, like, closed source models, and this created a gap between open and closed source. But in reality, we're actually really …”
Anastasios Angelopoulos Dec 31, 2025 ▶ 11:51 [State of Evals] LMArena's $1.7B Vision — Anastasios Angelopoulos, LMArena
a16z Assertion Supported
Angelopoulos: LMArena makes style control the default AI evaluation method
“That's why we're making style control default.”
Anastasios Angelopoulos May 29, 2025 ▶ 12:55 Beyond Leaderboards: LMArena’s Mission to Make AI Reliable
a16z Assertion Supported
Angelopoulos: LMArena router model outperforms all constituent models on Chatbot Arena
“When you train a prompt to leaderboard model, which is like, let's say a seven billion parameter model, and then you use it to route on just questions on the arena and everybody's questions, that model does better than any of the constituent models that were u…”
Anastasios Angelopoulos May 29, 2025 ▶ 1:29:10 Beyond Leaderboards: LMArena’s Mission to Make AI Reliable
a16z Assertion Supported
Angelopoulos: LMArena prompt router yields double the performance per dollar
“Now, if you trace the performance, the best performance that, you know, any individual model can give you as part of the router as a function of cost. That's like two X worse than the router. In other words, the router is giving you double the bang for your bu…”
Anastasios Angelopoulos May 29, 2025 ▶ 1:30:12 Beyond Leaderboards: LMArena’s Mission to Make AI Reliable
LATENT SPACE Assertion Supported
Angelopoulos: OpenAI o1 crushed Chatbot Arena, proving the benchmark isn't saturated
“So there's this model and it crushed the benchmark. You know, it's just like really like a big gap. And what that's telling us is that it's not saturated yet. And so it's still measuring some signal that was encouraging point.”
Anastasios Angelopoulos Nov 1, 2024 ▶ 27:20 In the Arena: How LMSys changed LLM Benchmarking Forever
LATENT SPACE Assertion Supported
Angelopoulos: The Chatbot Arena leaderboard is currently not an apples-to-apples comparison
“None of the leaderboard currently is apples to apples, because you have, like, Gemini Flash, you have, you know, all sorts of tiny models, like Llama Like, eight B and four or five B are not apples to apples.”
Anastasios Angelopoulos Nov 1, 2024 ▶ 28:03 In the Arena: How LMSys changed LLM Benchmarking Forever
20VC Assertion Supported
Angelopoulos: Chinese AI labs face severe hardware constraints and rely on black-market chips
“So they're way hardware constrained over there. And they've been trying to like black market import chips because of this. And you see this in the news, right? The information just reported on this.”
Anastasios Angelopoulos Aug 2, 2026 ▶ 14:01 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
20VC Assertion Supported
Angelopoulos: Google's Gemma sits on the performance-versus-cost Pareto curve
“Gemma, by the way, is pretty good in terms of efficiency. If you look at arena, you'll see the, on the Pareto curves of like performance versus cost. Gemma's on there.”
Anastasios Angelopoulos Aug 2, 2026 ▶ 24:37 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
20VC Assertion Supported
Angelopoulos: Harvey's CEO views model labs as his biggest competitive worry
“Harvey, the CEO of Harvey himself is saying that, you know, his biggest competitive worry is the model labs.”
Anastasios Angelopoulos Aug 2, 2026 ▶ 56:44 Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market
a16z Assertion Supported
Angelopoulos: Bradley-Terry models converge for AI evaluation, unlike Elo scores
“Okay, let's move from Elo to Bradley Terry because we're actually performing an estimate here instead of just like You know, and the ELO score moves over time. It doesn't converge, but Rally Terry models converge and how do we then construct confidence interva…”
Anastasios Angelopoulos May 29, 2025 ▶ 38:49 Beyond Leaderboards: LMArena’s Mission to Make AI Reliable
a16z Prediction Held up
Angelopoulos: LMArena will remain open-source as a commercial company
“We're going to keep publishing papers. We're going to keep releasing open source. We're going to keep releasing open data.”
Anastasios Angelopoulos May 29, 2025 ▶ 1:36:21 Beyond Leaderboards: LMArena’s Mission to Make AI Reliable
LATENT SPACE Assertion Supported
Angelopoulos: Arena received grants from Sequoia and a16z before incorporating
“He was not, you know, A-sixteen was not the only one to do this. We also had a great grant from Sequoia, but Ansh was in particular quite, quite supportive of us and, you know, gave us some resources in order to continue building out Arena before we even We're…”
Anastasios Angelopoulos Dec 31, 2025 ▶ 1:52 [State of Evals] LMArena's $1.7B Vision — Anastasios Angelopoulos, LMArena
LATENT SPACE Assertion Supported
Gradio scaled Arena to 1 million monthly active users before migration
“Gradio scaled us to a million Mal.”
Anastasios Angelopoulos Dec 31, 2025 ▶ 8:48 [State of Evals] LMArena's $1.7B Vision — Anastasios Angelopoulos, LMArena
LATENT SPACE Assertion Supported
Angelopoulos: Chatbot Arena scores are calculated via logistic regression
“The arena score that we show on our leaderboard is a particular type of linear model, right? It's a linear model that takes, it's a logistic regression that takes model identities and fits them against human preference, right? So it regresses human preference …”
Anastasios Angelopoulos Nov 1, 2024 ▶ 14:57 In the Arena: How LMSys changed LLM Benchmarking Forever
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.