The Ledger, every show
Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.
shows 




every show 44 of 44
Angelopoulos: Fake AI applicants passed Arena's live technical engineering interviews
“They come in, they're like, hey, I want to be an infrastructure engineer at Arena, which is a great job that we're hiring for you. But then the other side of it is some guy that looks perfectly normal. They're getting, you know, they're passing all of our tech…”
Angelopoulos: Open-source AI pressure could cause insolvency for leading frontier labs
“And I think that the reason to be worried is because if the open source ecosystem somehow makes the cost saving opportunity for businesses much more salient and therefore decreases the revenue of companies like OpenAI and Anthropic within the enterprise, that …”
Angelopoulos: Kimi K3 Beat Top US Closed Models in Web Development
“And for the first time ever, we saw a couple of weeks ago that Kimi K three actually beat the best closed source American models on a, you know, pretty important subset of tasks, for example, front end coding, like web development, which a huge fraction of dev…”
Angelopoulos: Most global AI inference spend remains on proprietary first-party APIs
“If you look at the whole space of all inference, most of it is still being consumed on first party APIs and on proprietary models. That's why anthropic revenue has been just a total hockey stick. It's not, you know, it's not like they're being completely canni…”
Angelopoulos: Software will cease to be a business moat within five years
“Software is no longer really a moat because it can be produced instantaneously, right? Let's project out five years. That's what's going to happen. And so what moats exist? Network effects exist and data moats exist.”
Angelopoulos: US Will Have a Multi-Hundred Billion Dollar Open-Source AI Company
“I believe that we're going to have at least one massive, you know, multi-hundred billion, if not trillion dollar American company focused on American first open source.”
Angelopoulos: Thinking Machines' Inkling ranks #1 in US open source, behind nine Chinese models
“And within that time, they'd become the number one American open source model. But then the less generous take would be the companies existed for a year and a half. And then they've come up with, yes, the number one American open source, but there's nine Chine…”
Angelopoulos: US Likely to Restrict Chinese Open AI Models Within Three Years
“Yeah, my guess my guess would be that we will. I, I'm not saying I support it, but I think that it is likely where the world is headed. If I had to like place a bet, it would be there, but I think it's very uncertain at the moment.”
Angelopoulos: Anthropic currently enjoys "disgustingly high" inference gross margins
“For example, one of the things that's going to happen is that like right now, Anthropic has like disgustingly high growth, gross margins in their inference.”
Angelopoulos: Going public will exert downward pricing pressure on Anthropic's inference
“And after they go public, the whole world is going to see that, right? Like we're going to see their margins because those are going to be public information. And that's going to exert downward pricing pressure on their inference.”
Angelopoulos: Enterprises Will Need Guardian AI Models Because Humans Are Too Slow
“So I think we need guardian models and also, you know, agents within our businesses.
What is a guardian model?
Something that can witness the traces basically that's looking over the shoulder of every agent within a business and then saying, okay, this is a sa…”
Angelopoulos: Two-thirds of 75 AI neo-labs will fail or be acqui-hired
“There's at least 75 Neolabs. And For sure, like two thirds of those are going to be worth nothing or like they're going to be bought out for parts, right? That's going to be like an aqua hire.”
Angelopoulos: AI data market will reach $100B to $1T by 2030
“I believe it's going to be at least a hundred billion dollars by 2030, if not a trillion.”
Angelopoulos: Arena has 30M+ monthly visitors, bigger than Hugging Face and xAI
“People don't know this, but Arena's one of the largest consumer AI apps in the world. We're bigger than, like, XAI. We're bigger than, like, Hugging Face, and Manus, and GenSpark, where it's so massive, like if you, like outside in, it's like 30 plus million m…”
Angelopoulos: Nvidia is probably leading the race to a $10T valuation
“I think it's hard to say not NVIDIA. I think NVIDIA is probably in the lead there.”
Angelopoulos: Enterprise AI adoption will 10X Nvidia reliably
“But I think the enterprise adoption of AI is going to be another 10 Xer for the industry. I think it'll 10 X NVIDIA very reliably.”
Arena's anonymous Nano Banana test moved Google's stock and product roadmap
“I mean, that moment alone changed Google's like roadmap. Market share. Seriously. I mean, Google stock, billions of dollars are moving because of Nano.”
Angelopoulos: Chatbot Arena is immune to model overfitting by design
“Static benchmarks overfit. Why? It's because as Jan said earlier, you're giving the student the same test over and over. You have a model, you test it, you know, you look at whether or not it's improved on a static data set. Then you find another model, you te…”
Angelopoulos: China Unlikely to Ban Chinese AI Models in the US
“I don't really see them banning the use of Chinese models in the U.S. I don't think it makes sense for them.”
Angelopoulos: Enterprises fear relying on frontier AI labs and Chinese open-source
“It's not only true that they're terrified of working with the frontier labs, but they're also terrified of working with the Chinese open source.”
Angelopoulos: Frontier AI labs spend 10% to 20% of GPU compute budgets on data
“Companies are spending on it, usually within Frontier Labs, at about 10 to 20% about the amount that they're spending on GPUs.”
Angelopoulos: AI Will Eradicate Diseases Like Open Problems in Math
“I think that the, like, level of just human flourishing that's going to happen as we start to one by one eradicate diseases the same way that we're currently eradicating open problems in math is going to be incredible.”
Angelopoulos: Value in AI Biology Will Accrue to Data Layer
“That's exactly one of the areas where the data layer, where you can clearly see that the data layer is where value is going to accrue. Because the GPUs Are the same GPUs in both cases. The problem is that, that data infrastructure, the flywheel, the data colle…”
Arena sampled open-source models at 60/40, debunking Leaderboard Illusion paper
“But, you know, there, for example said that we were, that we only sampled, like, nine percent open source models and, like, you know, 60%, like, closed source models, and this created a gap between open and closed source. But in reality, we're actually really …”
Angelopoulos: LMArena makes style control the default AI evaluation method
“That's why we're making style control default.”
Angelopoulos: Future AI evaluation will shift to personalized user leaderboards
“Absolutely. Absolutely. It should be personalized just for you. You should understand which models are best for you.”
Angelopoulos: LMArena router model outperforms all constituent models on Chatbot Arena
“When you train a prompt to leaderboard model, which is like, let's say a seven billion parameter model, and then you use it to route on just questions on the arena and everybody's questions, that model does better than any of the constituent models that were u…”
Angelopoulos: LMArena prompt router yields double the performance per dollar
“Now, if you trace the performance, the best performance that, you know, any individual model can give you as part of the router as a function of cost. That's like two X worse than the router. In other words, the router is giving you double the bang for your bu…”
Angelopoulos: OpenAI o1 crushed Chatbot Arena, proving the benchmark isn't saturated
“So there's this model and it crushed the benchmark. You know, it's just like really like a big gap. And what that's telling us is that it's not saturated yet. And so it's still measuring some signal that was encouraging point.”
Angelopoulos: The Chatbot Arena leaderboard is currently not an apples-to-apples comparison
“None of the leaderboard currently is apples to apples, because you have, like, Gemini Flash, you have, you know, all sorts of tiny models, like Llama Like, eight B and four or five B are not apples to apples.”
Angelopoulos: Chinese AI labs face severe hardware constraints and rely on black-market chips
“So they're way hardware constrained over there. And they've been trying to like black market import chips because of this. And you see this in the news, right? The information just reported on this.”
Angelopoulos: China Has Already Restricted American AI Models Domestically
“And by the way, it's worth noting that China has already restricted the use of American models within China, right? So if you look at the two by two matrix of US China restrict, not restrict, you know, like export import stuff They have already restricted the …”
Angelopoulos: Google's Gemma sits on the performance-versus-cost Pareto curve
“Gemma, by the way, is pretty good in terms of efficiency. If you look at arena, you'll see the, on the Pareto curves of like performance versus cost. Gemma's on there.”
Angelopoulos: Harvey's CEO views model labs as his biggest competitive worry
“Harvey, the CEO of Harvey himself is saying that, you know, his biggest competitive worry is the model labs.”
Arena processes tens of millions of conversations monthly, totaling 250 million
“We have probably two hundred and fifty million conversations that happen over the course of the platform. We're on the order of, you know, mid tens of millions of conversations every month that are happening on the platform.”
Angelopoulos: Academic paper figures will soon be generated by AI models
“Soon we're not going to be even making them for our papers. We're, they're just going to be, our paper figures are going to be made by Emily.”
Angelopoulos: LMArena has released more real-world AI data than almost anyone
“We've probably released more data than basically anybody on the real world use cases of AI.”
Angelopoulos: Human evaluators prefer longer AI responses given equal content
“It's true that people vote for longer responses, you know, preferentially over shorter responses, even given the same contents or well-known human bias.”
Angelopoulos: LMArena measures user preference, not AGI progress
“We don't claim to be an AGI benchmark. We are faithfully representing the preferences of our community.”
Angelopoulos: Bradley-Terry models converge for AI evaluation, unlike Elo scores
“Okay, let's move from Elo to Bradley Terry because we're actually performing an estimate here instead of just like You know, and the ELO score moves over time. It doesn't converge, but Rally Terry models converge and how do we then construct confidence interva…”
Angelopoulos: Chatbot Arena has 1M+ monthly users and 150M+ conversations
“A lot of people don't know this, but ShopBot Arena is Used by like a million plus monthly users. We get like, you know, tens of thousands of votes on a daily basis. We have like over like, you know, a hundred fifty million conversations that have been had on t…”
Angelopoulos: LMArena will remain open-source as a commercial company
“We're going to keep publishing papers. We're going to keep releasing open source. We're going to keep releasing open data.”
Angelopoulos: Real-world testing will remain fundamental for evaluating AI agents
“The fundamental is organic, real-world testing with feedback. That's not going to change. I can tell you that that is not going to change.”
Angelopoulos: Five-model selection bias is tiny compared to voter variability
“We don't do that right now, partially because we kind of have know from simulations that the amount of selection bias you incur with these five things is just not huge. It's not huge in comparison to the variability that you get from the, from just regular hum…”
Angelopoulos: Arena received grants from Sequoia and a16z before incorporating
“He was not, you know, A-sixteen was not the only one to do this. We also had a great grant from Sequoia, but Ansh was in particular quite, quite supportive of us and, you know, gave us some resources in order to continue building out Arena before we even We're…”
LMArena funds all model inference running on its platform
“We fund all of the inference on the platform.”