LMSYS Chatbot Arena
part of LMSYS
4 statements across 4 episodes · 3 bullish · 0 bearish · 4 people on the record · first statement Jan 11, 2024 by Nathan Lambert · said 87 times in 25 episodes since 2023 · across every show →
Mentions by year
brought up most by Shawn Wang (43), Anastasios Angelopoulos (13), Nathan Lambert (9), Alessio Fanelli (2), Vibhu Sapra (1), Vasek Mlejnsky (1), Tri Dao (1), Thomas Scialom (1)
2026 3 mentions in 2 episodes 2 per episode
2025 57 mentions in 11 episodes 5 per episode
- [State of Evals] LMArena's $1.7B Vision — Anastasios Angelopoulos, LMArena
- The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
- [State of Research Funding] Beyond NSF, Slingshots, Open Frontiers — Andy Konwinski, Laude Institute
- ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
- [State of Code Evals] After SWE-bench, Code Clash & SOTA Coding Benchmarks recap — John Yang
- [State of Context Engineering] Agentic RAG, Context Rot, MCP, Subagents — Nina Lopatina, Contextual
- Greg Brockman on OpenAI's Road to AGI
- ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
- 3 more episodes that year, every mention in 2025 →
2024 26 mentions in 11 episodes 2 per episode
- The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
- In the Arena: How LMSys changed LLM Benchmarking Forever
- Training Llama 2, 3 & 4: The Path to Open Source AGI — with Thomas Scialom of Meta AI
- The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
- Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
- The State of AI Startups in 2024 [LS Live @ NeurIPS]
- [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz
- Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
- 3 more episodes that year, every mention in 2024 →