frontier labs

also referred to as: frontier lab

8 statements across 8 episodes · 2 bullish · 2 bearish · 6 people on the record · first statement Jan 1, 2025 by Shawn Wang · across every show →

Everything said about frontier labs, oldest first

Jan 1, 2025
Insight
Swyx: Frontier AI labs distinguish themselves by adopting new benchmarks
“The labs that are not that frontier will keep measuring themselves on last year's benchmarks. And then the labs that are actually frontier will tell you about benchmarks you've never heard of.”
Shawn Wang Jan 1, 2025 ▶ 1:08:52 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
Jul 31, 2025 positive
Assertion Supported
Lambert: Frontier AI labs still rely on human preference data
“Every time I check in with people at frontier labs, they're like, yeah, we still use human preference data.”
Nathan Lambert Jul 31, 2025 ▶ 12:12 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Aug 29, 2025 bearish
Opinion
Morcos: Frontier AI lab data teams are systematically under-resourced
“I think you, what you see in all the frontier labs is that they have data teams. And if you talk to the folks that work on those data teams, what you'll kind of systematically hear is that typically they're under resourced relative to the gains that they're de…”
Ari Morcos Aug 29, 2025 ▶ 33:20 Better Data is All You Need — Ari Morcos, Datology
Sep 1, 2025
Disclosure
Landgraf: Ona will not build its own models, relying on frontier labs
“I think if you look from a business model perspective on this, I think at least right now, we are not planning to build our own models. I think ultimately that's a way out. You know, if you also are building your own models, we're not doing that. We are really…”
Johannes (Johannes Landgraf) Sep 1, 2025 ▶ 23:35 ⚡️Launching Ona: Coding Agent with Fully Sandboxed Cloud Environment
Oct 16, 2025 bearish
Opinion
Corbitt: Generic LLM-as-a-Judge Models Won't Beat Frontier Labs
“I'm pretty bearish on like Hey, this is a model that is trained as an LMS judge, but it's a generic LMS judge that can be used to judge anything. I just don't think you're going to beat the frontier labs on that.”
Kyle Corbitt Oct 16, 2025 ▶ 56:35 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Oct 18, 2025 bullish
Prediction Not checkable as stated
Merrill: Frontier AI labs will center operations around vertical products
“And with the Claude codes and the codec CLIs and the deep researchers, researchers, you starting to see some evidence that the products are going to be a much more central part of how these frontier labs operate.”
Mike Merrill Oct 18, 2025 ▶ 26:19 Terminal-Bench: Pushing Claude Code, OpenAI Codex, Factory Droid, et al to the limits
Oct 20, 2025 neutral
Assertion Contradicted
Swix: Every frontier lab now distills dense models into MoEs
“I think like, I think this is the pattern for every frontier lab now.”
Shawn Wang Oct 20, 2025 ▶ 36:32 ⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF
Nov 8, 2025
Assertion Not checkable as stated
Merrill: Terminal-Bench is used by every frontier AI lab
“It's used by all Frontier Labs in some capacity.”
Mike Merrill Nov 8, 2025 ▶ 2:18 Terminal-Bench 2.0: the most impt coding agent benchmark of 2025 gets a v2! Launch + Q&A w/ founders
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.