frontier models
also referred to as: frontier model
19 statements across 17 episodes · 7 bullish · 9 bearish · 17 people on the record · first statement Aug 4, 2025 by Stefano Ermon · across every show →
Everything said about frontier models, oldest first
Aug 4, 2025 bullish
Ermon: Power constraints will drive diffusion models to replace frontier LLMs
“If it happens, it's gonna be driven by efficiency. Like we're all constrained by essentially power. And if you have, I mean, at the end of the day, it's all an inference game, right? Okay. Training is expensive, but then the thing that matters is being able to…”
Nov 22, 2025 bullish
Wagner: Fine-tuned open-source models can beat frontier models on quality
“I think you're now with, you know, we'll be seeing the trend going also with the coding agents, you know, fine-tuning open source models and like way faster as a result and more accurate. You can actually beat frontier models on quality that way.”
Dec 7, 2025 positive
Ubl: Vercel's Composite Models Are Faster Than Agentic Loops
“Basically what we do is we have this like composite model architecture. We run the frontier model and then we run the fine tune model after to fix its errors. That doesn't perform better than an agentic loop, but it's orders of magnitude faster, right?”
Dec 18, 2025 bullish
Zhang: AI models will handle simple vision natively, using tools for complexity
“I think at least I want to bet on, you know, running their work natively together, the future for simple, I would say for simple or even intermediate difficult vision tasks. For example, kind of counting with less than 20 objects. I think for this kind of simp…”
Jan 28, 2026 negative
Feb 10, 2026 neutral
Self-reflection training data is now core to all frontier foundation models
“What this suggests about the GPT training data is that the self-reflection data has now actually become pretty much core to the training of all frontier models, because we're seeing that happen in non-instruct models across the board.”
Feb 12, 2026
Jeff Dean: Capable small models require first building frontier models
“Through distillation, which is a key technique for making the smaller models more capable, you know, you have to have the frontier model in order to then distill it into your smaller model. So it's not like an either or choice. You sort of need that in order t…”
Mar 5, 2026 negative
Mar 5, 2026 negative
Huber: Frontier models repeat mistakes if failed actions remain in context
“A few of the insights is, like, everyone, frontier model is not good at search. Humans have this natural explore-exploit trade-off, where we kind of understand, like, when to stop doing something. Also, humans are pretty good at, like, forgetting, actually, li…”
Apr 18, 2026 bearish
May 28, 2026 neutral
Yan: Devin requires orchestrating multiple frontier models for end-to-end app testing
“Well, in some cases we found that actually no one frontier model can actually do this full end-to-end task itself. We've seen cases where we actually had had to orchestrate different frontier models together to kind of solve this problem together.”
Jun 4, 2026 positive
Jun 18, 2026 negative
Midha: Frontier AI Models Were Terrible at Analyzing Condensed Matter Physics Data
“We had started benchmarking frontier models on physics and science capabilities, and they were not very good. They were good at, like, doing things like summarization of papers, but if you said, hey, could you, like, analyze the scientific data coming out of a…”
Jun 21, 2026 bullish
Jun 22, 2026 negative
Kolter: Frontier models fail at red teaming due to safety refusals
“So generally speaking, the issue with this is that frontier models are extremely bad at automated red teaming because they have a lot of safeguards built into them. So if you try to use them to jailbreak another model, they will actually refuse their safety tr…”
Jun 24, 2026 bullish
Zaharia: Databricks' document vision model is ~100x cheaper than frontier models
“Our team built this document sort of vision model that takes a page and gives you back a nice JSON with all the components. And it's very competitive. It's like probably like a hundred X cheaper than those frontier models and still better.”
Jul 13, 2026 bearish
Biderman: Model accuracy will still degrade at 10M context window scale
“But two is like, for the agentic tasks of 18 months from now, inside those major repositories of knowledge, and asking the models more and more things in underspecified ways, I suspect that the accuracy of the models would go down. The phenomenon of context fr…”
Jul 13, 2026 negative
Biderman: Harmless enterprise queries on frontier models cost thousands of dollars
“And now you can solve these tasks with frontier models and compaction. And when you ask them to do so, they will consume thousands of dollars for queries that we think are harmless. That every employee in the company would be able to answer.”
Aug 21, 2026 bearish
Park: Frontier models hit only 20-30% accuracy predicting niche human behavior
“Where in some cases, the model performance of frontier models go all the way down to 20, 30%. Especially if you go into that more niche population on topics that our customers will actually care about. On more gen pop, it might be around 50 to 60%.”