model
also referred to as: models · the model
36 statements across 32 episodes · 16 bullish · 11 bearish · 32 people on the record · first statement Oct 21, 2023 by Kanjun Qiu · across every show →
Everything said about model, oldest first
Oct 21, 2023 positive
Apr 11, 2024 neutral
Byun: Early LLMs prioritized answering questions over faithfulness to source text
“At the time, the models hadn't been trained at all to be faithful to a text. So they were just generating. So then when you ask them a question, they tried too hard to ask, answer the question, and didn't try hard enough to answer the question given the text o…”
Jun 11, 2024 bullish
Conover: AI models can parse documents to identify second-order derivative bets
“It is very straightforward to take a model and say, parse through all of these documents and find second order derivative bets and say, oh, it turns out that energy is like very, very adjacent to investments in AI and may not be priced in the same way that GPU…”
Aug 28, 2024 positive
Carlini: LLMs decompile obscure binaries into readable Python code
“It can turn the compiled source code, which is impossible for any human to understand into the Python code that is entirely reasonable to understand. And, you know, it doesn't run. It has a bunch of problems, but like, it's so much nicer that it's immediately …”
Oct 11, 2024
Oct 11, 2024 bearish
Nov 11, 2024 negative
Polu: Airbyte's Notion connector output is not useful for AI models
“And the reality is that if you look at Notion, Airby does the job of taking Notion and putting it in a structured way, but that's a way that is not really usable to actually make it available to models in a useful way. Because you get all the blocks, details, …”
Nov 15, 2024
Dec 24, 2024 neutral
Fu: Efficient AI Architectures Are Dead on Arrival Without Hardware Co-Design
“Even if your model is theoretically more efficient, if somebody goes and runs it and it's two times slower one of the things that, that we've learned is that if you're in that situation, it's just going to be dead on arrival. So you want to be designing your a…”
Dec 25, 2024 negative
Neubig: AI Models Are Poor At Pixel-Based Web Navigation
“The first way is this, the simplest way and the newest way, but it doesn't work very well, which is you take a screenshot of the website and then you click on a particular pixel value on the website and like models are not very good at that at the moment.
Like…”
Feb 1, 2025 bearish
Nguyen: Website clicks will drop as internet access shifts to AI models
“In my opinion, like, people in, like, few years will click On, like, websites way less. I want to see the plot of, like, website clicks over time, but then my prediction is, like, it will go down and, like, people's access to the internet will be through the m…”
Mar 4, 2025 positive
Pokémon is ideal for testing AI agents because delays bring no penalty
“Pokemon's actually really nice because, like, if you don't do anything for five seconds, like, there's typically not a consequence by the nature of, like, doing inference on a model every, like, snapshot of time. It's actually a pretty good game to be able to …”
May 7, 2025 positive
May 7, 2025 bullish
Cherny: Foundation models will eventually subsume external memory and RAG architectures
“Everything is the model. Like that's the thing that wins in the end. And it just, as the model gets better, it's it subsumes everything else. So, you know, at some point the model will encode its own knowledge graph. It'll encode its own like KV story if you j…”
May 9, 2025 bullish
May 29, 2025 bearish
External scaffolding provides higher leverage for coding agents than fine-tuning models
“But our take in general is that freezing the model at a specific quality level and freezing the model at a specific data set just feels like it's lower leverage than continuing to iterate on all these external systems.”
Jul 11, 2025 neutral
Jul 16, 2025 bullish
Jul 31, 2025 positive
Lambert: RLVR on math does not degrade knowledge benchmark performance
“I think part of the intuition of RLVR is that the model is good at knowing which prompt area it is, which is why the models don't get worse on knowledge benchmarks if you're trading on like just math or precise instruction following. So the model just kind of …”
Aug 15, 2025 positive
Brockman: LLMs consistently generalize to untrained preferences
“In order to get them to be able to operate according to different preferences and values, we just need to show that to them during training, and they are able to sort of generalize to different preferences and values that we didn't actually train against, and …”
Aug 29, 2025 negative
Aug 29, 2025 bullish
Sep 30, 2025 bearish
Krieger: Current vision models lack the precision of skilled visual designers
“The models don't see as well as they could. They see, okay, you know, you ask them analyze a complex photo and they're able to do it, but I want them to be as persnickety as a like really good visual designer. Like, no, that looks, the baseline looks a little …”
Dec 26, 2025 positive
Dec 28, 2025 neutral
Dec 30, 2025 bearish
Nair: AI models lag orders of magnitude behind human one-shot error learning
“It seems like we're kind of, like, a few orders of magnitude of, like, kind of data efficiency, basically, away from, like, that kind of, like, you know, you do something once, or, like, you make a mistake like, you, yeah, you introduce, like, a bug in your co…”
Feb 5, 2026 neutral
Mar 17, 2026 neutral
Rieseberg: Unclear if agent hyper-optimizations remain relevant in next-gen models
“Will those gaps still exist in the next few generations of models? It's like a little unclear to me though. Because right now these like hyper optimizations we make, I'm not sure for how long they're still really relevant.”
Mar 17, 2026 bearish
Rieseberg: Specialized AI wrapper apps won't survive as models generalize
“I think we're going to see a lot of like applications and companies that do very impressive things with AI that in the short term might seem very effective because they're very specialized to individual use cases. But I think once models get better at generali…”
May 5, 2026 bullish
Frontier AI Models Can Solve Six-Month Graduate Physics Starter Problems
“And I think the issue is that many such problems now, I would say these models can probably crush. Yeah. These are problems that we usually take again, you know, timescale for a theoretical physics paper is six months to a year. That's pretty typical.”
Jun 1, 2026
Jun 4, 2026 bearish
Jun 22, 2026 negative
Fredrikson: Evaluation-aware AI models often execute harmful actions because it is a simulation
“If you make, if you're testing the model for robustness or safety, right? And it's aware that it's being tested because you've set things up in a very artificial way, right? Like the email addresses are at example.com. The webpage is clearly not a real webpage…”
Jun 22, 2026 positive
Jun 25, 2026 bullish
Chen: OpenAI's three-year goal is models conducting end-to-end research
“When we look at our kind of three-year roadmap, right the end goal that we want to reach is one where You know, the models are just doing end-to-end research, and I think a part of that problem is just being able to have the model come up with good taste.”
Jul 16, 2026 bullish
Beam: Cross-domain training reduces domain data requirements in science models
“And so again, the core bet that we're making is that is true for science. That if the model is trained on an increasingly broad set of data, the amount of data that you need in a given domain, that data requirement is reduced. In some cases will be reduced to …”