reasoning models
also referred to as: reasoning model
9 statements across 8 episodes · 6 bullish · 1 bearish · 9 people on the record · first statement Oct 17, 2024 by Sarah Guo · across every show →
Everything said about reasoning models, oldest first
Oct 17, 2024 bullish
Sarah Guo predicts massive excitement when reasoning models gain tool integration
“I do think on how it is received and how people use it, as the labs, you know, attach tools and function calling to the ability to do this type of planning and iterative reasoning, I think people are going to get a lot more excited.”
Nov 21, 2024 bullish
Reasoning Models Will Become Highly Robust Within Two to Three Years
“I think right now it's extremely inefficient and it's quite brittle, similar to the early versions of language models. But over the next two or three years, it's gonna become incredibly robust and unlock just a whole new set of problems.”
Mar 5, 2025 negative
Hendrycks: Reasoning models have reached expert-level virology capabilities
“The AIs are getting very good at STEM PhD level types of topics, and that includes virology. So I think that they are sort of rounding the corner on being able to provide expert level capabilities in terms of their knowledge of the literature, Or even helping …”
Mar 13, 2025 bullish
Dohmke: Improved model reasoning will push SWE-bench scores near 100%
“As the models get better in reasoning we're going to get closer to a hundred percent of this VBench, which is that benchmark out of 12 repos open source Python repos a team in Princeton identified 2200 or so issue pull request pairs. Effectively, all the model…”
Apr 24, 2025
Fulford: Training reasoning models on math and coding generalizes to writing
“So I think in general you will always get a model better, better at a specific task if you train on that task, but we also see a lot of generalization from training on one kind of task to, you know, other domains. So you can train a reasoning model on mostly m…”
May 1, 2025 bullish
Mitchell: AI reasoning improvements will not be limited to math and code
“So like there, I think there's some reason for spikiness, but I think some people will probably go too far with this and saying like, oh yes, these models will only be really good at math and code. And like, not, you know, like everything else is like, you can…”
May 1, 2025 positive
McKinzie: Tools prevent reasoning models from degrading during test-time compute
“We've in the past for our reasoning models talked a lot about test time scaling, and I think for a lot of problems you know, without tools, test time scaling might occasionally work and, but at some point the model is just kind of ranting in its internal chain…”
Oct 9, 2025 neutral