Reasoning Models

topic on 13 shows · 66 statements across 53 episodes

the Y Combinator Startup Podcast BG2 Pod We Live to Build the Knowledge Project Latent Space the Neon Show No Priors WTF is with Nikhil Kamath Invest Like the Best the MAD Podcast the a16z Podcast Big Technology TBPN

The latest 60 statements about Reasoning Models, every show

NEON SHOW Assertion Not checkable as stated
Sankar: Agentic AI use cases took off after reasoning models dropped
“All kinds of agentic use cases, which really started becoming a thing after the reasoning model dropped.”
Prukalpa Sankar Sep 10, 2026 ▶ 1:50 Why Startups Fails Even After Finding PMF | Prukalpa Sankar, Atlan
a16z Assertion Not checkable as stated
George says post-reasoning AI models drove an adoption takeoff for Harvey
“Now, fast forward, post reasoning models, like that totally flipped. And you could see, you know, absolute takeoff of adoption, right? And so a bunch of different things happen at the same time. Lawyers got way more value out of the product. You could see it i…”
David George Sep 10, 2026 ▶ 24:28 Why Investors Are Rethinking Everything for the AI Era
a16z Insight
Sawhney: Backtracking is a general-purpose reasoning tool, not math-specific
“A lot of these behaviors that we're describing mathematically, like backtracking, or kind of starting again, I mean, these are not really specific to mathematics. I mean, we're seeing them specifically in mathematics in these examples, but kind of, they're gen…”
Mehtaab Sawhney Sep 8, 2026 ▶ 13:26 Inside OpenAI’s Breakthroughs in Mathematical Reasoning
MAD Insight
Kolter: Reasoning models are much harder to jailbreak via probability optimization
“Reasoning models were much more effective because you can't really do the same trick of optimizing for a probability with a reasoning model that has a whole trace of reasoning that happens in the middle and kind of reflect a bit more. So it's much harder to br…”
Zico Kolter May 7, 2026 ▶ 49:51 OpenAI Board Member Zico Kolter: Modern AI Is Just 200 Lines of Code
MAD Insight
Kolter: Major AI breakthroughs require both massive scale and luck
“Reasoning models were the next big breakthrough. Those are rare. They do take kind of a, you know, both, both a massive scale and kind of a bit of luck to get there.”
Zico Kolter May 7, 2026 ▶ 1:07:44 OpenAI Board Member Zico Kolter: Modern AI Is Just 200 Lines of Code
WE LIVE TO BUILD Assertion Not checkable as stated
May: Reasoning AI tactics increase token usage 20% of the time
“Sometimes when you try to use like one of these reasoning models where you add chain of thought or one of these tactics, they'll actually use more tokens about 20% of the time. It'll be more expensive.”
Rob May Apr 10, 2026 ▶ 19:13 Why the 7th Biggest Player in a Huge Market Beats the Winner in a Small One
Lopopolo: Reasoning models eliminate need for rigid state-machine scaffolding
“And this I think is like the fundamental difference between reasoning models and the four ones and four O's of the past where these models could not think. So you kind of had to put them in boxes with a predefined set of state transitions. Whereas here we have…”
Ryan Lopopolo Apr 7, 2026 ▶ 11:59 Extreme Harness Engineering: 1M LOC, 1B toks/day, 0% human code or review — Ryan Lopopolo, OpenAI
Bose: Enterprise context grows more valuable as AI reasoning models improve
“The real value we're providing, again, is with the enterprise-weight context and the shared memory. And so that becomes instantly more valuable as the reasoning model gets better.”
Arnab Bose Apr 7, 2026 ▶ 27:44 Should Software Companies Embrace AI or fight it? — With Asana Chief Product Officer Arnab Bose
NO PRIORS Opinion
Fedus: Reasoning and coding agents connect software AI to physical domains
“And I think those were foundational technologies necessary To then connect these systems to the physical world. Like it was just not impossible, not possible with like the AI technology of.”
Liam Fedus Apr 3, 2026 ▶ 6:24 AI for Atoms: How Periodic Labs is Revolutionizing Materials Engineering with Co-Founder Liam Fedus
BG2 Opinion
Turley: ChatGPT reasoning currently serves only power users but is transformative
“When you look at reasoning in ChatGPT today, it's irrelevant to a very small group of people. It's relevant for the people who are trying to get the most out of ChatGPT, but I fundamentally believe that reasoning, it's transformative”
Nick Turley Mar 15, 2026 ▶ 12:55 ChatGPT – The Super Assistant Era | BG2 Guest Interview · Bg2 Pod
BG2 Assertion Not checkable as stated
Turley: Early OpenAI reasoning model swore after making a mistake
“We were showing this chain of thought as it was streaming out of the model and the model swore and said, like, oh, damn it, may have to adjust, because it realized it had made a mistake in the puzzle. And the fact that it did that, But in particular, the fact …”
Nick Turley Mar 15, 2026 ▶ 1:01:05 ChatGPT – The Super Assistant Era | BG2 Guest Interview · Bg2 Pod
MAD Disclosure
LeCroix: Reasoning models are a major priority for Mistral AI
“Reasoning is a big priority”
Timothée LeCroix Feb 12, 2026 ▶ 45:23 Mistral AI vs. Silicon Valley: The Rise of Sovereign AI
Pineau: Reasoning models fail at multi-tier hierarchical planning
“That's the part that the reasoning models don't do. They do really well at like one level of granularity... But the going back and forth between different levels of sort of resolution of action, it's really hard. So on the technical terms, we call it hierarchi…”
Joelle Pineau Feb 6, 2026 ▶ 15:55 AI's Research Frontier: Memory, World Models, & Planning — With Joelle Pineau
LATENT SPACE Assertion Supported
Reasoning models consume 10x more tokens on average than non-reasoning models
“So, earlier this year, and probably when you and George last spoke for the AI engineers world's fair, we had this great slide that was super easy, where we would show that the average reasoning model is using 10 times the number of tokens per query in our inte…”
Micah Hill-Smith Jan 9, 2026 ▶ 1:06:27 Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith
AI progress would have completely stalled in 2024 without reasoning models
“Had reasoning not come along, there would have been no AI progress from mid 24 Through, essentially, Gemini three. There would have been none. Everything would have stalled.”
Gavin Baker Dec 9, 2025 ▶ 8:31 GPUs, TPUs, & The Economics of AI Explained | Gavin Baker Interview · Invest Like The Best
MAD Insight
Kaiser: Reasoning models are the second major milestone after Transformers
“One point was, of course, the Transformers when it started, but the other point was reasoning models.”
Łukasz Kaiser Nov 26, 2025 ▶ 3:53 What’s Next for AI? OpenAI’s Łukasz Kaiser (Transformer Co-Author)
MAD Disclosure
Kaiser: OpenAI began working on reasoning models around three years ago
“So we started working on it maybe three years ago”
Łukasz Kaiser Nov 26, 2025 ▶ 4:06 What’s Next for AI? OpenAI’s Łukasz Kaiser (Transformer Co-Author)
MAD Insight
Łukasz Kaiser: Reasoning models require verifiable data, excelling in math and coding
“So currently, and current for at least the Most basic ways we use it currently, it needs to be fairly verifiable. So there is an, is your answer correct or not? You prepare data for that. You can do that in mathematics, coding very well. You can do this in sci…”
Łukasz Kaiser Nov 26, 2025 ▶ 13:49 What’s Next for AI? OpenAI’s Łukasz Kaiser (Transformer Co-Author)
MAD Insight
Kaiser: Test-time compute increases AI capabilities faster than pre-training
“Using more tokens to think increases your capability, and it increases it, given the computation, way faster than pre-training, right?”
Łukasz Kaiser Nov 26, 2025 ▶ 46:49 What’s Next for AI? OpenAI’s Łukasz Kaiser (Transformer Co-Author)
Lenz: Most enterprises avoid reasoning models due to high latency
“Most enterprises don't really want to use reasoning models. The latencies is too high”
Barak Lenz Oct 11, 2025 ▶ 33:15 Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21
NO PRIORS Insight
Zelikman: Training reasoning models on just positive examples causes a plateau
“So if you only train on like the positive examples, then you end up in this kind of like potential minimum where there's just no more data that it can actually solve.”
Eric Zelikman Oct 9, 2025 ▶ 5:44 No Priors Ep. 135 | With Humans& Founder Eric Zelikman
a16z Assertion Not checkable as stated
Altman: OpenAI continues achieving fundamental breakthroughs in deep learning and reasoning
“And deep learning has been this miracle that keeps on giving, and we have kept finding, like, breakthrough after breakthrough. Again, when we got the reasoning model breakthrough, like, I also thought that was like, we're never gonna get another one like that.…”
Sam Altman Oct 8, 2025 ▶ 12:20 Sam Altman on Sora, Energy, and Building an AI Empire
Dwivedi: Deterministic workflows cannot solve complex enterprise incident debugging
“No amount of workflows will suffice for a big enterprise. Like you have to link together some of the missing pieces, some of the poorly instrumented data, and that requires world knowledge and a few iterations with the world knowledge.”
Raaz Dwivedi Oct 5, 2025 ▶ 32:39 ⚡️Traversal: Causal ML and Reinforcement Learning
a16z Assertion Not checkable as stated
Patel: Reasoning models are driving a surge in AI inference demand
“Inference demand has been skyrocketing this year, right? These reasoning models, the revenue it's been skyrocketing this year”
Dylan Patel Sep 22, 2025 ▶ 1:37:11 Dylan Patel on the AI Chip Race - NVIDIA, Intel & the US Government vs. China
Swix: Reasoning Models Have Better Context Utilization Than Standard LLMs
“I have a theory also that reasoning models have better context utilization because they can loop back. Normal auto-aggressive models, they just kind of go left to right, but reasoning models, in theory, they can loop back and look for things that they needed c…”
Shawn Wang Aug 19, 2025 ▶ 18:28 Long Live Context Engineering - with Jeff Huber of Chroma
BIG TECHNOLOGY Assertion Not checkable as stated
Lightcap: Most free ChatGPT users haven't used reasoning models yet
“Most of them have actually not experienced the power of the reasoning models. They mostly are using GPT-IV-O and, you know, they mostly are kind of using it for this very kind of you know, turn-based kind of like very quick you know, back and forth, almost sea…”
Brad Lightcap Aug 8, 2025 ▶ 17:08 OpenAI COO Brad Lightcap: GPT-5's Capabilities, Why It Matters, and Where AI Goes Next
a16z Prediction Not checkable as stated
Dwarkesh: Reasoning models will outperform GPT-4o on real-world deductive tasks
“I think a reasoning model, I think a reasoning model will be more reliable and be better at solving that kind of problem than Poirot.”
Dwarkesh Patel Aug 4, 2025 ▶ 48:25 Dwarkesh Patel and Noah Smith on AGI and the Economy
Lambert: The RL algorithm is not the most important component in reasoning models
“I definitely don't think the algorithm tends to be the most important thing.”
Nathan Lambert Jul 31, 2025 ▶ 19:28 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Lambert: North star of reasoning models is dynamic token budget calibration
“I think that has to be the north star for most people working on reasoning, which is the model will just Spend the right amount of tokens on it.”
Nathan Lambert Jul 31, 2025 ▶ 20:44 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Nadella: Reasoning models paired with human synthesis are the true frontier
“Now, like having a sophisticated reasoning model and your prefrontal cortex work together whereas a lot of the mundane Stuff is getting done by even some core agent or what have you. That I think is definitely the frontier.”
Satya Nadella Jun 25, 2025 ▶ 15:16 Satya Nadella: Microsoft's AI Bets, Hyperscaling, Quantum Computing Breakthroughs · Y Combinator
Brown: Deep Research proves reasoning models work in unverifiable domains
“And that is very clearly a domain where you don't have an easily verifiable metric for success. It's very like, what is the best research report that you could generate? And yet these models are doing extremely well at this domain. So I think that's like an ex…”
Noam Brown Jun 19, 2025 ▶ 7:32 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
TBPN Insight
Chen: AI reasoning only emerges at large scale requiring massive compute
“When you look at reasoning you just don't see that happen at small scale, right? There's like a certain scale at which it starts becoming signal bearing and that requires you to have resources, right?”
Mark Chen Jun 7, 2025 ▶ 1:22:48 Weekly Recap - Elon Vs Trump, Ukraine's Drone Attack, Cluely Update & OpenAI CRO
MAD Insight
Gomez: Creating AI reasoning models is dramatically cheaper than pre-training
“It's easy to create a reasoning model. It's dramatically cheaper than pre-training. And so it's accessible. And so there's this huge intelligence uplift that comes for really quite little effort.”
Aidan Gomez Jun 5, 2025 ▶ 17:57 Inside the Paper That Changed AI Forever - Cohere CEO Aidan Gomez on 2025 Agents
MAD Prediction Not checkable as stated
Gomez: AI reasoning models will expand into medicine and physical sciences
“We've just scratched the surface at the moment. It's mostly focused on, you know math problems and this sort of thing. There is a whole world of applications that we need to make it work work in medicine, you know, everything from the pure sciences, physics, c…”
Aidan Gomez Jun 5, 2025 ▶ 18:18 Inside the Paper That Changed AI Forever - Cohere CEO Aidan Gomez on 2025 Agents
Will Brown: AI reasoning models are merely a stepping stone toward autonomous agents
“The thing that's going to make the next wave of stuff be powerful is just, like, everyone wants better agents. Everyone wants models that can, like, go off and do stuff. And, like, reasoning was kind of, like, a precursor to that a little bit.”
Will Brown May 23, 2025 ▶ 1:29 ⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
Reasoning models acting as reward models are key to agent RL
“And the most, one of the most promising ways, I think, towards doing this is having the reward models also be able to answer harder questions by themselves being reasoning models.”
Will Brown May 9, 2025 ▶ 11:54 ⚡️Open Questions in Agentic RL — Will Brown (Prime Intellect)
Marcus: 'Reasoning' AI models only mimic patterns without genuine abstractions
“Now, the reason I wouldn't call them reasoning models, though you're right that many people do, is what I think they're doing is basically copying patterns of human reasoning. They're getting data about how humans reason certain things, but the depth of reason…”
Gary Marcus May 7, 2025 ▶ 11:55 Are We at the End of Ai Progress? — With Gary Marcus
NO PRIORS Insight
McKinzie: Tools prevent reasoning models from degrading during test-time compute
“We've in the past for our reasoning models talked a lot about test time scaling, and I think for a lot of problems you know, without tools, test time scaling might occasionally work and, but at some point the model is just kind of ranting in its internal chain…”
Brandon McKinzie May 1, 2025 ▶ 3:45 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
NO PRIORS Opinion
Mitchell: AI reasoning improvements will not be limited to math and code
“So like there, I think there's some reason for spikiness, but I think some people will probably go too far with this and saying like, oh yes, these models will only be really good at math and code. And like, not, you know, like everything else is like, you can…”
Eric Mitchell May 1, 2025 ▶ 31:16 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
LATENT SPACE Assertion Supported
Factorio benchmark results show reasoning models underperform expectations on extended planning
“One thing we have found in preliminary results is that the reasoning models don't seem to do as well as you'd expect in this setting. And I think that's probably because the way we set this up, it's a bit like we're already making it do reasoning traces over a…”
Jack Hopkins Apr 27, 2025 ▶ 12:12 ⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins
NO PRIORS Insight
Fulford: Training reasoning models on math and coding generalizes to writing
“So I think in general you will always get a model better, better at a specific task if you train on that task, but we also see a lot of generalization from training on one kind of task to, you know, other domains. So you can train a reasoning model on mostly m…”
Isa Fulford Apr 24, 2025 ▶ 7:23 No Priors Ep. 112 | With OpenAI Deep Research, Isa Fulford
Patel: Generating pre-answer reasoning tokens yields superior AI performance
“Models now will think for some time before they answer. And this enables much better performance on all sorts of tasks, whether it be coding or math or understanding science or understanding complex Social dilemmas, right? All sorts of different topics they're…”
Dylan Patel Apr 23, 2025 ▶ 26:02 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Taylor: Reasoning models generate net new ideas, breaking the data wall
“What's really interesting about, you know, reasoning and reasoning models is I think I feel really optimistic these models are generating net new ideas, and so it really affords the opportunity to break through some of these, the data wall as well.”
Bret Taylor Apr 15, 2025 ▶ 45:37 Bret Taylor: A Vision for AI’s Next Frontier
Pokrass: Pair reasoning models for planning with smaller models for execution
“I do think reasoning models for planning and using kind of more targeted models to execute is definitely a good architecture.”
Michelle Pokrass Apr 15, 2025 ▶ 28:48 GPT 4.1: The New OpenAI Workhorse
GPT-4.1 excels at exploring repositories, while reasoning models dominate targeted file changes
“Basically, where GPT, 4.1, can it kind of explore, go through a repo? It's been trained to do that particularly well. Whereas you know, to just get some code and produce a change, a reasoning model might do better because it can kind of reason over the entire …”
Michelle Pokrass Apr 15, 2025 ▶ 31:20 GPT 4.1: The New OpenAI Workhorse
LATENT SPACE Assertion Supported
OpenAI currently restricts reinforcement fine-tuning exclusively to its reasoning models
“No, that's reinforcement fine tuning is only for reasoning models.”
Michelle Pokrass Apr 15, 2025 ▶ 39:32 GPT 4.1: The New OpenAI Workhorse
BIG TECHNOLOGY Prediction Not checkable as stated
Suleyman: Reasoning models taking desktop action will be cheap and abundant within 15 years
“That's the transition that's going to happen over the next 15 years is that it is going to be a cheap and basically abundant resource to have these reasoning models that can take action in your workplace, that can orchestrate your apps and get things done for …”
Mustafa Suleyman Apr 4, 2025 ▶ 47:59 Microsoft AI CEO Mustafa Suleyman: Building AI Personality
BIG TECHNOLOGY Assertion Supported
Hendrycks: Recent AI reasoning models score in 90th percentile on wet-lab guidance
“We are finding that with the most recent reasoning models quite unlike the models from two years ago, like the initial GPT-IV, the most recent reasoning models are getting around 90th percentile compared to these expert level virologists in their area of exper…”
Dan Hendrycks Mar 28, 2025 ▶ 11:16 AI's Rising Risks: Hacking, Virology, Loss of Control — With Dan Hendrycks
WTF Prediction Not checkable as stated
Srinivas: Many reasoning models will exist, but few products will integrate personal context well
“There's gonna be a bunch of great reasoning models, but there's not gonna be a hundred products that really package, personal context all the API integration, services integrations native integration to your phone to be an assistant really well.”
Aravind Srinivas Mar 23, 2025 ▶ 1:10:02 Nikhil Kamath ft. Perplexity CEO, Aravind Srinivas | WTF Online Ep 1. · Nikhil Kamath
a16z Prediction Not checkable as stated
Acharya: Voice plus reasoning models will rapidly eliminate unwanted hallucinations
“This is where I think the capability is just going to get better and better faster than we appreciate. You know, with the language models, they're prone to hallucination, and there are certain conversations like the therapy one that benefit from the hallucinat…”
Anish Acharya Mar 18, 2025 ▶ 23:46 Why AI Voice Feels More Human Than Ever
Shankar: Evals are necessary to train AI reasoning models
“You need evals to train your reasoning models.”
Shreya Shankar Mar 13, 2025 ▶ 27:13 [Lightning Pod] Evals: How to Improve AI Consistently — with Hamel Husain and Shreya Shankar
NO PRIORS Prediction Not checkable as stated
Dohmke: Improved model reasoning will push SWE-bench scores near 100%
“As the models get better in reasoning we're going to get closer to a hundred percent of this VBench, which is that benchmark out of 12 repos open source Python repos a team in Princeton identified 2200 or so issue pull request pairs. Effectively, all the model…”
Thomas Dohmke Mar 13, 2025 ▶ 2:25 No Priors Ep 106 | With GitHub CEO Thomas Dohmke
a16z Assertion Supported
Appenzeller: Reasoning models now dominate top AI model rankings
“If you look at the slide here that shows the current ranking of one of the best AI models that we have today, you'll see that pretty much the whole top of the rankings has been taken over by reasoning models.”
Guido Appenzeller Mar 5, 2025 ▶ 0:51 DeepSeek, Reasoning Models, and the Future of LLMs
a16z Assertion Not checkable as stated
Appenzeller: Switching to reasoning models would increase inference compute needs 20x
“Very roughly, if we all, if everybody would switch tomorrow from whatever they have today to a reasoning model, we would need 20 times more inference.”
Guido Appenzeller Mar 5, 2025 ▶ 22:30 DeepSeek, Reasoning Models, and the Future of LLMs
NO PRIORS Assertion Not checkable as stated
Hendrycks: Reasoning models have reached expert-level virology capabilities
“The AIs are getting very good at STEM PhD level types of topics, and that includes virology. So I think that they are sort of rounding the corner on being able to provide expert level capabilities in terms of their knowledge of the literature, Or even helping …”
Dan Hendrycks Mar 5, 2025 ▶ 5:43 No Priors Ep. 105 | With Director of the Center of AI Safety Dan Hendrycks
LATENT SPACE Prediction Not checkable as stated
Kilpatrick: Reasoning will solve multi-item retrieval in long context
“And like, it feels like, again, like back to this, the thread around these capabilities, like it feels like long context with reasoning is like finally going to be that thing where like, it actually just like blows the lid off of it. And like, it makes the use…”
Logan Kilpatrick Feb 28, 2025 ▶ 11:43 Gemini 2.0 Flash and Flash Thinking: the new SOTA models for the agentic era
Chen: AI models cannot learn reasoning from scratch without pre-trained knowledge
“You need knowledge in order to build reasoning on top of it. Right. a model can't kind of go in blind and just learn reasoning from scratch. So we find these two paradigms to be fairly complementary and we think, you know, they have feedback loops on each oth…”
Mark Chen Feb 27, 2025 ▶ 4:45 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
Chen: GPT-4.5 outshines reasoning models like o1 in creative writing
“And, you know, we find that in a lot of areas like creative writing, for instance Again, this is stuff that we want to test over the next one or two months but we find that there are areas like creative writing where this model outshines reasoning models.”
Mark Chen Feb 27, 2025 ▶ 6:35 OpenAI's Chief Research Officer on GPT 4.5's Debut, Scaling Laws, And Teaching EQ to Models
LATENT SPACE Prediction Not checkable as stated
Roucher: Visual AI models will likely jump the reasoning S-curve in 2025
“But as with text agents, we've really found that we made a jump on the S curve with reasoning models. I think it's going to be the same with the next visual models, basically better base models just allow you to jump over this S curve. And probably I think it'…”
Aymeric (Emmerich) Feb 13, 2025 ▶ 21:21 smol agents are all you need
Karina Nguyen: Verification difficulty makes alignment crucial for reasoning models
“The question of like alignment is actually more important for this like complex reasoning models to like, how do we help humans to like verify the outputs of these models is quite important.”
Karina Nguyen Feb 1, 2025 ▶ 21:41 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.