Opinion certainty 3/5 debate potential 3/5

Mitchell: AI reasoning improvements will not be limited to math and code

Eric Mitchell · No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie · May 1, 2025 · at 31:16

OpenAI research scientist Eric Mitchell addresses whether post-training and RL produce narrow, domain-specific capability spikes rather than general intelligence.

0:00 / 0:18exact quote · 18.8s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“So like there, I think there's some reason for spikiness, but I think some people will probably go too far with this and saying like, oh yes, these models will only be really good at math and code. And like, not, you know, like everything else is like, you can't get better at them. And I think that is probably not the right intuition to have.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Eric Mitchell

Insight
Mitchell: LLMs allocate compute efficiently by deferring tasks to specialized tools
“I think, like, part of this is you can just allocate compute a lot more efficiently because you can defer stuff that the model doesn't have comparative advantage to doing to a tool that is, like, really well suited to doing that thing.”
Eric Mitchell Oct 31, 2025 ▶ 12:28 No Priors Ep. 138 | The Best of 2025 (So Far) with Sarah Guo and Elad Gil
Assertion Supported
Mitchell: o3 autonomously executes multi-step tasks using integrated tools
“Not only is the model it's on its own smarter than our previous O series models, which is great, but it's also able to use all these tools that like further enhance its abilities and whether that's doing like research on something where you want up-to-date inf…”
Eric Mitchell May 1, 2025 ▶ 2:03 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
Disclosure
Mitchell: OpenAI plans to unify models and remove the ChatGPT switcher
“You know, I think for us, like unification of our models is something that, you know, Sam has talked about publicly that, you know, we have this big crazy model switcher in ChatGPT and there are a lot of choices and you know, we have a model that might be good…”
Eric Mitchell May 1, 2025 ▶ 5:24 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
Insight
Mitchell: OpenAI limits model agency due to asymmetric error costs
“There's a reason we don't go hog wild and say, like, oh yes, here's, like, the keys to the kingdom, like, have at it. There are still, you know, asymmetric costs to, like, the time you can save and the types of errors you can make, and so we're trying to, like…”
Eric Mitchell May 1, 2025 ▶ 17:47 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
Insight
Mitchell: Physical time bottlenecks make AI tasks harder than simulatable domains
“Stuff that is really bottlenecked by like time, like the physical world is also, you know, just harder than stuff that we can simulate really well.”
Eric Mitchell May 1, 2025 ▶ 21:15 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
Insight
Mitchell: High-quality evaluation benchmarks are underappreciated compared to training data
“I mean, yeah, like you want, you know, good data to train on and that's of course valuable for making the model better, but I think it is often neglected how also important it is to have high quality data, which is like a different definition of high quality w…”
Eric Mitchell May 1, 2025 ▶ 32:58 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.