Isa Fulford leads OpenAI's deep research and ChatGPT agent team on post-training. She explains how chain-of-thought reasoning from RL on STEM problems enables real-world agent capabilities.
“When we saw the reinforcement learning algorithm working really well on math and physics problems and coding problems, It became pretty clear, like, just from reading through the chain of thought, like, okay, this thing's actually, like, thinking and reasoning and backtracking, and to build something that's able to, like, navigate the real world, it also needs to have that ability. So we realized, okay, like, this is a thing that's gonna actually let us get to useful agents.”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Isa Fulford
AssertionContradicted
Isa Fulford: Deep Research was the first AI model to do comprehensive browsing
“Deep Research, it was the first model to do, like, very comprehensive browsing.”
Isa FulfordAug 8, 2025▶ 7:31GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Disclosure
Fulford: OpenAI recycles agent model datasets to train frontier reasoning models
“We're able to take the data sets that we've created for The, you know, frontier agent models and then contribute it back to the frontier reasoning models.”
Isa FulfordAug 8, 2025▶ 7:37GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
AssertionNot checkable as stated
Fulford: ChatGPT agent's browser and terminal access enable most human computer tasks
“The ChatGPT agent, for example, has such a general tool. It has a browser and a terminal, and between those two things, you can basically do most of the tasks that A human does on a computer.”
Isa FulfordAug 8, 2025▶ 16:30GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Insight
Fulford: OpenAI bootstraps browsing models to generate synthetic training data
“For initial deep research, there's not really any data sets that exist for browsing in the same way that you have a math data set that already exists. So we have to create all this data. But once you have good browsing models or good computer use models, you c…”
Isa FulfordAug 8, 2025▶ 31:30GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Insight
Isa Fulford: OpenAI defies startup wisdom by targeting universal users
“I mean, it's like everything they tell you not to do at a startup is just like your user is anyone.”
Isa FulfordAug 8, 2025▶ 11:44GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Insight
Isa Fulford: Reinforcement learning for specific model capabilities is data-efficient
“Training a model to be good at a specific capability is very data efficient. You don't need that many examples to teach it something new.”
Isa FulfordAug 8, 2025▶ 7:16GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 1,000 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.