Deep Research

9 statements across 6 episodes · 7 bullish · 0 bearish · 7 people on the record · first statement Feb 18, 2025 by Arush Sehgal · said 32 times in 8 episodes since 2025 · across every show →

Mentions by year

brought up most by Nathan Lambert (9), Josh McGrath (6), Bret Taylor (4), Shawn Wang (1), Paul Klein (1)

tap a year for its mentions
002044082025episodesmentions
0482025episodes it came up in
0024482025episodesmentions per episode
2025 32 mentions in 8 episodes 4 per episode

every mention, scene by scene, with the transcript →

Everything said about Deep Research, oldest first

Feb 18, 2025 positive
Insight
Sehgal: AI deep research provides most lift on niche, non-Wikipedia topics
“We love to test, like, super niche random things, like, things where there's, like, No Wikipedia page already about this topic or something like that, right? Because that's where you'll see the most lift from a feature like this.”
Arush Sehgal Feb 18, 2025 ▶ 2:59 Why is everyone cloning Deep Research?
Feb 18, 2025 positive
Insight
Sridhar: Deep Research targets multi-tab exploratory queries rather than direct searches
“There are things that, you know exactly what you're looking for and their search is still probably, you know, a very, you know, probably one of the best places to go. I think where deep research really shines is that, like, Multiple facets to your question, an…”
Mukund Sridhar Feb 18, 2025 ▶ 2:27 Why is everyone cloning Deep Research?
Feb 18, 2025 neutral
Insight
Sridhar: Multi-minute AI research creates unique UX alignment and web navigation challenges
“This is one of the first times, you know, something takes about five, six minutes trying to perform your research, so there's a few challenges that brings, like, you want to make sure you're spending that time in the computer doing what the user wants, so ther…”
Mukund Sridhar Feb 18, 2025 ▶ 1:43 Why is everyone cloning Deep Research?
Jun 19, 2025 positive
Insight
Brown: Deep Research proves reasoning models work in unverifiable domains
“And that is very clearly a domain where you don't have an easily verifiable metric for success. It's very like, what is the best research report that you could generate? And yet these models are doing extremely well at this domain. So I think that's like an ex…”
Noam Brown Jun 19, 2025 ▶ 7:32 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
Jul 23, 2025 positive
Assertion Not checkable as stated
McCloy: Enterprise users are outsourcing B2B software vendor bake-offs to ChatGPT
“One specific example that takes people a little bit by surprise is, like, there are a huge number of people doing, like, B to B software bake-offs in chat to VT. You know, if you're at a big company and your boss is like, Hey, we need to buy something to solve…”
Robert McCloy Jul 23, 2025 ▶ 34:05 AI is Eating Search
Jul 31, 2025 neutral
Opinion
Lambert: Deep Research relies on modular RL tasks rather than end-to-end outcomes
“I think the deep research blog post kind of hints that they do a bunch of small scale RL and then poof, the system works. Which I think is much more of what's happening is people train on a bunch of small things and they do some prompting and they see that whe…”
Nathan Lambert Jul 31, 2025 ▶ 7:35 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
Oct 18, 2025 bullish
Prediction Not checkable as stated
Merrill: Frontier AI labs will center operations around vertical products
“And with the Claude codes and the codec CLIs and the deep researchers, researchers, you starting to see some evidence that the products are going to be a much more central part of how these frontier labs operate.”
Mike Merrill Oct 18, 2025 ▶ 26:19 Terminal-Bench: Pushing Claude Code, OpenAI Codex, Factory Droid, et al to the limits
Dec 31, 2025 positive
Prediction Not checkable as stated
McGrath: Specialized and Frontier Reasoning AI Models Will Eventually Converge
“You know, I think if you look at like deep research, the original one and GPT-Five thinking on like high reasoning today, I think you'll see that like eventually the models all sort of converge in their capabilities.”
Josh McGrath Dec 31, 2025 ▶ 6:13 [State of Post-Training] From GPT-4.1 to 5.1: RLVR, Agent & Token Efficiency — Josh McGrath, OpenAI
Dec 31, 2025 positive
Assertion Supported
McGrath: GPT-5 Thinking Matches or Beats Deep Research on Published Evals
“I mean, I think if you look at our published evals, they're, they look, like, basically on par if it's not better, so, like, I mean, that's personally what I do.”
Josh McGrath Dec 31, 2025 ▶ 6:46 [State of Post-Training] From GPT-4.1 to 5.1: RLVR, Agent & Token Efficiency — Josh McGrath, OpenAI
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.