GPT-5

includes GPT 5 Codex, GPT 5 Pro

13 statements across 4 episodes · 8 bullish · 3 bearish · 4 people on the record · first statement Aug 8, 2025 by Christina Kim · said 109 times in 21 episodes since 2025 · across every show →

Mentions by year, the whole family

brought up most by Erik Torenberg (20), Nathan Labenz (12), Olivia Moore (11), Sherman Wu (9), Christina Kim (7), Isa Fulford (6), Mark Chen (5), Marc Andreessen (5)

tap a year for its mentions
0050101002020252026episodesmentions
0102020252026episodes it came up in
0031062020252026episodesmentions per episode
2026 9 mentions in 4 episodes 2 per episode
2025 100 mentions in 17 episodes 6 per episode

every mention, scene by scene, with the transcript →

Everything said about GPT-5, oldest first

Aug 8, 2025 neutral
Insight
Kim: Real-world usage will replace saturated benchmarks to measure AI progress
“I feel like we've almost saturated a lot of these evals, and the real, like, metric of, like, how good our models are getting is, I think, gonna be, like, usage, right?”
Christina Kim Aug 8, 2025 ▶ 9:15 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 8, 2025 bullish
Prediction Not checkable as stated
Kim: AI prompt-based app generation will spur surge in indie businesses
“I think we're just gonna have a lot more, I would expect, like, maybe a lot more, like, indie type of, like, Businesses built around this because of the fact that, like, you just need to have the idea, write a simple prompt, and then you get the full fledged a…”
Christina Kim Aug 8, 2025 ▶ 8:26 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 8, 2025 positive
Assertion Not checkable as stated
Kim: GPT-5 internal testers felt insulted by instant answers to hard questions
“I think we hear this with GPT-Five internally when people are testing and they're like, oh, I thought I asked like a really hard question. I feel like a little bit insulted that I thought for like two seconds or like when it doesn't even want to think at all.”
Christina Kim Aug 8, 2025 ▶ 0:27 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 8, 2025 positive
Opinion
Kim: GPT-5's creative writing capability is tender and touching
“That's one of my favorite improvements in GBT five. The writing, I honestly find it's very tender and touching, especially for a lot of the creative writing that we want to do.”
Christina Kim Aug 8, 2025 ▶ 17:04 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 8, 2025 bullish
Opinion
Kim: The leap from GPT-4 to GPT-5 is OpenAI's most impressive yet
“Maybe I'm biased, recency biased, but I think to jump to four to five is most impressive for me, because I guess with 3.5 when we first released it, the most common use case for me then also was still just for coding. And, but now, like, Even though four was b…”
Christina Kim Aug 8, 2025 ▶ 20:07 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 8, 2025 positive
Disclosure
Kim: GPT-5 is a step change for personal coding and writing
“I use it for coding and writing all the time, and it's just a huge stuff change.”
Christina Kim Aug 8, 2025 ▶ 2:13 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 8, 2025 positive
Opinion
Kim: GPT-5 front-end coding is a massive leap over o3
“If you compare it to O three's front end coding capability, this is just totally next level.”
Christina Kim Aug 8, 2025 ▶ 3:43 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
Aug 13, 2025 neutral
Insight
Benchmark-topping AI models are not necessarily what consumers want for chat
“I don't necessarily think the like smartest model that scores the best on sort of all of these objective benchmarks of intelligence will be the model that people want to chat with.”
Justine Moore Aug 13, 2025 ▶ 8:45 This Week in AI: GPT-5 Ships, 4o Pulled Back, Grok Imagine Goes Social
Aug 13, 2025 positive
Assertion Supported
OpenAI's GPT-5 scored highest on the physician-trained HealthBench medical benchmark
“They talked about how GPT-V was kind of the highest scoring model on this thing called HealthBench, which is a benchmark they trained with like, 250 plus physicians. To measure how good an LLM is at answering medical questions.”
Justine Moore Aug 13, 2025 ▶ 11:44 This Week in AI: GPT-5 Ships, 4o Pulled Back, Grok Imagine Goes Social
Oct 14, 2025 negative
Assertion Not checkable as stated
Labenz: OpenAI's router failure caused bad initial GPT-5 outputs
“The problem at launch was that that router was broken. So all of the queries were going to the dumb model, and so a lot of people literally just got Bad outputs, which were worse than oh three because they were getting non thinking responses.”
Nathan Labenz Oct 14, 2025 ▶ 21:35 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
Oct 14, 2025 positive
Opinion
Labenz: GPT-4 to GPT-5 capability leap matches GPT-3 to GPT-4
“And if you look back to GPT three, you know, there's a huge leap. I would contend that the leap is similar from GPT four to five.”
Nathan Labenz Oct 14, 2025 ▶ 5:03 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
Oct 23, 2025 bearish
Opinion
Masad: GPT-5 shows no reasoning progress on open-ended controversial topics
“Go you know, dig up GPT-IV or other models and go to GPT-V. You're not gonna find that much difference of, okay, let's reason together. Let's try to figure out what was the origins of COVID. Because it's still an unanswered question, you know? And I don't see …”
Amjad Masad Oct 23, 2025 ▶ 43:37 Marc Andreessen & Amjad Masad on “Good Enough” AI, AGI, and the End of Coding
Oct 23, 2025 negative
Opinion
Masad: GPT-5 regressed in human tone compared to GPT-4
“My feeling is that you know, GPT-Five got good at verifiable domains. It didn't feel that much better at anything else. The more human angle of it felt like it regressed”
Amjad Masad Oct 23, 2025 ▶ 41:39 Marc Andreessen & Amjad Masad on “Good Enough” AI, AGI, and the End of Coding
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.