GPT-2

7 statements across 6 episodes · 3 bullish · 2 bearish · 6 people on the record · first statement Jan 2, 2024 by Suhail Doshi · said 73 times in 37 episodes since 2023 · across every show →

Mentions by year

brought up most by Shawn Wang (13), Nathan Lambert (5), Dylan Patel (4), Ashvin Nair (4), Andrej Karpathy (4), Alessio Fanelli (4), Andreas Stuhlmüller (3), William Beauchamp (2)

tap a year for its mentions
0020840152023202420252026episodesmentions
08152023202420252026episodes it came up in
001.37.52.5152023202420252026episodesmentions per episode
2026 9 mentions in 6 episodes 2 per episode
2025 19 mentions in 11 episodes 2 per episode
2024 36 mentions in 15 episodes 2 per episode
2023 9 mentions in 5 episodes 2 per episode

every mention, scene by scene, with the transcript →

Everything said about GPT-2, oldest first

Jan 2, 2024 bearish
Opinion
Doshi: Image Generative AI Is Stuck in a 'GPT-2 Moment'
“So I think that we continue to feel like graphics and these foundation models for anything really related to pixels, but also definitely images continues to be very under invested. It feels a little like graphics is in like this GPT two moment, right? Like eve…”
Suhail Doshi Jan 2, 2024 ▶ 18:07 The AI-First Graphics Editor - with Suhail Doshi of Playground AI
Apr 11, 2024 neutral
Assertion Not checkable as stated
Stuhlmüller: Elicit continues to use T5-based models
“We do also use, like, T-Five-based models, even, even now but started, yeah, started with GPT-II.”
Andreas Stuhlmüller Apr 11, 2024 ▶ 13:18 Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Sep 21, 2024 positive
Assertion Supported
Karpathy: llm.c was 20% faster and used 30% less memory than PyTorch
“At the time of that post, we were using, in LL and that's in 30% less memory, and we were 20% faster in training, just the truth.”
Andrej Karpathy Sep 21, 2024 ▶ 19:10 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Sep 21, 2024 positive
Assertion Supported
Karpathy: llm.c trains GPT-2 on one H100 node in 24 hours for $600
“You can train it on a single node of H-one-hundreds in about 24 hours, and that costs roughly 600 dollars.”
Andrej Karpathy Sep 21, 2024 ▶ 18:02 llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
Jun 19, 2025 negative
What-if
Noam Brown: Reasoning paradigms would have failed on GPT-2
“If you try to do the reasoning paradigm on top of GPT-II, I don't think it would have gotten you almost anything.”
Noam Brown Jun 19, 2025 ▶ 9:35 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
Aug 15, 2025 positive
Assertion Partly supported
Brockman: Arc Institute trained 40B DNA model on 13T base pairs
“I'd say that maybe the neural net we produced, you know, it's a 40 B neural net trained on, you know, like 13 trillion base pairs or something like that. The results to be felt like GPT one, maybe starting to be GPT two level, right? It's like accessible or, a…”
Greg Brockman Aug 15, 2025 ▶ 17:53 Greg Brockman on OpenAI's Road to AGI
Dec 30, 2025 neutral
Opinion
Nair: AI robotics is currently in its 'GPT-1 to GPT-2' era
“Yeah, like I would say that robotics is in kind of like the GPT-one to GPT-two area right now.”
Ashvin Nair Dec 30, 2025 ▶ 4:54 [State of RL/Reasoning] IMO/IOI Gold, OpenAI o3/GPT-5, and Cursor Composer — Ashvin Nair, Cursor
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.