GPT-1, every mention
21 scenes · ← back to GPT-1
tap a year for its mentions
every year anyone Shunyu Yao 2Shawn Wang 2Jeremy Howard 2Jack Morris 2Yi Tay 1Spooks (Swyx) 1Nikunj Handa 1Marc Andreessen 1Itamar Friedman 1Greg Brockman 1
Verbatim, from the transcripts: the passages where GPT-1 comes up
🔬 Training Transformers to solve 95% failure rate of Cancer Trials — Ron Alfa & Daniel Bear, Noetik
- ▶ 1:14:50 Brandon Anderson GPT-II, GPT-III, GPT-III, you know, GPT-I, two, and three, like, there was a clear progression there.
Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
- ▶ 6:02 Marc Andreessen But then, uh, um, Alec Radford did GPT-one in, what, probably?
[State of RL/Reasoning] IMO/IOI Gold, OpenAI o3/GPT-5, and Cursor Composer — Ashvin Nair, Cursor
- ▶ 4:54 Ashvin Nair Yeah, like I would say that robotics is in kind of like the GPT-one to GPT-two area right now.
Greg Brockman on OpenAI's Road to AGI
- ▶ 18:02 Greg Brockman The results to be felt like GPT one, maybe starting to be GPT two level, right?
Information Theory for Language Models: Jack Morris
- ▶ 2:20 Jack Morris GPT two, GPT one from open AI were like interesting, but I think most people were into BERT at that time.
- ▶ 1:09:50 Jack Morris And then the second thing was Transformers and BERT, and this Attention is All You Need paper, 2017, the first GPT, 2018, which is web scale pre-training.
Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
- ▶ 32:28 Spooks (Swyx) We had David Luan on, who I think was VP Eng at the time of GPT-I and II.
The new OpenAI Agents Platform: CUA, Web Search, Responses API, Agents SDK!!
- ▶ 18:39 Nikunj Handa It's like the GPT-II of computer use or maybe GPT-I of computer use right now.
0 to over $8M ARR in 2 months as a Claude Wrapper (Bolt.new, Qodo)
- ▶ 54:40 Itamar Friedman For us, like, Alpha Codium one is like GPT one.
[Paper Club] BERT: Bidirectional Encoder Representations from Transformers
- ▶ 22:24 unnamed speaker Um, you can see from this in the, at least when it was released, BERT-Large was state-of-the-art, even beating out, uh, GPT-ONE.
Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI
- ▶ 34:50 Shawn Wang You know, like, I think that, that is the, the core insight of the GPTs, the GPC one, two, three, that was
Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph
- ▶ 2:05 Shunyu Yao So he was actually the second author of GPT-ONE when he was like a visiting scientist at OpenAI. 2 times in the scene
The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- ▶ 19:09 Alessio Fanelli They're kind of like speed running GPT one, GPT two, GPT three in open source.
The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
Breaking down the OG GPT Paper by Alec Radford
- ▶ 0:29 unnamed speaker So with that out of the way, let's get directly to what we want to discuss today, which is like the GPT one paper by the folks at OpenAI.
- ▶ 12:59 unnamed speaker Uh, so let, let's talk a bit about GPT. 6 times in the scene
- ▶ 30:55 unnamed speaker And this, this, like, this defeats, this defeats the whole purpose of the GPT work.
- ▶ 46:37 unnamed speaker Like, let's say you have this dataset that is one trillion tokens or one billion tokens, and you measure perplexity of GPT-I on this dataset, and you can measure the perplexity of LAMA-II on this dataset. 2 times in the scene
- ▶ 50:45 unnamed speaker They did, after the benchmarks, they now have a good model, so they are, they try, they are trying to understand why their model is good, and why their GPT-E-ONE is, is suddenly SOTA. 4 times in the scene
Why Google failed to make GPT-3 -- with David Luan of Adept
- ▶ 7:12 Shawn Wang Um, did you overlap with GPT-I?
The End of Finetuning — with Jeremy Howard of Fast.ai
- ▶ 15:38 Jeremy Howard and he had released GPT, 2 times in the scene