DeepSeek-R1, every mention
6 scenes · ← back to DeepSeek-R1
tap a year for its mentions
every year anyone Sebastian Raschka 7Mitch Trojanowski 1Jeremy Howard 1
Verbatim, from the transcripts: the passages where DeepSeek-R1 comes up
How to Build Autonomous, Long-Horizon AI Agents | Basis
- ▶ 20:03 Mitch Trojanowski Versus if you look, you know, if you fast forward a bit and you look at the, like the DeepSeq, um, R-one paper, uh, where they effectively laid out, you know, what I think all the labs were doing at that time, or at least OpenAI was doing…
State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
- ▶ 19:55 Sebastian Raschka So RLVR was kind of like popularized by deep seek R one, which was based on deep seek version three. 3 times in the scene
- ▶ 26:03 Sebastian Raschka And so my statement that it is not so promising or useful was mainly based on the R one paper where they had a final paragraph at the bottom.
- ▶ 31:06 Sebastian Raschka Because if you also look at the numbers of how much it costs, uh, just GPU hours deep seek version three, they had like a five million dollar price tag on that given the, I think two dollars per GPU, they assume whether that's a correct… 2 times in the scene
- ▶ 37:29 Sebastian Raschka I mean, if you look at a version three and then R one, and then they had version 3.2 model with the sparse attention mechanism, and then also this math version two with a self refinement and everything.
Jeremy Howard on Building 5,000 AI Products with 14 People (Answer AI Deep-Dive)
- ▶ 33:42 Jeremy Howard Um, like it helps, like, understand the details of the technology well, so I kind of, these things like DeepSea Car One or whatever don't seem, they don't come out of the blue, they don't seem like wild jumps or whatever, you know, I can,…