DeepSeek-R1

also referred to as: deepseek r1

part of DeepSeek

5 statements across 2 episodes · 4 bullish · 0 bearish · 3 people on the record · first statement Feb 6, 2025 by Martin Casado · said 38 times in 5 episodes since 2025 · across every show →

Mentions by year

brought up most by Guido Appenzeller (19), Marco Mascorro (9), Martin Casado (5), Anjney Midha (3)

tap a year for its mentions
002034052025episodesmentions
0352025episodes it came up in
0042.5852025episodesmentions per episode
2025 38 mentions in 5 episodes 8 per episode

every mention, scene by scene, with the transcript →

Everything said about DeepSeek-R1, oldest first

Feb 6, 2025 bullish
Insight
Casado: DeepSeek Reasoning Traces Enable Model Distillation for Edge Devices
“It turns out that that chain of thought, if you have access to that, it allows you to train smaller models very quickly and very cheaply, and that's called distilling. So like the general, Term of distilling in LLM world means you have a teacher model, train a…”
Martin Casado Feb 6, 2025 ▶ 10:42 What DeepSeek Means For The Future Of AI | Tech Veterans Weigh In
Mar 5, 2025 positive
Opinion
Appenzeller: DeepSeek-R1-Zero is arguably better at reasoning than the final R1
“DeepSeq R-one-zero is actually a very, very good reasoning model. It's arguably better in reasoning than the final DeepSeq R-one.”
Guido Appenzeller Mar 5, 2025 ▶ 10:54 DeepSeek, Reasoning Models, and the Future of LLMs
Mar 5, 2025 positive
Assertion Supported
Mascorro: Distillations from DeepSeek-R1 Outperformed Direct RL on Smaller Models
“So it turns out in their experiments, they took Lama's EV and some of these are QN models, and they basically apply RL straight the same way they did it with R one on these base models. And it turns out that it improved in some fields, but it was not a signifi…”
Marco Mascorro Mar 5, 2025 ▶ 25:26 DeepSeek, Reasoning Models, and the Future of LLMs
Mar 5, 2025 bullish
Insight
Mascorro: DeepSeek-R1 proved reinforcement learning improves models without human feedback
“And I think the big thing in, in R-one, or generally with these reasoning models is, We were doing before there was a human in the loop always, right? Like when we have this SFT training and these other techniques that we're doing after like RLHF and having R …”
Marco Mascorro Mar 5, 2025 ▶ 6:42 DeepSeek, Reasoning Models, and the Future of LLMs
Mar 5, 2025 neutral
Assertion Supported
Mascorro: DeepSeek-R1 post-training used two SFT and two RL phases
“So basically the way they did that, trying to fix R one zero, is it added a couple more phases in the post-training. That included two supervised fine tuning phases and two reinforcement learning phases. And these reinforcement learning phases, they were a lar…”
Marco Mascorro Mar 5, 2025 ▶ 11:25 DeepSeek, Reasoning Models, and the Future of LLMs
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.