DeepSeek-V3, every mention
13 scenes · ← back to DeepSeek-V3
tap a year for its mentions
every year anyone Yining Zhang 14Micah Hill-Smith 3Ronak Malde 1Elie Bakouch 1Alessio Fanelli 1
Verbatim, from the transcripts: the passages where DeepSeek-V3 comes up
⚡️Every product of the future will be a living system — Ronak Malde, Trajectory.ai
- ▶ 14:26 Ronak Malde And then all of a sudden, DeepSeq v three comes out.
Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith
- ▶ 25:36 Micah Hill-Smith Well, a couple of weeks, it was, it was Boxing Day in New Zealand, uh, when, when Deep Seek v three came out and I like, we'd been tracking Deep Seek and a bunch of the other global players that were less known. 3 times in the scene
⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF
- ▶ 8:54 Elie Bakouch And for example, a good, uh, a good way to view that is that, uh, DeepSeq rig three is still using the same Adam parameter than, uh, Lama two.
⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
- ▶ 6:45 Alessio Fanelli For me, it was Mistral Small being the best open source model above Foro and DeepSeq VIII.
Outlasting Noam Shazeer, Crowdsourcing Chai AI w/ 1.4m DAU — with William Beauchamp, Chai Research
- ▶ 22:54 unnamed speaker V-three, yeah.
DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing)
- ▶ 0:43 unnamed speaker And, you know, you are lead software engineer on the model performance team, and you guys recently shipped DeepSeq v. three as one of the many models that you do host. 2 times in the scene
- ▶ 1:22 Yining Zhang Yeah, because, uh, DeepSeq VIII, I think, is currently considered the leading open source LLMs based on the benchmark results and the chat area results. 6 times in the scene
- ▶ 4:56 Yining Zhang So before the DeepSig-Fee-Seventy-B, I think we haven't encountered that issue for that so large weights. 4 times in the scene
- ▶ 18:28 unnamed speaker Can you maybe quickly run people through how do you go from taking the deep seek V three weights to like actually run it? 2 times in the scene
- ▶ 26:57 Yining Zhang I think for the common use case, maybe not, not the DeepSeq VIII, for the common use case, I think SGLAN's performance is better than FLM, and its usability is better than TensorFlow TLM. 2 times in the scene
- ▶ 35:50 unnamed speaker And when you think about a model that is, you know, as large as DeepSeq v three, especially, uh, like having better KVcache reutilization is great.
- ▶ 50:03 Yining Zhang When we released the DeepSeq feed story support, we have some community user, something like Cursor. 2 times in the scene
- ▶ 55:36 unnamed speaker I think this is a really good dive into both BaseNet and SG Lang, and a little bit of DeepSig V three, which people are very interested in.