DeepSeek-V3, every mention

13 scenes · ← back to DeepSeek-V3

tap a year for its mentions
0013225420252026episodesmentions
02420252026episodes it came up in
00326420252026episodesmentions per episode

every year anyone Yining Zhang 14Micah Hill-Smith 3Ronak Malde 1Elie Bakouch 1Alessio Fanelli 1

Verbatim, from the transcripts: the passages where DeepSeek-V3 comes up

loading…

⚡️Every product of the future will be a living system — Ronak Malde, Trajectory.ai Jun 21, 2026 · 1 mention

Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith Jan 9, 2026 · 3 mentions

  • ▶ 25:36 Micah Hill-Smith Well, a couple of weeks, it was, it was Boxing Day in New Zealand, uh, when, when Deep Seek v three came out and I like, we'd been tracking Deep Seek and a bunch of the other global players that were less known. 3 times in the scene

⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF Oct 20, 2025 · 1 mention

  • ▶ 8:54 Elie Bakouch And for example, a good, uh, a good way to view that is that, uh, DeepSeq rig three is still using the same Adam parameter than, uh, Lama two.

⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo Jul 14, 2025 · 1 mention

  • ▶ 6:45 Alessio Fanelli For me, it was Mistral Small being the best open source model above Foro and DeepSeq VIII.

Outlasting Noam Shazeer, Crowdsourcing Chai AI w/ 1.4m DAU — with William Beauchamp, Chai Research Jan 26, 2025 · 1 mention

DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing) Jan 19, 2025 · 20 mentions

  • ▶ 0:43 unnamed speaker And, you know, you are lead software engineer on the model performance team, and you guys recently shipped DeepSeq v. three as one of the many models that you do host. 2 times in the scene
  • ▶ 1:22 Yining Zhang Yeah, because, uh, DeepSeq VIII, I think, is currently considered the leading open source LLMs based on the benchmark results and the chat area results. 6 times in the scene
  • ▶ 4:56 Yining Zhang So before the DeepSig-Fee-Seventy-B, I think we haven't encountered that issue for that so large weights. 4 times in the scene
  • ▶ 18:28 unnamed speaker Can you maybe quickly run people through how do you go from taking the deep seek V three weights to like actually run it? 2 times in the scene
  • ▶ 26:57 Yining Zhang I think for the common use case, maybe not, not the DeepSeq VIII, for the common use case, I think SGLAN's performance is better than FLM, and its usability is better than TensorFlow TLM. 2 times in the scene
  • ▶ 35:50 unnamed speaker And when you think about a model that is, you know, as large as DeepSeq v three, especially, uh, like having better KVcache reutilization is great.
  • ▶ 50:03 Yining Zhang When we released the DeepSeq feed story support, we have some community user, something like Cursor. 2 times in the scene
  • ▶ 55:36 unnamed speaker I think this is a really good dive into both BaseNet and SG Lang, and a little bit of DeepSig V three, which people are very interested in.
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.