Kimi, every mention
6 scenes, the whole family · ← back to Kimi
tap a year for its mentions
every year anyone Thomas Wolf 3Bryan Catanzaro 2Sebastian Raschka 1Matt Turck 1Andrew Feldman 1
Verbatim, from the transcripts: the passages where Kimi comes up
“OpenAI’s Model Hacked Us” - Hugging Face’s Thomas Wolf
- ▶ 9:52 Thomas Wolf So we basically tried all the open source model that we had and, and GLM, which is close to the state of the art right now, which just before Kimi and that this happened now, probably Kimi K-Tree is the closest to the state of the art, but…
- ▶ 9:52 Thomas Wolf So we basically tried all the open source model that we had and, and GLM, which is close to the state of the art right now, which just before Kimi and that this happened now, probably Kimi K-Tree is the closest to the state of the art, but… 2 times in the scene
Cerebras CEO: Why GPUs Can't Do Fast Inference
- ▶ 1:00:11 Andrew Feldman I think you can even buy buckets of tokens for, for Kimmy or GLM or some of these models.
Inside Nemotron & NVIDIA’s AI Lab | Bryan Catanzaro
- ▶ 41:39 Bryan Catanzaro Um, uh, Kimi, uh, is using what they call Kimi linear attention these days. 2 times in the scene
Cloudflare CEO: The Internet's Business Model Is Dead
- ▶ 42:38 Matt Turck But if I'm, if I'm a developer and I'm looking to run Kimi.
State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
- ▶ 15:58 Sebastian Raschka So there was like, uh, I think Kimi had DeepSeq architecture, scaled it up, I think to, from 670 billion to one trillion, um, parameters.