Mamba, every mention
23 scenes across 5 shows · ← back to Mamba
tap a year for its mentions
Latent Space 30
the MAD Podcast 4
No Priors 4
20VC 1
the a16z Podcast 1
every year every show
Latent Space 30
No Priors 4
the MAD Podcast 4
the a16z Podcast 1
20VC 1
Verbatim, from the transcripts: passages where Mamba comes up on Latent Space, No Priors, the MAD Podcast, the a16z Podcast, 20VC
Inside Nemotron & NVIDIA’s AI Lab | Bryan Catanzaro
- ▶ 39:37 Matt Turck So it's a combination of transformer and member state space, which is a slightly more exotic, uh, form of, of, of architecture.
AI-Native Engineering: 100% adoption, 5x search throughput, unlimited tokens — Mikhail Parakhin
- ▶ 1:05:46 Mikhail Parakhin Transformer, like in Mamba fashion, they probably the best architecture I'm aware of, like, period.
State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
- ▶ 3:09 Sebastian Raschka At the same time, you have other alternatives popping up, like, you know, diffusion models as text diffusion, uh, particular or Mamba models, state space models and so forth.
How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
- ▶ 1:12 unnamed speaker I think some people might be familiar with some of the like Mamba adjacent models that you released in the past, but how should people think about what you work on and the current research direction you're in?
- ▶ 22:42 Quentin Anthony So like it's very, Grok is very inflexible hardware so that you kind of, if you want to do like a Mamba SSM on it, you're going to have a really hard time because instead of having like a low level CUDA compiler, like everything is like
- ▶ 58:55 Quentin Anthony That's why we jumped on Mamba. 2 times in the scene
Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21
- ▶ 5:42 Barak Lenz But then as we were doing the, the ablations, I, I came across the Mamba paper and it actually was pointed out to me by, by several people. 8 times in the scene
The AI Agenda: GPT5 leaks and the business of AI News — Steph Palazzolo, The Information
- ▶ 1:10:54 Shawn Wang There are some notable exceptions, primarily together with the Mombard architecture, um, recursal with RWKV.
Information Theory for Language Models: Jack Morris
- ▶ 1:09:01 Jack Morris That's like, for whatever reason, like the kind of most glamorous thing people think you can do as a researcher, like Mamba, it's like,
Building the Next Generation of Conversational AI
- ▶ 1:08:03 Ankit Kumar There are other kind of architectures that are slowly gaining in some popularity, Mamba and SSMs and so on.
2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
- ▶ 11:34 Dan Fu Um, if you've heard of the Mamba model, this, this also takes this selection to, to the next level by actually making some changes in that, uh, fundamental recurrent state space.
- ▶ 24:50 Eugene Cheah And, and I believe together AI worked on the law cats for, for the Mamba side of things, and, and we took some ideas from there as well, and we essentially did that for RWKV. 2 times in the scene
AGI, The Future of AI Agents And The Next Wave of Opportunities in AI | Richard Socher, CEO, You.com
- ▶ 25:58 Richard Socher Ja, es gibt die Mamba-Samba-Modelle und Neural-State-Space-Modelle, die sehr interessiert sind. 2 times in the scene
The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- ▶ 19:37 unnamed speaker I will call out, um, you know, they, they have some interesting experimentations with Mamba, and Mistral Nemo is actually on the efficiency frontier chart that I drew that is still relevant.
Ethan Mollick: Why OpenAl Abandons Products, The Biggest Opportunities They Have Not Taken | E1184 · 20VC with Harry Stebbings
- ▶ 10:28 Ethan Mollick Let's say that LLM's top out, and it turns out we have to switch to, you know, Mamba or some other, like, you know, uh, other, like, who cares?
No Priors Ep. 70 | With Cartesia Co-Founders Karan Goel & Albert Gu
- ▶ 0:07 Sarah Guo We're excited to talk to Karen Gohl and Albert Gu, the co-founders of Cartesia, and authors behind such revolutionary models as S-For and Mamba.
- ▶ 2:24 Albert Gu Recently, uh, proposed a model called Mamba, um, which was, uh, kind of brought these to language modeling, um, and showed really good results there.
- ▶ 8:00 Albert Gu Newer versions of these models like, um, Mamba, which was the most recent one that, that's been out for a few months, that's a lot better at modeling the same types of data as transformers.
- ▶ 12:02 Albert Gu So, um, I actually just heard from some collaborators today that they applied, um, a mama based model, um, on DNA modeling.
A Comprehensive Overview of Large Language Models - Latent Space Paper Club
- ▶ 26:50 unnamed speaker Cuma sedikit ekstra, ini adalah salah satu topik yang sangat biasa yang mereka ingin memulai semasa mereka masuk keadaan seperti model Mamba anda.
Building an open AI company - with Ce and Vipul of Together AI
- ▶ 42:54 Vipul Ved Prakash We know that this, you know, system B is better for, uh, Mixtrol, and system C is going to be better for Stripe Tine or Mamba.
- ▶ 55:37 Shawn Wang Um, and, uh, I mean, first of all, I'll throw that open to you, but second of all, I think what Mamba did for me was change that perception of that it's only about a long context. 5 times in the scene
The Four Wars of the AI Stack - Dec 2023 Recap
- ▶ 35:25 unnamed speaker And then, obviously, uh, RWKV and Mamba and Stripe, Stripe Tyena from, um, together. 4 times in the scene