FlashAttention
also referred to as: flash attention
7 statements across 4 episodes · 5 bullish · 1 bearish · 4 people on the record · first statement Aug 3, 2023 by Tri Dao · said 53 times in 17 episodes since 2023 · across every show →
Mentions by year
brought up most by Diego Bachman (9), Quentin Anthony (6), Tri Dao (5), Chris Lattner (3), Alessio Fanelli (3), Shawn Wang (2), Jeremy Howard (2), Andrej Karpathy (2)
2025 19 mentions in 4 episodes 5 per episode
- ⚡️ Beyond Transformers with Power Retention
- How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
- The Shape of Compute (Chris Lattner of Modular)
- ⚡️Mercury: Ultra-Fast Diffusion LLMs — Estefano Ermon, CEO Inception Labs
- every mention in 2025, scene by scene →
2024 14 mentions in 9 episodes 2 per episode
- Building an open AI company - with Ce and Vipul of Together AI
- llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE
- Answer.ai & AI Magic with Jeremy Howard
- The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
- 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
- [Paper Club] Upcycling Large Language Models into Mixture of Experts
- Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph
- How to train a Million Context LLM — with Mark Huang of Gradient.ai
- 1 more episode that year, every mention in 2024 →