Llama
also referred to as: llama models
includes Llama 3, Llama 2, Llama 4, Llama 3.1, Llama 1, Llama 70B, Llama 405B, Llama Stack, Llama 3.2, Llama 3.3, Llama 7B, Llama Guard and 5 more
23 statements across 21 episodes · 7 bullish · 7 bearish · 20 people on the record · first statement Dec 5, 2023 by Dylan Patel · said 455 times in 94 episodes since 2023 · across every show →
Mentions by year, the whole family
brought up most by Shawn Wang (63), Alessio Fanelli (35), Thomas Scialom (22), Nathan Lambert (21), Soumith Chintala (15), George Hotz (13), Yining Zhang (9), Lin Qiao (7)
2026 28 mentions in 15 episodes 2 per episode
- 🔬 "The Most Innovative Diffusion Research Is Happening in Drug Discovery, Not Image Generation"
- Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
- The AI Frontier: from open weights to open research — Eiso Kant, Poolside AI
- AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
- Mistral: Voxtral TTS, Forge, Leanstral, & Mistral 4 — w/ Pavan Kumar Reddy & Guillaume Lample
- Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup
- The AI Memory Problem: Why Long Context Isn’t Enough — Dan Biderman, Engram Co-founder & CEO
- The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO
- 7 more episodes that year, every mention in 2026 →
2025 90 mentions in 28 episodes 3 per episode
- DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing)
- Better Data is All You Need — Ari Morcos, Datology
- 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
- SAM 3: The Eyes for AI — Nikhila & Pengchuan (Meta Superintelligence), ft. Joseph Nelson (Roboflow)
- ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
- The Utility of Interpretability — Emmanuel Amiesen
- The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
- The Magic of LLM Distillation — Rishabh Agarwal, Google DeepMind
- 20 more episodes that year, every mention in 2025 →
2024 299 mentions in 43 episodes 7 per episode
- Training Llama 2, 3 & 4: The Path to Open Source AGI — with Thomas Scialom of Meta AI
- [LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models
- Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
- The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
- The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
- How to train a Million Context LLM — with Mark Huang of Gradient.ai
- A Brief History of the Open Source AI Hacker - with Ben Firshman of Replicate
- 35 more episodes that year, every mention in 2024 →
2023 38 mentions in 8 episodes 5 per episode
- Ep 18: Petaflops to the People — with George Hotz of tinycorp
- FlashAttention-2: Making Transformers 800% faster AND exact
- RAG is a hack - with Jerry Liu of LlamaIndex
- The End of Finetuning — with Jeremy Howard of Fast.ai
- The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis
- Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
- RWKV: Reinventing RNNs for the Transformer Era
- Why AI Agents Don't Work (yet) - with Kanjun Qiu of Imbue
- every mention in 2023, scene by scene →