Nvidia
includes NVIDIA H100, NVIDIA A100, NVIDIA Rubin, NVIDIA Blackwell, NVIDIA H200, NVIDIA Dynamo, NVIDIA Cosmos, NVIDIA GTC, NVIDIA Hopper, NVIDIA T4, NVIDIA GB300, NVIDIA A10 and 32 more
80 statements across 35 episodes · 38 bullish · 14 bearish · 34 people on the record · first statement Jun 20, 2023 by George Hotz · said 721 times in 100 episodes since 2023 · across every show →
Mentions by year, the whole family
brought up most by Shawn Wang (80), Dylan Patel (55), Kyle Kranen (43), Ali Taha (26), Ethan He (24), Chris Lattner (24), George Hotz (23), Philip Kiely (20)
2026 306 mentions in 30 episodes 10 per episode
- Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup
- Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
- Inside xAI: Building Grok Imagine in 3 Months, Videogen vs World Models, and Video Agents— Ethan He
- Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
- Dylan Patel Explains the AI War While Cooking | In-Context Cooking
- The Inference Frontier: from 100 to 10,000 tokens per second — Sean Lie, Cerebras CTO
- The AI Frontier: from open weights to open research — Eiso Kant, Poolside AI
- Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith
- 22 more episodes that year, every mention in 2026 →
2025 191 mentions in 33 episodes 6 per episode
- ⚡️Accelerators @ 3x NVIDIA H200 perf, Made in the USA - Thomas Sohmers + Mitesh Agrawal, Positron AI
- The Shape of Compute (Chris Lattner of Modular)
- SF Compute: Commoditizing Compute
- A Technical History of Generative Media
- 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
- How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
- DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing)
- SAM 3: The Eyes for AI — Nikhila & Pengchuan (Meta Superintelligence), ft. Joseph Nelson (Roboflow)
- 25 more episodes that year, every mention in 2025 →
2024 122 mentions in 31 episodes 4 per episode
- [Paper Club] Weight Streaming on Wafer-Scale Clusters (w/ Sarah Chieng of Cerebras)
- Building an open AI company - with Ce and Vipul of Together AI
- State of the Art: Training 70B LLMs on 10,000 H100 clusters
- Why Google failed to make GPT-3 -- with David Luan of Adept
- The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- [Paper Club] SWE-Bench [OpenAI Verified/Multimodal] + MLE-Bench with Jesse Hu
- [LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models
- Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
- 23 more episodes that year, every mention in 2024 →
2023 102 mentions in 6 episodes 17 per episode
- The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis
- Ep 18: Petaflops to the People — with George Hotz of tinycorp
- Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
- FlashAttention-2: Making Transformers 800% faster AND exact
- Why AI Agents Don't Work (yet) - with Kanjun Qiu of Imbue
- RWKV: Reinventing RNNs for the Transformer Era
- every mention in 2023, scene by scene →