VLLM
product on 6 shows · 14 statements across 9 episodes · said 115 times in 32 episodes since 2023
Latent Space 75
the a16z Podcast 25
the MAD Podcast 11
20VC 2
the Y Combinator Startup Podcast 1
No Priors 1
Mentions by year, every show
Latent Space 75
the a16z Podcast 25
the MAD Podcast 11
20VC 2
No Priors 1
the Y Combinator Startup Podcast 1
2026 45 mentions in 9 episodes 5 per episode
-
How Open Source Became AI's Backbone | Inferact with a16z -
Dylan Patel: NVIDIA's New Moat & Why China is "Semiconductor Pilled” -
Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten -
Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup -
Nebius Co-Founder on AI Infrastructure Bubbles | How Price Elastic is Demand for Compute · 20VC with Harry Stebbings -
⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind -
Multi-GPU Kernels, Intelligence per Watt, Heterogeneous Inference, and More | YC Paper Club · Y Combinator -
The AI Memory Problem: Why Long Context Isn’t Enough — Dan Biderman, Engram Co-founder & CEO - 1 more episode that year, every mention in 2026 →
2025 47 mentions in 11 episodes 4 per episode
-
The Shape of Compute (Chris Lattner of Modular) -
DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing) -
Outlasting Noam Shazeer, Crowdsourcing Chai AI w/ 1.4m DAU — with William Beauchamp, Chai Research -
One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation -
🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R) -
A Technical History of Generative Media -
Dylan Patel on GPT-5’s Router Moment, GPUs vs TPUs, Monetization -
No Priors Ep. 127 | With SemiAnalysis Founder and CEO Dylan Patel - 3 more episodes that year, every mention in 2025 →
2024 22 mentions in 11 episodes 2 per episode
-
[Paper Club] Writing in the Margins: Chunked Prefill KV Caching for Long Context Retrieval -
Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini -
Windsurf: The Enterprise AI IDE -
Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI -
Answer.ai & AI Magic with Jeremy Howard -
A Brief History of the Open Source AI Hacker - with Ben Firshman of Replicate -
Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024] -
In the Arena: How LMSys changed LLM Benchmarking Forever - 3 more episodes that year, every mention in 2024 →
2023 1 mention in 1 episode
every mention on every show, scene by scene, with the transcript →