vLLM

3 statements across 1 episodes · 3 bullish · 0 bearish · 2 people on the record · first statement Aug 5, 2026 by Simon Moe · said 25 times in 4 episodes since 2023 · across every show →

Mentions by year

brought up most by Simon Moe (12), Elena Burger (9), Matt Bornstein (1), Ion Stoica (1), Dylan Patel (1), Arthur Mensch (1)

tap a year for its mentions
001312522023202420252026episodesmentions
0122023202420252026episodes it came up in
001312522023202420252026episodesmentions per episode
2026 22 mentions in 1 episode
2025 2 mentions in 2 episodes 1 per episode
2023 1 mention in 1 episode

every mention, scene by scene, with the transcript →

Everything said about vLLM, oldest first

Aug 5, 2026 positive
Assertion Supported
Mo: Major chipmakers use vLLM as an internal benchmark
“And additionally, VLM also work closely with all the hardware vendors. So that means across like NVIDIA, AMD, Google, and Amazon, Intel, and a lot more, their newest chip will make sure VLM can run on them. And then a lot of cases they use VLM as a benchmark t…”
Simon Moe Aug 5, 2026 ▶ 9:29 How Open Source Became AI's Backbone | Inferact with a16z
Aug 5, 2026 bullish
Assertion Partly supported
Simon Mo: vLLM supports over 1,000 model architectures
“For VRM, we support more than a thousand model architecture up to today, and a lot of those are proprietary, but also a lot of those are open-weight, right?”
Simon Moe Aug 5, 2026 ▶ 8:54 How Open Source Became AI's Backbone | Inferact with a16z
Aug 5, 2026 positive
Assertion Not checkable as stated
Burger: vLLM Runs on 500,000 GPUs at Any Moment
“Today we're here with Simon Moe, co-founder of Infraact, and a lead maintainer of VLLM, the open source inference engine, now running on half a million GPUs at any moment.”
Elena Burger Aug 5, 2026 ▶ 1:01 How Open Source Became AI's Backbone | Inferact with a16z
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.