TensorRT-LLM, every mention

2 scenes (2026) · ← back to TensorRT-LLM

tap a year for its mentions
00132253202420252026episodesmentions
023202420252026episodes it came up in
0041.583202420252026episodesmentions per episode

every year 2026 anyone Yining Zhang 10Kyle Kranen 2Chris Lattner 2Stefano Ermon 1Ben Firshman 1Ali Taha 1

Verbatim, from the transcripts: the passages where TensorRT-LLM comes up

loading…

Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten Aug 3, 2026 · 1 mention

  • ▶ 23:58 Ali Taha Like oftentimes, um, the image you run, like NVIDIA will release an image, for instance, and if we will upstream the changes from the latest RTLM image into our stack, we'll find that it, it fixes it.

Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup Mar 8, 2026 · 2 mentions

  • ▶ 28:05 Kyle Kranen Dynoa sort of came about at NVIDIA because myself and a couple others were sort of talking about these concepts that like, you know, you have inference engines like VLM, SGLang, TensorRTLM, um, 2 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.