cuDNN, every mention

4 scenes · ← back to cuDNN

tap a year for its mentions
00214220242025episodesmentions
01220242025episodes it came up in
00214220242025episodesmentions per episode

every year anyone Andrej Karpathy 4Quentin Anthony 1Chris Lattner 1

Verbatim, from the transcripts: the passages where cuDNN comes up

loading…

How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony Nov 3, 2025 · 1 mention

  • ▶ 11:31 Quentin Anthony So I would use something like, um, uh, like QDNN or something like that with some, uh, gem backend.

The Shape of Compute (Chris Lattner of Modular) Jun 13, 2025 · 1 mention

llm.c's Origin and the Future of LLM Compilers - Andrej Karpathy at CUDA MODE Sep 21, 2024 · 4 mentions

  • ▶ 14:24 Andrej Karpathy It turns out CodeEman has a very good flash retention implementation, so we switch to that.
  • ▶ 18:17 Andrej Karpathy Uh, so, uh, you do need QDNN, which is the most heavy dependency, but QDNN is optional, so if you'd like to roll your own manual attention, that is possible in LLM.C, but QDNN is kind of like the hairiest dependency, but after that it's… 3 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.