GPU HBM

1 statements across 1 episodes · 0 bullish · 0 bearish · 1 people on the record · first statement Jul 13, 2026 by Dan Biderman · across every show →

Everything said about GPU HBM, oldest first

Jul 13, 2026 neutral
Assertion Partly supported
Biderman: Processing a Wikipedia article in Llama 70B consumes 80GB HBM
“If you take a Lama, a 70 B model, and you load one article from Wikipedia, which is a few tens of kilobytes, and you have the model read this The brain state of the model when reading this few tens of kilobytes is like, 80 gigabytes. 80 gigabytes on, on the HB…”
Dan Biderman Jul 13, 2026 ▶ 22:39 The AI Memory Problem: Why Long Context Isn’t Enough — Dan Biderman, Engram Co-founder & CEO
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.