Nvidia B200, every mention
4 scenes · ← back to Nvidia B200
tap a year for its mentions
every year anyone Ali Taha 2Sean Lie 1Dylan Patel 1
Verbatim, from the transcripts: the passages where Nvidia B200 comes up
The Inference Frontier: from 100 to 10,000 tokens per second — Sean Lie, Cerebras CTO
Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
- ▶ 38:45 Ali Taha Like, if you, if you obviously have a thing where you're serving it on just, like, a node of H-one-hundreds, and then you throw, like, you know, you short the model across, like, four nodes of B-to-hundreds.
- ▶ 55:08 Ali Taha Like on a B 200 is one 80 gigawatts per GPU, and then a node of eight, you're talking like one 80 times eight.
Dylan Patel Explains the AI War While Cooking | In-Context Cooking
- ▶ 42:46 Dylan Patel It was a large GPU, um, and it was like having, it was like the best memory, the best networking, everything, sort of the best as possible, um, and sort of like one size fits all, uh, with the main line of like A-one hundred, H-one…