Llama 2 70B, every mention
3 scenes · ← back to Llama 2 70B
tap a year for its mentions
every year anyone Dylan Patel 6Elie Bakouch 1
Verbatim, from the transcripts: the passages where Llama 2 70B comes up
⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF
- ▶ 9:05 Elie Bakouch I don't know if it's just me, but I feel that like the, the, the, the hyperpenters for Lama two AB, for example, shouldn't be the optimal one for, uh, this, uh, mega DeepSeq model with, uh, with a lot of, uh, of parameter.
The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis
- ▶ 16:54 Dylan Patel Llama-seventy-b was two million batch size, and like, you talk to someone at one of the frontier labs, and they're like, ha, right? 4 times in the scene
- ▶ 26:06 Dylan Patel Hey, to run Llama's seventy billion requires two terabytes a second of memory bandwidth, 2.1, at reading, human reading speed. 2 times in the scene