Llama 2 70B, every mention

3 scenes · ← back to Llama 2 70B

tap a year for its mentions
003161202320242025episodesmentions
011202320242025episodes it came up in
0030.561202320242025episodesmentions per episode

every year anyone Dylan Patel 6Elie Bakouch 1

Verbatim, from the transcripts: the passages where Llama 2 70B comes up

loading…

⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF Oct 20, 2025 · 1 mention

  • ▶ 9:05 Elie Bakouch I don't know if it's just me, but I feel that like the, the, the, the hyperpenters for Lama two AB, for example, shouldn't be the optimal one for, uh, this, uh, mega DeepSeq model with, uh, with a lot of, uh, of parameter.

The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis Dec 5, 2023 · 6 mentions

  • ▶ 16:54 Dylan Patel Llama-seventy-b was two million batch size, and like, you talk to someone at one of the frontier labs, and they're like, ha, right? 4 times in the scene
  • ▶ 26:06 Dylan Patel Hey, to run Llama's seventy billion requires two terabytes a second of memory bandwidth, 2.1, at reading, human reading speed. 2 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.