Llama 3

part of Llama includes Llama 3 70B

8 statements across 3 episodes · 3 bullish · 2 bearish · 3 people on the record · first statement Apr 22, 2024 by Meta's Generative AI Head · said 30 times in 10 episodes since 2024 · across every show →

Mentions by year, the whole family

brought up most by Alex Kantrowitz (18), Andrew Bosworth (5), Meta's Generative AI Head (3), Aaron Levie (2), Dylan Patel (1), Aravind Srinivas (1)

tap a year for its mentions
00134258202420252026episodesmentions
048202420252026episodes it came up in
002448202420252026episodesmentions per episode

every mention, scene by scene, with the transcript →

Everything said about Llama 3, oldest first

Apr 22, 2024 positive
Assertion Supported
Meta trained Llama 3 8B and 70B models on 15 trillion tokens
“So if you look at something like the eight billion and seventy billion, They were trained on almost 15 trillion tokens and tokens roughly you can imagine as a word. So roughly like 15 trillion words, which is an incredible outcome.”
Meta's Generative AI Head Apr 22, 2024 ▶ 5:43 Meta's Generative AI Head: How We Trained Llama 3
Apr 22, 2024
Assertion Partly supported
Meta used 100x more compute to train Llama 3 than Llama 2
“So actually, I think it's I believe it's a hundred times more compute.”
Meta's Generative AI Head Apr 22, 2024 ▶ 10:04 Meta's Generative AI Head: How We Trained Llama 3
Apr 22, 2024 negative
Insight
Ahmad Al-Dahle: AI models degrade user experience by over-moralizing refusals
“Some of these models, for example tend to do a lot of moralization or like really take a perspective or a point of view. And we worked on and I'm continuing to work on and innovate on how, how the model responds and how it refuses, which I think is also part o…”
Meta's Generative AI Head Apr 22, 2024 ▶ 15:22 Meta's Generative AI Head: How We Trained Llama 3
Apr 22, 2024 neutral
Assertion Not checkable as stated
Ahmad Al-Dahle: Llama 3's performance exactly matched Meta's scaling law predictions
“I don't think anything about the model has really personally surprised me in terms of its performance. I think we kind of expected to be here. You know, we do a lot of like rigorous scaling laws and rigorous prediction of what we think the metrics will look li…”
Meta's Generative AI Head Apr 22, 2024 ▶ 27:55 Meta's Generative AI Head: How We Trained Llama 3
Apr 22, 2024 bullish
Disclosure
Meta releases 8-billion and 70-billion parameter Llama 3 models
“We are releasing an updated eight billion parameter model plus a seventy billion parameter model. And these are state of the art.”
Meta's Generative AI Head Apr 22, 2024 ▶ 0:09 Meta's Generative AI Head: How We Trained Llama 3
Apr 22, 2024 positive
Disclosure
Ahmad Al-Dahle: Meta is currently training a 400-billion parameter Llama 3 model
“We also are talking a little bit about one of the larger models that we're training that is already achieving, you know exceptional performance which is a, it's a model that's over four hundred billion parameters.”
Meta's Generative AI Head Apr 22, 2024 ▶ 9:07 Meta's Generative AI Head: How We Trained Llama 3
May 15, 2024 neutral
Assertion Supported
Meta used AI-generated synthetic data to train its Llama 3 model
“Like to train Lama three meta use synthetic data, like data created from basically, you know, from AI itself”
Alex Kantrowitz May 15, 2024 ▶ 22:22 AI Scaling, Alignment, and the Path to Superintelligence — With Dwarkesh Patel
Jul 8, 2026 negative
Disclosure
Bosworth: Meta emptied its AI research pipeline to build Llama 3
“When we were pulling Llama three together, we had really pulled in all the research, all the every, we pulled out every single stop we had and unwittingly kind of killed the pipeline. So researchers, you know, the way that it works is you build a base And you'…”
Andrew Bosworth Jul 8, 2026 ▶ 1:57 Meta CTO Andrew Bosworth: Our Path To Frontier AI, Renting Models, Consumer AI's Struggles
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.