Prediction Held up AI assessment confidence: 95% certainty 3/5 debate potential 2/5

Patel: Meta's next Llama model will match DeepSeek-V3's cost efficiency

Dylan Patel · Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel · Apr 23, 2025 · at 30:27

Dylan Patel, CEO of SemiAnalysis, discusses market reactions to DeepSeek-V3 and anticipates Meta's next open-source Llama release.

0:00 / 0:10exact quote · 10.6s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“And Meta's Meta is going to release their new llama soon enough. Right. And that one is going to be, you know, a similar level of cost decrease probably similar areas, deep seek V three.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Dylan Patel

Opinion
Patel: Most AI models lean left due to Bay Area origins
“Most AI models are made in the Bay Area, so they tend to just be left leaning, right? But also the internet in general is a little bit left leaning because it skews younger than older.”
Dylan Patel Apr 23, 2025 ▶ 19:44 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Insight
Patel: Building cheaper AI requires massive frontier models for synthetic data
“You can't actually make that cheaper model without making the better model, bigger model. So you can generate data to help you make the cheaper model, right?”
Dylan Patel Apr 23, 2025 ▶ 33:21 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Assertion Not checkable as stated
Patel: $10B AI data centers aim to automate software engineering, not chatbots
“No one is trying to make with these, you know, with these ten billion dollar data centers, they're not trying to make chat models, right? They're not trying to make models that people chat with, just to be clear, right? They're trying to solve things like soft…”
Dylan Patel Apr 23, 2025 ▶ 34:22 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Assertion Not checkable as stated
Patel: OpenAI's Orion training run failed to reach GPT-5 performance levels
“There were hopes that Orion could be used for GPT-V but its improvement was, like, not enough to be, like, really a GPT-V. Furthermore, it was trained on the classical method, which is, like which is a ton of pre-training, and then some reinforcement learning …”
Dylan Patel Apr 23, 2025 ▶ 36:12 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Opinion
Patel: Language is a representation for reasoning, not human thought itself
“Language is not actually how our brain thinks. It's just a representation for which it to, you know, reason over.”
Dylan Patel Apr 23, 2025 ▶ 4:58 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Insight
Patel: AI models ingest dangerous data during pre-training for world knowledge
“So you don't want to just filter out everything so that the model doesn't know anything about it but at the same time, you don't want it to output, you know, how to build a bomb so there's like a fine balance here, and that's why pre-training is defined as pre…”
Dylan Patel Apr 23, 2025 ▶ 12:57 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.