May 24, 2025 · 32m · big-technology

Google DeepMind CTO: Advancing AI Frontier, New Reasoning Methods, Video Generation’s Potential

Koray Kavukcuoglu · 21m spoken Alex Kantrowitz · 7m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

In this episode of the Big Technology Podcast at Google I/O, Google DeepMind CTO Koray Kavukcuoglu discusses the research roadmap toward AGI, detailing advances in multimodal foundation models, test-time reasoning with DeepThink, and generative video systems.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Alex holds 27.1% of the talking time here. How this is scored →

Alex as informed peer 4.9 Guest teaching 4.6 Guest disagreement 2.0 Alex pushing back 3.8
05100:0010:0020:0030:001:16–7:04 · Alex as informed peer 6/10 Google DeepMind's Unified AGI Mission and Research Pillars Kantrowitz presses on whether scaling laws are hitting diminishing returns, citing industry-wide delays across OpenAI, Meta, and Anthropic. Kavukcuoglu provides a balanced breakdown of how compute scale interacts with algorithmic architecture, data quality, and inference techniques.7:04–9:13 · Alex as informed peer 4/10 Model Progress Across the Industry and the Gemini Roadmap Kavukcuoglu reframes the scaling discussion around the tripartite nature of scaling laws (data, compute, parameters) rather than single-dimension scaling. Kantrowitz listens as the guest outlines DeepMind's progress from Gemini 1.5 to 2.5.9:13–12:18 · Alex as informed peer 6/10 Debating AGI Pathways and Yann LeCun's Critique Kantrowitz challenges Kavukcuoglu with Yann LeCun's assertion that scaling LLMs cannot achieve human-level intelligence. Kavukcuoglu counters by calling it an unproven hypothesis and noting that no serious frontier lab is solely scaling LLMs without broader architectural innovations.12:18–16:05 · Alex as informed peer 5/10 DeepThink and Parallel Reasoning in Gemini 2.5 Pro Kantrowitz asks about the DeepThink announcement and inference-time compute. Kavukcuoglu clarifies that DeepThink is a reasoning mode rather than a standalone product, educating the host on parallel hypothesis exploration versus single chain-of-thought processing.16:05–18:44 · Alex as informed peer 5/10 Native Multimodality and Continuous Frontier Velocity Kantrowitz asks if frontier model improvement velocity is slowing down compared to earlier leaps like GPT-3 to GPT-4. Kavukcuoglu disputes any perception of slowdown, arguing that native multimodality and unified reasoning architectures are accelerating overall capability.18:44–22:18 · Alex as informed peer 4/10 Quantifying Model Improvement and User Feedback Kantrowitz poses a hypothetical question about the real-world value of 10% or 50% model improvements. Kavukcuoglu reframes the question around non-linear evaluation metrics and feedback loops between research and product deployments.22:18–26:24 · Alex as informed peer 5/10 Generative Video Breakthroughs: Veo 3, Physics, and Flow Kantrowitz highlights Google's Veo 3 and Flow video tools, noting synchronized audio and storyboard capabilities. Kavukcuoglu explains the technical milestones from physical object dynamics in Veo 2 to orthogonal audio-visual modeling in Veo 3.26:24–30:20 · Alex as informed peer 6/10 Open Weights vs. Closed Frontier Model Governance Kantrowitz asks how Google navigates the tension between open-source contributions and proprietary frontier models like Gemini, referencing DeepSeek and transformer history. Kavukcuoglu articulates Google's dual strategy of releasing Gemma open weights while restricting Gemini weights for safety and governance.30:20–31:33 · Alex as informed peer 3/10 The Potential of 'Vibe Coding' for Accessibility Kantrowitz ends on a light note asking about vibe coding. Kavukcuoglu enthusiastically endorses the concept, explaining how it democratizes software creation for non-technical users.1:16–7:04 · Guest teaching 4/10 Google DeepMind's Unified AGI Mission and Research Pillars Kantrowitz presses on whether scaling laws are hitting diminishing returns, citing industry-wide delays across OpenAI, Meta, and Anthropic. Kavukcuoglu provides a balanced breakdown of how compute scale interacts with algorithmic architecture, data quality, and inference techniques.7:04–9:13 · Guest teaching 5/10 Model Progress Across the Industry and the Gemini Roadmap Kavukcuoglu reframes the scaling discussion around the tripartite nature of scaling laws (data, compute, parameters) rather than single-dimension scaling. Kantrowitz listens as the guest outlines DeepMind's progress from Gemini 1.5 to 2.5.9:13–12:18 · Guest teaching 6/10 Debating AGI Pathways and Yann LeCun's Critique Kantrowitz challenges Kavukcuoglu with Yann LeCun's assertion that scaling LLMs cannot achieve human-level intelligence. Kavukcuoglu counters by calling it an unproven hypothesis and noting that no serious frontier lab is solely scaling LLMs without broader architectural innovations.12:18–16:05 · Guest teaching 6/10 DeepThink and Parallel Reasoning in Gemini 2.5 Pro Kantrowitz asks about the DeepThink announcement and inference-time compute. Kavukcuoglu clarifies that DeepThink is a reasoning mode rather than a standalone product, educating the host on parallel hypothesis exploration versus single chain-of-thought processing.16:05–18:44 · Guest teaching 4/10 Native Multimodality and Continuous Frontier Velocity Kantrowitz asks if frontier model improvement velocity is slowing down compared to earlier leaps like GPT-3 to GPT-4. Kavukcuoglu disputes any perception of slowdown, arguing that native multimodality and unified reasoning architectures are accelerating overall capability.18:44–22:18 · Guest teaching 5/10 Quantifying Model Improvement and User Feedback Kantrowitz poses a hypothetical question about the real-world value of 10% or 50% model improvements. Kavukcuoglu reframes the question around non-linear evaluation metrics and feedback loops between research and product deployments.22:18–26:24 · Guest teaching 4/10 Generative Video Breakthroughs: Veo 3, Physics, and Flow Kantrowitz highlights Google's Veo 3 and Flow video tools, noting synchronized audio and storyboard capabilities. Kavukcuoglu explains the technical milestones from physical object dynamics in Veo 2 to orthogonal audio-visual modeling in Veo 3.26:24–30:20 · Guest teaching 5/10 Open Weights vs. Closed Frontier Model Governance Kantrowitz asks how Google navigates the tension between open-source contributions and proprietary frontier models like Gemini, referencing DeepSeek and transformer history. Kavukcuoglu articulates Google's dual strategy of releasing Gemma open weights while restricting Gemini weights for safety and governance.30:20–31:33 · Guest teaching 2/10 The Potential of 'Vibe Coding' for Accessibility Kantrowitz ends on a light note asking about vibe coding. Kavukcuoglu enthusiastically endorses the concept, explaining how it democratizes software creation for non-technical users.1:16–7:04 · Guest disagreement 2/10 Google DeepMind's Unified AGI Mission and Research Pillars Kantrowitz presses on whether scaling laws are hitting diminishing returns, citing industry-wide delays across OpenAI, Meta, and Anthropic. Kavukcuoglu provides a balanced breakdown of how compute scale interacts with algorithmic architecture, data quality, and inference techniques.7:04–9:13 · Guest disagreement 2/10 Model Progress Across the Industry and the Gemini Roadmap Kavukcuoglu reframes the scaling discussion around the tripartite nature of scaling laws (data, compute, parameters) rather than single-dimension scaling. Kantrowitz listens as the guest outlines DeepMind's progress from Gemini 1.5 to 2.5.9:13–12:18 · Guest disagreement 4/10 Debating AGI Pathways and Yann LeCun's Critique Kantrowitz challenges Kavukcuoglu with Yann LeCun's assertion that scaling LLMs cannot achieve human-level intelligence. Kavukcuoglu counters by calling it an unproven hypothesis and noting that no serious frontier lab is solely scaling LLMs without broader architectural innovations.12:18–16:05 · Guest disagreement 2/10 DeepThink and Parallel Reasoning in Gemini 2.5 Pro Kantrowitz asks about the DeepThink announcement and inference-time compute. Kavukcuoglu clarifies that DeepThink is a reasoning mode rather than a standalone product, educating the host on parallel hypothesis exploration versus single chain-of-thought processing.16:05–18:44 · Guest disagreement 3/10 Native Multimodality and Continuous Frontier Velocity Kantrowitz asks if frontier model improvement velocity is slowing down compared to earlier leaps like GPT-3 to GPT-4. Kavukcuoglu disputes any perception of slowdown, arguing that native multimodality and unified reasoning architectures are accelerating overall capability.18:44–22:18 · Guest disagreement 2/10 Quantifying Model Improvement and User Feedback Kantrowitz poses a hypothetical question about the real-world value of 10% or 50% model improvements. Kavukcuoglu reframes the question around non-linear evaluation metrics and feedback loops between research and product deployments.22:18–26:24 · Guest disagreement 1/10 Generative Video Breakthroughs: Veo 3, Physics, and Flow Kantrowitz highlights Google's Veo 3 and Flow video tools, noting synchronized audio and storyboard capabilities. Kavukcuoglu explains the technical milestones from physical object dynamics in Veo 2 to orthogonal audio-visual modeling in Veo 3.26:24–30:20 · Guest disagreement 2/10 Open Weights vs. Closed Frontier Model Governance Kantrowitz asks how Google navigates the tension between open-source contributions and proprietary frontier models like Gemini, referencing DeepSeek and transformer history. Kavukcuoglu articulates Google's dual strategy of releasing Gemma open weights while restricting Gemini weights for safety and governance.30:20–31:33 · Guest disagreement 0/10 The Potential of 'Vibe Coding' for Accessibility Kantrowitz ends on a light note asking about vibe coding. Kavukcuoglu enthusiastically endorses the concept, explaining how it democratizes software creation for non-technical users.1:16–7:04 · Alex pushing back 5/10 Google DeepMind's Unified AGI Mission and Research Pillars Kantrowitz presses on whether scaling laws are hitting diminishing returns, citing industry-wide delays across OpenAI, Meta, and Anthropic. Kavukcuoglu provides a balanced breakdown of how compute scale interacts with algorithmic architecture, data quality, and inference techniques.7:04–9:13 · Alex pushing back 3/10 Model Progress Across the Industry and the Gemini Roadmap Kavukcuoglu reframes the scaling discussion around the tripartite nature of scaling laws (data, compute, parameters) rather than single-dimension scaling. Kantrowitz listens as the guest outlines DeepMind's progress from Gemini 1.5 to 2.5.9:13–12:18 · Alex pushing back 6/10 Debating AGI Pathways and Yann LeCun's Critique Kantrowitz challenges Kavukcuoglu with Yann LeCun's assertion that scaling LLMs cannot achieve human-level intelligence. Kavukcuoglu counters by calling it an unproven hypothesis and noting that no serious frontier lab is solely scaling LLMs without broader architectural innovations.12:18–16:05 · Alex pushing back 4/10 DeepThink and Parallel Reasoning in Gemini 2.5 Pro Kantrowitz asks about the DeepThink announcement and inference-time compute. Kavukcuoglu clarifies that DeepThink is a reasoning mode rather than a standalone product, educating the host on parallel hypothesis exploration versus single chain-of-thought processing.16:05–18:44 · Alex pushing back 5/10 Native Multimodality and Continuous Frontier Velocity Kantrowitz asks if frontier model improvement velocity is slowing down compared to earlier leaps like GPT-3 to GPT-4. Kavukcuoglu disputes any perception of slowdown, arguing that native multimodality and unified reasoning architectures are accelerating overall capability.18:44–22:18 · Alex pushing back 4/10 Quantifying Model Improvement and User Feedback Kantrowitz poses a hypothetical question about the real-world value of 10% or 50% model improvements. Kavukcuoglu reframes the question around non-linear evaluation metrics and feedback loops between research and product deployments.22:18–26:24 · Alex pushing back 2/10 Generative Video Breakthroughs: Veo 3, Physics, and Flow Kantrowitz highlights Google's Veo 3 and Flow video tools, noting synchronized audio and storyboard capabilities. Kavukcuoglu explains the technical milestones from physical object dynamics in Veo 2 to orthogonal audio-visual modeling in Veo 3.26:24–30:20 · Alex pushing back 4/10 Open Weights vs. Closed Frontier Model Governance Kantrowitz asks how Google navigates the tension between open-source contributions and proprietary frontier models like Gemini, referencing DeepSeek and transformer history. Kavukcuoglu articulates Google's dual strategy of releasing Gemma open weights while restricting Gemini weights for safety and governance.30:20–31:33 · Alex pushing back 1/10 The Potential of 'Vibe Coding' for Accessibility Kantrowitz ends on a light note asking about vibe coding. Kavukcuoglu enthusiastically endorses the concept, explaining how it democratizes software creation for non-technical users.

speaking balance: gold is Alex, purple is the guest (3 minute bins)

0:00 · Alex 52.4% · guest 47.6%0:00 · Alex 52.4% · guest 47.6%3:00 · Alex 32.7% · guest 67.3%3:00 · Alex 32.7% · guest 67.3%6:00 · Alex 35.4% · guest 64.6%6:00 · Alex 35.4% · guest 64.6%9:00 · Alex 13.1% · guest 86.9%9:00 · Alex 13.1% · guest 86.9%12:00 · Alex 29.6% · guest 70.4%12:00 · Alex 29.6% · guest 70.4%15:00 · Alex 18.7% · guest 81.3%15:00 · Alex 18.7% · guest 81.3%18:00 · Alex 22.4% · guest 77.6%18:00 · Alex 22.4% · guest 77.6%21:00 · Alex 36.4% · guest 63.6%21:00 · Alex 36.4% · guest 63.6%24:00 · Alex 24.5% · guest 75.5%24:00 · Alex 24.5% · guest 75.5%27:00 · Alex 10.1% · guest 89.9%27:00 · Alex 10.1% · guest 89.9%30:00 · Alex 19.9% · guest 80.1%30:00 · Alex 19.9% · guest 80.1%
Sharpest disagreement ▶ 9:37 Kavukcuoglu dismisses LeCun's critique as a strawman hypothesis

Kavukcuoglu firmly rejects the premise that labs are merely scaling LLMs to reach AGI, dismissing LeCun's argument as an untested hypothesis not reflective of actual frontier research.

Hardest push from Alex ▶ 9:13 Kantrowitz challenges DeepMind's path with LeCun's AGI critique

Kantrowitz confronts the guest directly by citing former lab director Yann LeCun's emphatic claim that scaling LLMs cannot reach human-level intelligence.

Biggest teaching moment ▶ 13:05 Kavukcuoglu corrects product framing and explains parallel reasoning

Kavukcuoglu clarifies that DeepThink is a reasoning mode rather than a new standalone product and educates the host on parallel hypothesis generation versus sequential chain of thought.

Alex holds their own ▶ 5:41 Kantrowitz cites industry-wide model stalls to question scaling limits

Kantrowitz demonstrates strong domain command by listing unreleased frontier models across OpenAI, Meta, and Anthropic to challenge the guest on scaling diminishing returns.

the scores for every segment, with the reasoning behind each
ChapterTopicAlex as informed peerGuest teachingGuest disagreementAlex pushing backWhy
Google DeepMind's Unified AGI Mission and Research Pillars 6425 Kantrowitz presses on whether scaling laws are hitting diminishing returns, citing industry-wide delays across OpenAI, Meta, and Anthropic. Kavukcuoglu provides a balanced breakdown of how compute scale interacts with algorithmic architecture, data quality, and inference techniques.
Model Progress Across the Industry and the Gemini Roadmap 4523 Kavukcuoglu reframes the scaling discussion around the tripartite nature of scaling laws (data, compute, parameters) rather than single-dimension scaling. Kantrowitz listens as the guest outlines DeepMind's progress from Gemini 1.5 to 2.5.
Debating AGI Pathways and Yann LeCun's Critique 6646 Kantrowitz challenges Kavukcuoglu with Yann LeCun's assertion that scaling LLMs cannot achieve human-level intelligence. Kavukcuoglu counters by calling it an unproven hypothesis and noting that no serious frontier lab is solely scaling LLMs without broader architectural innovations.
DeepThink and Parallel Reasoning in Gemini 2.5 Pro 5624 Kantrowitz asks about the DeepThink announcement and inference-time compute. Kavukcuoglu clarifies that DeepThink is a reasoning mode rather than a standalone product, educating the host on parallel hypothesis exploration versus single chain-of-thought processing.
Native Multimodality and Continuous Frontier Velocity 5435 Kantrowitz asks if frontier model improvement velocity is slowing down compared to earlier leaps like GPT-3 to GPT-4. Kavukcuoglu disputes any perception of slowdown, arguing that native multimodality and unified reasoning architectures are accelerating overall capability.
Quantifying Model Improvement and User Feedback 4524 Kantrowitz poses a hypothetical question about the real-world value of 10% or 50% model improvements. Kavukcuoglu reframes the question around non-linear evaluation metrics and feedback loops between research and product deployments.
Generative Video Breakthroughs: Veo 3, Physics, and Flow 5412 Kantrowitz highlights Google's Veo 3 and Flow video tools, noting synchronized audio and storyboard capabilities. Kavukcuoglu explains the technical milestones from physical object dynamics in Veo 2 to orthogonal audio-visual modeling in Veo 3.
Open Weights vs. Closed Frontier Model Governance 6524 Kantrowitz asks how Google navigates the tension between open-source contributions and proprietary frontier models like Gemini, referencing DeepSeek and transformer history. Kavukcuoglu articulates Google's dual strategy of releasing Gemma open weights while restricting Gemini weights for safety and governance.
The Potential of 'Vibe Coding' for Accessibility 3201 Kantrowitz ends on a light note asking about vibe coding. Kavukcuoglu enthusiastically endorses the concept, explaining how it democratizes software creation for non-technical users.

Statements from this episode (8)

Disclosure
DeepMind CTO: Gemini and Veo are our primary pathways to AGI
“So like with Gemini models, with VO models, all those kinds of generative AI models and exciting things that is happening there. That's our main line of AGI research that's going on.”
Koray Kavukcuoglu May 24, 2025 ▶ 2:12
Insight
DeepMind CTO: AI algorithms and architecture matter just as much as scale
“When we are thinking about our architectures, Like the architectural elements, the algorithms that we put in there that come, that make up the model, right? They are as important as the scale.”
Koray Kavukcuoglu May 24, 2025 ▶ 4:27
Opinion
DeepMind CTO denies AI slowdown, says progress across the field remains stellar
“Rather than thinking about scaling only in one dimension, there's actually many different ways to think about it, and investing in those, and we can see the returns that I think across the field, really, not just not just here at Google, but across the field, …”
Koray Kavukcuoglu May 24, 2025 ▶ 7:56
Opinion
Kavukcuoglu: No major AI lab relies solely on scaling LLMs for AGI
“But also, I don't think that there's any research lab that is trying to only do scaling up the LLM.”
Koray Kavukcuoglu May 24, 2025 ▶ 9:45
Disclosure
Gemini 2.5 Pro uses DeepThink to generate parallel hypotheses, not single chains
“It is a mode that we are enabling our 2.5 pro model so that, like, it can spend a lot more time during inference time to think, to build hypotheses, and the important thing is to build parallel hypotheses rather than a single chain of thought. It can build par…”
Koray Kavukcuoglu May 24, 2025 ▶ 13:11
Insight
DeepMind CTO: A 10% gain in mathematical reasoning vastly expands AI applicability
“So when we say by improving 10%, if we can improve 10% by its understanding in math, right, understanding of really highly complex reasoning problems, I think that is a huge improvement because then that actually expands the general knowledge that, that would …”
Koray Kavukcuoglu May 24, 2025 ▶ 19:47
Assertion Not checkable as stated
Kavukcuoglu: Veo 2 is DeepMind's first video model to truly understand physics
“With VO two, I think for the first time we could comfortably say that for many, many cases, right, the model has understood the dynamics of the world well.”
Koray Kavukcuoglu May 24, 2025 ▶ 23:43
Opinion
DeepMind CTO: 'Vibe coding' allows non-programmers to easily build software applications
“I find it really exciting, right? Like, I mean, because, like, what it does is all of a sudden it enables a lot of people who are not necessarily, who do not necessarily have that coding background to build applications.”
Koray Kavukcuoglu May 24, 2025 ▶ 30:30
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.