Sep 22, 2025 · 1h 38m · a16z

Dylan Patel on the AI Chip Race - NVIDIA, Intel & the US Government vs. China

Dylan Patel · 1h 13m spoken Sarah Wang · 8m spoken Guido Appenzeller · 5m spoken Erik Torenberg · 28s spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

In this episode of the a16z Podcast, SemiAnalysis founder Dylan Patel joins Sarah Wang, Erik Torenberg, and Guido Appenzeller to analyze the fast-evolving AI semiconductor landscape. Together, they explore NVIDIA's market dominance and Intel collaboration, US-China chip export controls, hyperscaler capital expenditure, and emerging hardware architectures.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. The host holds 0.5% of the talking time here. How this is scored →

The host as informed peer 2.8 Guest teaching 5.8 Guest disagreement 1.3 The host pushing back 1.4
05100:0020:0040:001:00:001:20:000:29–5:11 · The host as informed peer 3/10 NVIDIA's Investment in Intel and Future Product Collaboration Erik opens with NVIDIA's $5B investment in Intel. Dylan offers analysis on PC chiplet collaboration and Intel's financial position, while co-host Guido contributes context as former Intel Data Center CTO.5:11–15:01 · The host as informed peer 2/10 Huawei's AI Roadmap, US Sanctions, and Domestic Mandates Dylan provides a comprehensive history of Huawei's Ascend roadmaps, US sanction workarounds, and domestic stockpiles. Guido asks a clarifying question regarding China's export negotiation strategy.15:01–19:03 · The host as informed peer 2/10 HBM Supply Chain Bottlenecks and Manufacturing Yields Dylan breaks down high-bandwidth memory (HBM) supply bottlenecks, tracking Chinese equipment imports like etch tools and through-silicon vias. The host listens as Dylan explains manufacturing yield learning curves.19:03–22:24 · The host as informed peer 2/10 Jensen Huang's Huawei Strategy and Technological Galapagosing Dylan outlines Jensen Huang's strategy against Huawei, introducing Noah Smith's 'technological Galapagosing' concept. The hosts prompt the topic and let Dylan detail the thesis.22:24–29:20 · The host as informed peer 4/10 Hyperscaler CapEx Growth Projections and AI Value Creation Dylan presents his aggressive $455B-$500B CapEx numbers versus Wall Street consensus. Guido engages in peer-level discussion regarding white-collar productivity tax and AI economic value creation.29:20–39:55 · The host as informed peer 2/10 Jensen Huang's Leadership Style and NVIDIA's Historical Bets Dylan explains NVIDIA's historical bet-the-farm culture, aggressive supply ordering before contract confirmations, and Jensen's CFO/CEO dynamic.39:55–47:06 · The host as informed peer 4/10 NVIDIA's First-Pass Silicon Culture and Engineering Leadership Dylan highlights NVIDIA's engineering discipline of first-pass silicon success (A0 steppings). Guido adds host expertise by validating how Intel's repeated mask revisions (E2 steppings) created severe market delays.47:06–50:17 · The host as informed peer 2/10 Capital Allocation Strategy and Ecosystem Neutrality Dylan analyzes NVIDIA's massive balance sheet cash reserves and discusses why Jensen avoids buying entire startup funding rounds to remain ecosystem-neutral.50:17–55:53 · The host as informed peer 3/10 Power Infrastructure as the Ultimate AI Bottleneck Dylan argues NVIDIA should invest directly into power and data centers rather than cloud software layers. Guido adds historical context on Intel's customer concentration issues.55:53–1:02:56 · The host as informed peer 4/10 AWS Growth Re-Acceleration and Infrastructure Dominance Dylan outlines SemiAnalysis's thesis on AWS revenue re-acceleration. Guido pushes back intelligently, questioning whether legacy AWS data center capacity meets modern high-density liquid cooling requirements.1:02:56–1:07:03 · The host as informed peer 2/10 The Engineering Trade-offs of Custom AI Silicon Dylan details the engineering hurdles of developing custom AI chips like Trainium, pointing out how frontier labs write custom low-level assembly kernels for specific models.1:07:03–1:16:01 · The host as informed peer 3/10 Oracle's Cloud Expansion and Financial Engineering Dylan explains how SemiAnalysis predicted Oracle's surge by tracking satellite imagery, electrical permits, and transformer supply chains to map out gigawatt data center deals.1:16:01–1:22:07 · The host as informed peer 2/10 xAI Colossus II Buildout and Regulatory Maneuvering Dylan discusses xAI's Memphis cluster buildout and Elon Musk's tactic of expanding across the state border into Mississippi to bypass local power and municipal permitting bottlenecks.1:22:07–1:27:50 · The host as informed peer 3/10 GB200 TCO, Blast Radius Reliability, and SLA Architecture Dylan details the total cost of ownership (TCO) and severe reliability 'blast radius' issues of GB200 NVL72 architectures, explaining how failure rates require splitting high and low-priority workloads.1:27:50–1:34:49 · The host as informed peer 4/10 Disaggregated Pre-fill, Decode, and CPX Architecture Guido asks if NVIDIA's newly announced Rubin CPX pre-fill chip cannibalizes existing GPUs. Dylan explains the architectural split between compute-heavy pre-fill and memory-bandwidth-heavy decode.1:34:49–1:38:18 · The host as informed peer 2/10 GPU Procurement Dynamics and the Neocloud Market Dylan candidly compares buying GPUs to procurement dynamics in illegal drug markets. He wraps up by analyzing current Hopper versus Blackwell market tightness.0:29–5:11 · Guest teaching 4/10 NVIDIA's Investment in Intel and Future Product Collaboration Erik opens with NVIDIA's $5B investment in Intel. Dylan offers analysis on PC chiplet collaboration and Intel's financial position, while co-host Guido contributes context as former Intel Data Center CTO.5:11–15:01 · Guest teaching 7/10 Huawei's AI Roadmap, US Sanctions, and Domestic Mandates Dylan provides a comprehensive history of Huawei's Ascend roadmaps, US sanction workarounds, and domestic stockpiles. Guido asks a clarifying question regarding China's export negotiation strategy.15:01–19:03 · Guest teaching 6/10 HBM Supply Chain Bottlenecks and Manufacturing Yields Dylan breaks down high-bandwidth memory (HBM) supply bottlenecks, tracking Chinese equipment imports like etch tools and through-silicon vias. The host listens as Dylan explains manufacturing yield learning curves.19:03–22:24 · Guest teaching 5/10 Jensen Huang's Huawei Strategy and Technological Galapagosing Dylan outlines Jensen Huang's strategy against Huawei, introducing Noah Smith's 'technological Galapagosing' concept. The hosts prompt the topic and let Dylan detail the thesis.22:24–29:20 · Guest teaching 5/10 Hyperscaler CapEx Growth Projections and AI Value Creation Dylan presents his aggressive $455B-$500B CapEx numbers versus Wall Street consensus. Guido engages in peer-level discussion regarding white-collar productivity tax and AI economic value creation.29:20–39:55 · Guest teaching 6/10 Jensen Huang's Leadership Style and NVIDIA's Historical Bets Dylan explains NVIDIA's historical bet-the-farm culture, aggressive supply ordering before contract confirmations, and Jensen's CFO/CEO dynamic.39:55–47:06 · Guest teaching 6/10 NVIDIA's First-Pass Silicon Culture and Engineering Leadership Dylan highlights NVIDIA's engineering discipline of first-pass silicon success (A0 steppings). Guido adds host expertise by validating how Intel's repeated mask revisions (E2 steppings) created severe market delays.47:06–50:17 · Guest teaching 5/10 Capital Allocation Strategy and Ecosystem Neutrality Dylan analyzes NVIDIA's massive balance sheet cash reserves and discusses why Jensen avoids buying entire startup funding rounds to remain ecosystem-neutral.50:17–55:53 · Guest teaching 5/10 Power Infrastructure as the Ultimate AI Bottleneck Dylan argues NVIDIA should invest directly into power and data centers rather than cloud software layers. Guido adds historical context on Intel's customer concentration issues.55:53–1:02:56 · Guest teaching 6/10 AWS Growth Re-Acceleration and Infrastructure Dominance Dylan outlines SemiAnalysis's thesis on AWS revenue re-acceleration. Guido pushes back intelligently, questioning whether legacy AWS data center capacity meets modern high-density liquid cooling requirements.1:02:56–1:07:03 · Guest teaching 6/10 The Engineering Trade-offs of Custom AI Silicon Dylan details the engineering hurdles of developing custom AI chips like Trainium, pointing out how frontier labs write custom low-level assembly kernels for specific models.1:07:03–1:16:01 · Guest teaching 7/10 Oracle's Cloud Expansion and Financial Engineering Dylan explains how SemiAnalysis predicted Oracle's surge by tracking satellite imagery, electrical permits, and transformer supply chains to map out gigawatt data center deals.1:16:01–1:22:07 · Guest teaching 6/10 xAI Colossus II Buildout and Regulatory Maneuvering Dylan discusses xAI's Memphis cluster buildout and Elon Musk's tactic of expanding across the state border into Mississippi to bypass local power and municipal permitting bottlenecks.1:22:07–1:27:50 · Guest teaching 7/10 GB200 TCO, Blast Radius Reliability, and SLA Architecture Dylan details the total cost of ownership (TCO) and severe reliability 'blast radius' issues of GB200 NVL72 architectures, explaining how failure rates require splitting high and low-priority workloads.1:27:50–1:34:49 · Guest teaching 7/10 Disaggregated Pre-fill, Decode, and CPX Architecture Guido asks if NVIDIA's newly announced Rubin CPX pre-fill chip cannibalizes existing GPUs. Dylan explains the architectural split between compute-heavy pre-fill and memory-bandwidth-heavy decode.1:34:49–1:38:18 · Guest teaching 5/10 GPU Procurement Dynamics and the Neocloud Market Dylan candidly compares buying GPUs to procurement dynamics in illegal drug markets. He wraps up by analyzing current Hopper versus Blackwell market tightness.0:29–5:11 · Guest disagreement 1/10 NVIDIA's Investment in Intel and Future Product Collaboration Erik opens with NVIDIA's $5B investment in Intel. Dylan offers analysis on PC chiplet collaboration and Intel's financial position, while co-host Guido contributes context as former Intel Data Center CTO.5:11–15:01 · Guest disagreement 1/10 Huawei's AI Roadmap, US Sanctions, and Domestic Mandates Dylan provides a comprehensive history of Huawei's Ascend roadmaps, US sanction workarounds, and domestic stockpiles. Guido asks a clarifying question regarding China's export negotiation strategy.15:01–19:03 · Guest disagreement 1/10 HBM Supply Chain Bottlenecks and Manufacturing Yields Dylan breaks down high-bandwidth memory (HBM) supply bottlenecks, tracking Chinese equipment imports like etch tools and through-silicon vias. The host listens as Dylan explains manufacturing yield learning curves.19:03–22:24 · Guest disagreement 1/10 Jensen Huang's Huawei Strategy and Technological Galapagosing Dylan outlines Jensen Huang's strategy against Huawei, introducing Noah Smith's 'technological Galapagosing' concept. The hosts prompt the topic and let Dylan detail the thesis.22:24–29:20 · Guest disagreement 2/10 Hyperscaler CapEx Growth Projections and AI Value Creation Dylan presents his aggressive $455B-$500B CapEx numbers versus Wall Street consensus. Guido engages in peer-level discussion regarding white-collar productivity tax and AI economic value creation.29:20–39:55 · Guest disagreement 1/10 Jensen Huang's Leadership Style and NVIDIA's Historical Bets Dylan explains NVIDIA's historical bet-the-farm culture, aggressive supply ordering before contract confirmations, and Jensen's CFO/CEO dynamic.39:55–47:06 · Guest disagreement 1/10 NVIDIA's First-Pass Silicon Culture and Engineering Leadership Dylan highlights NVIDIA's engineering discipline of first-pass silicon success (A0 steppings). Guido adds host expertise by validating how Intel's repeated mask revisions (E2 steppings) created severe market delays.47:06–50:17 · Guest disagreement 1/10 Capital Allocation Strategy and Ecosystem Neutrality Dylan analyzes NVIDIA's massive balance sheet cash reserves and discusses why Jensen avoids buying entire startup funding rounds to remain ecosystem-neutral.50:17–55:53 · Guest disagreement 2/10 Power Infrastructure as the Ultimate AI Bottleneck Dylan argues NVIDIA should invest directly into power and data centers rather than cloud software layers. Guido adds historical context on Intel's customer concentration issues.55:53–1:02:56 · Guest disagreement 2/10 AWS Growth Re-Acceleration and Infrastructure Dominance Dylan outlines SemiAnalysis's thesis on AWS revenue re-acceleration. Guido pushes back intelligently, questioning whether legacy AWS data center capacity meets modern high-density liquid cooling requirements.1:02:56–1:07:03 · Guest disagreement 1/10 The Engineering Trade-offs of Custom AI Silicon Dylan details the engineering hurdles of developing custom AI chips like Trainium, pointing out how frontier labs write custom low-level assembly kernels for specific models.1:07:03–1:16:01 · Guest disagreement 1/10 Oracle's Cloud Expansion and Financial Engineering Dylan explains how SemiAnalysis predicted Oracle's surge by tracking satellite imagery, electrical permits, and transformer supply chains to map out gigawatt data center deals.1:16:01–1:22:07 · Guest disagreement 1/10 xAI Colossus II Buildout and Regulatory Maneuvering Dylan discusses xAI's Memphis cluster buildout and Elon Musk's tactic of expanding across the state border into Mississippi to bypass local power and municipal permitting bottlenecks.1:22:07–1:27:50 · Guest disagreement 2/10 GB200 TCO, Blast Radius Reliability, and SLA Architecture Dylan details the total cost of ownership (TCO) and severe reliability 'blast radius' issues of GB200 NVL72 architectures, explaining how failure rates require splitting high and low-priority workloads.1:27:50–1:34:49 · Guest disagreement 1/10 Disaggregated Pre-fill, Decode, and CPX Architecture Guido asks if NVIDIA's newly announced Rubin CPX pre-fill chip cannibalizes existing GPUs. Dylan explains the architectural split between compute-heavy pre-fill and memory-bandwidth-heavy decode.1:34:49–1:38:18 · Guest disagreement 2/10 GPU Procurement Dynamics and the Neocloud Market Dylan candidly compares buying GPUs to procurement dynamics in illegal drug markets. He wraps up by analyzing current Hopper versus Blackwell market tightness.0:29–5:11 · The host pushing back 1/10 NVIDIA's Investment in Intel and Future Product Collaboration Erik opens with NVIDIA's $5B investment in Intel. Dylan offers analysis on PC chiplet collaboration and Intel's financial position, while co-host Guido contributes context as former Intel Data Center CTO.5:11–15:01 · The host pushing back 1/10 Huawei's AI Roadmap, US Sanctions, and Domestic Mandates Dylan provides a comprehensive history of Huawei's Ascend roadmaps, US sanction workarounds, and domestic stockpiles. Guido asks a clarifying question regarding China's export negotiation strategy.15:01–19:03 · The host pushing back 1/10 HBM Supply Chain Bottlenecks and Manufacturing Yields Dylan breaks down high-bandwidth memory (HBM) supply bottlenecks, tracking Chinese equipment imports like etch tools and through-silicon vias. The host listens as Dylan explains manufacturing yield learning curves.19:03–22:24 · The host pushing back 1/10 Jensen Huang's Huawei Strategy and Technological Galapagosing Dylan outlines Jensen Huang's strategy against Huawei, introducing Noah Smith's 'technological Galapagosing' concept. The hosts prompt the topic and let Dylan detail the thesis.22:24–29:20 · The host pushing back 2/10 Hyperscaler CapEx Growth Projections and AI Value Creation Dylan presents his aggressive $455B-$500B CapEx numbers versus Wall Street consensus. Guido engages in peer-level discussion regarding white-collar productivity tax and AI economic value creation.29:20–39:55 · The host pushing back 1/10 Jensen Huang's Leadership Style and NVIDIA's Historical Bets Dylan explains NVIDIA's historical bet-the-farm culture, aggressive supply ordering before contract confirmations, and Jensen's CFO/CEO dynamic.39:55–47:06 · The host pushing back 1/10 NVIDIA's First-Pass Silicon Culture and Engineering Leadership Dylan highlights NVIDIA's engineering discipline of first-pass silicon success (A0 steppings). Guido adds host expertise by validating how Intel's repeated mask revisions (E2 steppings) created severe market delays.47:06–50:17 · The host pushing back 1/10 Capital Allocation Strategy and Ecosystem Neutrality Dylan analyzes NVIDIA's massive balance sheet cash reserves and discusses why Jensen avoids buying entire startup funding rounds to remain ecosystem-neutral.50:17–55:53 · The host pushing back 2/10 Power Infrastructure as the Ultimate AI Bottleneck Dylan argues NVIDIA should invest directly into power and data centers rather than cloud software layers. Guido adds historical context on Intel's customer concentration issues.55:53–1:02:56 · The host pushing back 3/10 AWS Growth Re-Acceleration and Infrastructure Dominance Dylan outlines SemiAnalysis's thesis on AWS revenue re-acceleration. Guido pushes back intelligently, questioning whether legacy AWS data center capacity meets modern high-density liquid cooling requirements.1:02:56–1:07:03 · The host pushing back 1/10 The Engineering Trade-offs of Custom AI Silicon Dylan details the engineering hurdles of developing custom AI chips like Trainium, pointing out how frontier labs write custom low-level assembly kernels for specific models.1:07:03–1:16:01 · The host pushing back 1/10 Oracle's Cloud Expansion and Financial Engineering Dylan explains how SemiAnalysis predicted Oracle's surge by tracking satellite imagery, electrical permits, and transformer supply chains to map out gigawatt data center deals.1:16:01–1:22:07 · The host pushing back 1/10 xAI Colossus II Buildout and Regulatory Maneuvering Dylan discusses xAI's Memphis cluster buildout and Elon Musk's tactic of expanding across the state border into Mississippi to bypass local power and municipal permitting bottlenecks.1:22:07–1:27:50 · The host pushing back 2/10 GB200 TCO, Blast Radius Reliability, and SLA Architecture Dylan details the total cost of ownership (TCO) and severe reliability 'blast radius' issues of GB200 NVL72 architectures, explaining how failure rates require splitting high and low-priority workloads.1:27:50–1:34:49 · The host pushing back 2/10 Disaggregated Pre-fill, Decode, and CPX Architecture Guido asks if NVIDIA's newly announced Rubin CPX pre-fill chip cannibalizes existing GPUs. Dylan explains the architectural split between compute-heavy pre-fill and memory-bandwidth-heavy decode.1:34:49–1:38:18 · The host pushing back 1/10 GPU Procurement Dynamics and the Neocloud Market Dylan candidly compares buying GPUs to procurement dynamics in illegal drug markets. He wraps up by analyzing current Hopper versus Blackwell market tightness.

speaking balance: gold is the host, purple is the guest (3 minute bins)

0:00 · the host 8% · guest 92%0:00 · the host 8% · guest 92%3:00 · the host 0% · guest 100%3:00 · the host 0% · guest 100%6:00 · the host 0% · guest 100%6:00 · the host 0% · guest 100%9:00 · the host 0% · guest 100%9:00 · the host 0% · guest 100%12:00 · the host 0% · guest 100%12:00 · the host 0% · guest 100%15:00 · the host 0% · guest 100%15:00 · the host 0% · guest 100%18:00 · the host 1.2% · guest 98.8%18:00 · the host 1.2% · guest 98.8%21:00 · the host 0.8% · guest 99.2%21:00 · the host 0.8% · guest 99.2%24:00 · the host 0% · guest 100%24:00 · the host 0% · guest 100%27:00 · the host 0% · guest 100%27:00 · the host 0% · guest 100%30:00 · the host 0% · guest 100%30:00 · the host 0% · guest 100%33:00 · the host 0% · guest 100%33:00 · the host 0% · guest 100%36:00 · the host 0% · guest 100%36:00 · the host 0% · guest 100%39:00 · the host 0% · guest 100%39:00 · the host 0% · guest 100%42:00 · the host 0% · guest 100%42:00 · the host 0% · guest 100%45:00 · the host 0% · guest 100%45:00 · the host 0% · guest 100%48:00 · the host 1.2% · guest 98.8%48:00 · the host 1.2% · guest 98.8%51:00 · the host 1.7% · guest 98.3%51:00 · the host 1.7% · guest 98.3%54:00 · the host 0% · guest 100%54:00 · the host 0% · guest 100%57:00 · the host 0% · guest 100%57:00 · the host 0% · guest 100%1:00:00 · the host 0% · guest 100%1:00:00 · the host 0% · guest 100%1:03:00 · the host 0% · guest 100%1:03:00 · the host 0% · guest 100%1:06:00 · the host 0% · guest 100%1:06:00 · the host 0% · guest 100%1:09:00 · the host 0% · guest 100%1:09:00 · the host 0% · guest 100%1:12:00 · the host 0% · guest 100%1:12:00 · the host 0% · guest 100%1:15:00 · the host 0% · guest 100%1:15:00 · the host 0% · guest 100%1:18:00 · the host 0% · guest 100%1:18:00 · the host 0% · guest 100%1:21:00 · the host 0% · guest 100%1:21:00 · the host 0% · guest 100%1:24:00 · the host 0% · guest 100%1:24:00 · the host 0% · guest 100%1:27:00 · the host 0% · guest 100%1:27:00 · the host 0% · guest 100%1:30:00 · the host 0% · guest 100%1:30:00 · the host 0% · guest 100%1:33:00 · the host 0% · guest 100%1:33:00 · the host 0% · guest 100%1:36:00 · the host 4.8% · guest 95.2%1:36:00 · the host 4.8% · guest 95.2%
Sharpest disagreement ▶ 1:35:33 Cocaine Procurement Analogy

Dylan bluntly compares buying enterprise AI GPUs to calling around for illegal drugs, using informal and provocative framing.

Hardest push from the host ▶ 1:00:25 Data Center Infrastructure Challenge

Guido directly challenges Dylan's call on AWS re-acceleration by questioning whether legacy AWS facilities have adequate density, water access, and liquid cooling capabilities.

Biggest teaching moment ▶ 1:32:00 Disaggregated Pre-fill vs Decode Mechanics

Dylan walks through the mathematical and architectural distinctions between pre-fill and decode compute requirements to clarify why CPX strips expensive HBM.

The host holds their own ▶ 44:00 Intel's Mask Revision Disasters

Guido leverages his background as former Intel Data Center CTO to demonstrate insider knowledge on how Intel reached 15 mask steppings (E2), causing multi-quarter delays.

the scores for every segment, with the reasoning behind each
ChapterTopicThe host as informed peerGuest teachingGuest disagreementThe host pushing backWhy
NVIDIA's Investment in Intel and Future Product Collaboration 3411 Erik opens with NVIDIA's $5B investment in Intel. Dylan offers analysis on PC chiplet collaboration and Intel's financial position, while co-host Guido contributes context as former Intel Data Center CTO.
Huawei's AI Roadmap, US Sanctions, and Domestic Mandates 2711 Dylan provides a comprehensive history of Huawei's Ascend roadmaps, US sanction workarounds, and domestic stockpiles. Guido asks a clarifying question regarding China's export negotiation strategy.
HBM Supply Chain Bottlenecks and Manufacturing Yields 2611 Dylan breaks down high-bandwidth memory (HBM) supply bottlenecks, tracking Chinese equipment imports like etch tools and through-silicon vias. The host listens as Dylan explains manufacturing yield learning curves.
Jensen Huang's Huawei Strategy and Technological Galapagosing 2511 Dylan outlines Jensen Huang's strategy against Huawei, introducing Noah Smith's 'technological Galapagosing' concept. The hosts prompt the topic and let Dylan detail the thesis.
Hyperscaler CapEx Growth Projections and AI Value Creation 4522 Dylan presents his aggressive $455B-$500B CapEx numbers versus Wall Street consensus. Guido engages in peer-level discussion regarding white-collar productivity tax and AI economic value creation.
Jensen Huang's Leadership Style and NVIDIA's Historical Bets 2611 Dylan explains NVIDIA's historical bet-the-farm culture, aggressive supply ordering before contract confirmations, and Jensen's CFO/CEO dynamic.
NVIDIA's First-Pass Silicon Culture and Engineering Leadership 4611 Dylan highlights NVIDIA's engineering discipline of first-pass silicon success (A0 steppings). Guido adds host expertise by validating how Intel's repeated mask revisions (E2 steppings) created severe market delays.
Capital Allocation Strategy and Ecosystem Neutrality 2511 Dylan analyzes NVIDIA's massive balance sheet cash reserves and discusses why Jensen avoids buying entire startup funding rounds to remain ecosystem-neutral.
Power Infrastructure as the Ultimate AI Bottleneck 3522 Dylan argues NVIDIA should invest directly into power and data centers rather than cloud software layers. Guido adds historical context on Intel's customer concentration issues.
AWS Growth Re-Acceleration and Infrastructure Dominance 4623 Dylan outlines SemiAnalysis's thesis on AWS revenue re-acceleration. Guido pushes back intelligently, questioning whether legacy AWS data center capacity meets modern high-density liquid cooling requirements.
The Engineering Trade-offs of Custom AI Silicon 2611 Dylan details the engineering hurdles of developing custom AI chips like Trainium, pointing out how frontier labs write custom low-level assembly kernels for specific models.
Oracle's Cloud Expansion and Financial Engineering 3711 Dylan explains how SemiAnalysis predicted Oracle's surge by tracking satellite imagery, electrical permits, and transformer supply chains to map out gigawatt data center deals.
xAI Colossus II Buildout and Regulatory Maneuvering 2611 Dylan discusses xAI's Memphis cluster buildout and Elon Musk's tactic of expanding across the state border into Mississippi to bypass local power and municipal permitting bottlenecks.
GB200 TCO, Blast Radius Reliability, and SLA Architecture 3722 Dylan details the total cost of ownership (TCO) and severe reliability 'blast radius' issues of GB200 NVL72 architectures, explaining how failure rates require splitting high and low-priority workloads.
Disaggregated Pre-fill, Decode, and CPX Architecture 4712 Guido asks if NVIDIA's newly announced Rubin CPX pre-fill chip cannibalizes existing GPUs. Dylan explains the architectural split between compute-heavy pre-fill and memory-bandwidth-heavy decode.
GPU Procurement Dynamics and the Neocloud Market 2521 Dylan candidly compares buying GPUs to procurement dynamics in illegal drug markets. He wraps up by analyzing current Hopper versus Blackwell market tightness.

Statements from this episode (42)

Insight
Patel: Buying AI GPUs resembles buying cocaine via informal networks
“How you buy GPUs is like buying cocaine. You call up a couple people, you text a couple people, you ask, yo, how much you got? What's the price?”
Dylan Patel Sep 22, 2025 ▶ 0:00
Opinion
Dylan Patel: Intel is crawling to NVIDIA in a full-circle reversal
“It's kind of poetic that everything's gone full circle and Intel's sort of crawling to NVIDIA.”
Dylan Patel Sep 22, 2025 ▶ 1:49
Opinion
Patel: Integrated x86 and Nvidia graphics laptop would be best on market
“An x-ay six laptop with NVIDIA graphics fully integrated would be probably the best product in the market.”
Dylan Patel Sep 22, 2025 ▶ 2:03
Assertion Not checkable as stated
Appenzeller: Intel lacks competitive AI chips and its Gaudi effort is done
“They can't, they don't have anything competitive, right? There was the Gaudi effort that's more or less done, right? There was the internal graphics chips, which never competed really at the high end, right?”
Guido Appenzeller Sep 22, 2025 ▶ 4:09
Opinion
Appenzeller: Intel-Nvidia partnership means AMD is 'fucked'
“I think AMD is fucked, right? I mean, they're, you're just, If your two arch nemesis suddenly team up, that's the worst possible news you can have, right? They were already struggling, right? Their cards are good, their software stack is not, right? They were …”
Guido Appenzeller Sep 22, 2025 ▶ 4:30
Opinion
Appenzeller: Intel-Nvidia alliance weakens Arm's core value proposition
“I think ARM is a little bit screwed as well, right? Because they are, their biggest selling point was sort of like, look, we can partner with everybody that doesn't want to partner with Intel.”
Guido Appenzeller Sep 22, 2025 ▶ 4:47
Assertion Supported
Patel: Huawei Sourced 2.9 Million TSMC Chips via Shell Entities
“They were able to acquire three million chips, 2.9 million chips from TSMC through these other entities, right? Roughly five hundred million dollars worth of orders”
Dylan Patel Sep 22, 2025 ▶ 7:55
Assertion Supported
Patel: NVIDIA Lost $20B+ China Revenue from H20 Ban
“Our revenue estimate for Nvidia in China for just H-twenty was north of twenty billion because that's what they were booking in capacity slash had to write off.”
Dylan Patel Sep 22, 2025 ▶ 8:36
Prediction Not checkable as stated
Patel: China Will Mass Produce 7nm AI Chips Despite Sanctions
“Even though the government says they're for 14 nanometer, the actual equipment that's banned is only for below seven nanometer, and so they'll be able to make a lot of seven nanometer AI chips and maybe even get to five with, you know, using existing equipment…”
Dylan Patel Sep 22, 2025 ▶ 9:56
Prediction Not checkable as stated
Patel: China Will Pause NVIDIA Bans During AI Chip Capacity Gap
“I think they'll be able to ramp. I think it'll take a little bit longer and there will be, like, a sort of a gap in between where China probably backtracks and says it's fine.”
Dylan Patel Sep 22, 2025 ▶ 12:10
Assertion Not checkable as stated
Patel: HBM production capacity remains a bottleneck for Huawei
“I think production capacity wise, it is still absolutely a bottleneck. They certain types of equipment required for making HBM need to be imported. They're working on domestic solutions, but as far as we know, they have not imported enough equipment for this.”
Dylan Patel Sep 22, 2025 ▶ 15:21
Assertion Supported
Patel: China stockpiled lithography at 30-40% of equipment imports
“China, because they wanted, they sort of, like, wanted to stockpile lithography, and they were worried about the coming ban, they were importing lithography at a much higher rate than that, right? Like, 30, 40% of their equipment imports were lithography, and …”
Dylan Patel Sep 22, 2025 ▶ 16:00
Assertion Partly supported
Patel: China has only sampled HBM2 and not started HBM3 production
“Yield, they haven't even started production of high speed, of HBM-III, right? They've only done some sampling of HBM-II.”
Dylan Patel Sep 22, 2025 ▶ 17:12
Prediction Open · timeframe Sep 2030
Patel: AI end market will far exceed total semiconductor market
“The end market of AI is going to be way larger than the end market of semiconductors and equipment.”
Dylan Patel Sep 22, 2025 ▶ 18:45
Assertion Partly supported
Patel: Huawei surpassed Apple in TSMC orders and market share before bans
“Yeah, well, like, I mean, like, every other, like, Huawei's beat Apple, right? They passed Apple up in TSMC orders, they passed Apple up in phone market share not in the US, but, like, in many parts of the world. Before the bans came down, and then even now th…”
Dylan Patel Sep 22, 2025 ▶ 19:22
Assertion Supported
Patel: Wall Street projects $360B hyperscaler CapEx next year
“The consensus for the banks is three hundred and sixty billion dollars of spend next year across all of them.”
Dylan Patel Sep 22, 2025 ▶ 23:09
Prediction Open · timeframe Dec 2026
Patel: Hyperscaler CapEx will reach $450B-$500B next year
“And my number is closer to, like, it's like four 5500 and that's based on, like, you know, all the research we do on, like, data centers and, like, tracking each individual data center in the supply chains, right?”
Dylan Patel Sep 22, 2025 ▶ 23:16
Prediction Not checkable as stated
Patel: OpenAI will not achieve profitability until 2029
“They're not turning a cash flow, they're not gonna be profitable until 2029.”
Dylan Patel Sep 22, 2025 ▶ 25:27
Assertion Contradicted
Patel: NVIDIA is the only $10B+ semiconductor company founded post-1990
“It's the only semiconductor company that's worth, you know, I think even north of ten billion dollars that was founded as late as it was. Like MediaTek was in the early nineties and then Nvidia and everyone else is like from the seventies mostly.”
Dylan Patel Sep 22, 2025 ▶ 35:31
Assertion Partly supported
Patel: NVIDIA almost always ships first-pass A0 silicon
“So, like, NVIDIA always ships A zero. Almost always. They sometimes ship A one.”
Dylan Patel Sep 22, 2025 ▶ 43:31
Assertion Not checkable as stated
Patel: NVIDIA added Tensor cores to Volta months before fab
“There's a story about how Volta, Which was the first NVIDIA chip with Tensor cores. You know, they saw all the AI stuff on the prior generation P. 100 Pascal. And they decided we should go all in on AI. And they added the Tensor cores to Volta, like only a han…”
Dylan Patel Sep 22, 2025 ▶ 45:25
Prediction Open · timeframe Sep 2028
Patel: NVIDIA will accumulate hundreds of billions in cash balance
“He's gonna have hundreds of billions of dollars of cash on his balance sheet.”
Dylan Patel Sep 22, 2025 ▶ 48:07
Assertion Not checkable as stated
Patel: Silicon Valley AI startups spend 75% of venture rounds on GPUs
“Most companies in the valley spend, what, 75% of their round on GPUs, right?”
Dylan Patel Sep 22, 2025 ▶ 48:53
Insight
Patel: Energy and data centers are NVIDIA's primary growth bottleneck
“Do you invest in data centers and energy? Yeah. Do you invest in, because that's the bottleneck for your growth really is, is A, well, how much people want to spend and can spend, and B, the ability to actually put them in data centers.”
Dylan Patel Sep 22, 2025 ▶ 53:08
Opinion
Patel: Apple has lacked visionary innovation under Tim Cook for a decade
“The reason why Apple hasn't done anything interesting in like, you know, nearly a decade is, is, you know, they've got a not visionary at the head. Tim Cook's greatest supply chain. And they're just plowing the money into buybacks.”
Dylan Patel Sep 22, 2025 ▶ 53:50
Disclosure
Appenzeller: Intel's customer concentration allowed hyperscalers to push down prices
“One of the biggest problems we had was that our customer base sucked, right? I mean, we were selling to, most of the chips went to the large hyperscalers, you know, which they're way too concentrated, and they build their own chips, and so you can push down yo…”
Guido Appenzeller Sep 22, 2025 ▶ 55:27
Prediction Held up
Patel: AWS year-over-year revenue growth will re-accelerate above 20%
“This is the lowest AWS revenue growth will be on a year-to-year basis. For at least the next year, right? And it's re-accelerating to north of 20% again because of all these massive data centers they have online with Tranium and GPUs, right?”
Dylan Patel Sep 22, 2025 ▶ 59:32
Prediction Not checkable as stated
Patel: Amazon will lose its position as largest data center operator within two years
“The company with the most data center capacity in the world, that, and still today, although they may get passed up in the next two years is Amazon. Actually, they will get passed up based on what we see is Amazon, but incrementally, Amazon still has the most …”
Dylan Patel Sep 22, 2025 ▶ 1:00:07
Opinion
Patel: AWS Trainium remains difficult for developers to program
“Oh, it's still bad. You know, it's tough to use.”
Dylan Patel Sep 22, 2025 ▶ 1:03:30
Assertion Supported
Patel: Oracle does not physically build its own data centers
“Oracle doesn't build their own data centers either, right? By the way, they get them from other companies. They co-engineer but they don't, Physically build them themselves.”
Dylan Patel Sep 22, 2025 ▶ 1:09:21
Assertion Supported
Patel: NVIDIA GB200 system costs $50,000 all-in per GPU
“If I'm talking about a GV 200, right? Each individual GPU is 1200 watts. But when you talk about the CPU, the whole system, it's roughly 2000 watts. At the same time, you know, all in, everything, simplicity's sake, 50,000 dollars per GPU, right? The GPU doesn…”
Dylan Patel Sep 22, 2025 ▶ 1:10:44
Assertion Partly supported
Patel: OpenAI signed up to pay Oracle $80+ billion annually
“With OpenAI, it's not, and so there's gotta be some, like, error bars as you go further out in terms of, like, Will OpenAI exist in 28, 29 30, and will they be able to pay the 80 plus billion dollars a year that they've signed up to Oracle with, right?”
Dylan Patel Sep 22, 2025 ▶ 1:13:08
Assertion Supported
Patel: Oracle purchases GPUs only 1-2 quarters before leasing
“The GPUs they purchase one to two quarters before they start renting them. So they, they're not, you know, the downside risk is pretty low for them in terms of If they don't get the deal, well, they don't get the revenue, but they're not, it's not like they ha…”
Dylan Patel Sep 22, 2025 ▶ 1:13:31
Assertion Not checkable as stated
Patel: Approximately ten 100,000-GPU AI clusters exist globally
“There are like 10 hundred K GPU clusters in the world.”
Dylan Patel Sep 22, 2025 ▶ 1:17:09
Assertion Partly supported
Patel: Elon Musk deployed 100,000 GPUs in Memphis within six months
“A hundred K GPUs in six months. He bought a factory in like February of 24 and had models training within six months, right?”
Dylan Patel Sep 22, 2025 ▶ 1:18:19
Assertion Supported
Patel: Musk bought a Mississippi power plant to bypass energy constraints
“His facility's like a mile away from Mississippi, and he bought a power plant in Mississippi and he's putting turbines there. Yeah. The regulation is completely different, right?”
Dylan Patel Sep 22, 2025 ▶ 1:20:46
Assertion Supported
Patel: NVIDIA GB200 yields 3x-4x performance per dollar on DeepSeek
“If you're running DeepSeq inference, the performance difference per GPU is like north of like six, seven X, and it continues to optimize you know, for DeepSeq inference. And so the, you know, then it's like, well, I'm only paying 60% more for six X, and it's l…”
Dylan Patel Sep 22, 2025 ▶ 1:23:50
Assertion Not checkable as stated
Patel: Next-gen GPU failure rates are flat or getting worse
“GPU failure rates at best are the same and likely worse, right? Gen on gen, because everything's getting hotter, faster, et cetera.”
Dylan Patel Sep 22, 2025 ▶ 1:25:12
Assertion Partly supported
Patel: HBM makes up over half of GPU cost
“HBM is more than half the cost of the GPU.”
Dylan Patel Sep 22, 2025 ▶ 1:34:28
Assertion Not checkable as stated
Patel: Multiple major neoclouds have sold out of NVIDIA Hopper capacity
“It turned out that multiple major neoclouds had sold out of hopper capacity.”
Dylan Patel Sep 22, 2025 ▶ 1:36:58
Assertion Not checkable as stated
Patel: Reasoning models are driving a surge in AI inference demand
“Inference demand has been skyrocketing this year, right? These reasoning models, the revenue it's been skyrocketing this year”
Dylan Patel Sep 22, 2025 ▶ 1:37:11
Assertion Not checkable as stated
Patel: NVIDIA Hopper rental prices bottomed and are creeping back up
“And actually prices for hopper bottomed like three or four months ago or like five or six months ago. And actually they've like crept up a little bit now.”
Dylan Patel Sep 22, 2025 ▶ 1:37:52
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.