Lie: AMD, Trainium, and TPU are all trying to build a better Nvidia Rubin
“AMD, Tranium, in many ways TPU, like all of these in my mind are all trying to build a better Reuben, right? And there's a huge amount of value in that.”
Google TPU Data Centers Require Double the Equity Check of NVIDIA GPUs
“TPUs are the second most financeable. It probably takes, I don't know, double the equity check at least. And then the rates on the rest of it are higher.”
Thompson: Google agreed to sell about 20% of its TPUs to Anthropic
“Google already made a deal to sell, like, 20% of their TPUs to Anthropic.”
Moe: LLM serving differs fundamentally from traditional ML workloads
“Serving large language model is a fundamentally different problem. Because serving it requires to run it on accelerators like GPUs or TPUs, and it is a computationally intensive process that will require a lot of engineering and ensuring that for each request,…”
Cahn: Google's biggest AI advantages are search cash and TPUs
“Google has two big advantages, right? It has a cache machine from search that can fund a lot of things. And then number two, it has TPU. And I cannot overstate what an advantage it is to have TPU. They've been building this chip for a long time. It's a really …”
Dean: Specialized inference hardware will surpass general GPUs and TPUs
“I think, ah, you're gonna see more and more, ah high performance and low energy inference hardware systems, because I think everyone is now realizing that inference is the key to making, you know, these agent-based systems be available to more and more people,…”
Chaubard: Chip specialization and ASICs will proliferate across AI workloads
“From my vantage point, we're gonna see this proliferation, and where we have sufficient demand now, because there's so much demand for tokens, Where it makes sense to specialize at the chip level that will pay out and allow you to go through the full new produ…”
Katti: Many OpenAI chip team members previously designed Google TPUs
“They have, many, many of the team have designed TPU chips at Google in the past.”
Feldman: Google TPU Co-Design with Gemini and DeepMind Is a Major Advantage
“One of the advantages Google has is that their TPU can be designed in collaboration with the team building Gemini or the team of DeepMind, and so they can inform their choices back and forth, and that's an enormously powerful thing that is surprisingly relativ…”
Eliahu-Ontiveros: Nvidia risks losing silicon share unless it backstops neocloud power
“It's actually becoming a pretty existential risk for NVIDIA because if they don't support the NeoClouds, they're at risk of basically seeing Google Amazon and others take tremendous market share on the silicon side, right? So you would see Tranium take, take m…”
Ross: GPUs are now better than Google TPUs
“So over time, these GPUs have, you know, gotten, you know, as good and better than TPUs, you know, as the career of the TPU, I have to admit, GPUs are now better.”
Wachen: Existing GPUs and TPUs are retrofitted pre-ChatGPT hardware
“Every GPU, every TPU, every AI chip that was serving these models were just, like, fundamentally built before this and are retrofit to serve these modern models.”
Baker: AWS Trainium is doing best among chips trying to beat GPUs
“And you have TPU, Tranium, and AMD, Which are all you know, essentially trying to be a better GPU, and today I think probably Tranium is doing the best.”
Anthropic signs 5GW TPU deal with Google and Broadcom starting in 2027
“So last month, you know, we signed a five gigawatt deal with Google and with Broadcom for TPUs starting in twenty-twenty-seven.”
Vahdat: Google co-designs TPUs directly with power sources, buildings, and Gemini
“In other words, for us, for let's say our TPUs, we co-design them with a building. We co-design them with the power generation source. We co-design them with the DeepMind team that builds Gemini models. So it's the software above, the models above that, the ch…”
Pichai: First-party hardware will be very important in regulated AI and robotics domains
“My lesson from Waymo and on the AI side with TPUs, et cetera, I think to really push the curve well, particularly in areas where you have safety, regulatory, everything, you want the firsthand experience of the product feedback cycle. So I think having first p…”
Kim: Stacking memory directly on GPUs will be a major hardware innovation
“Both chief scientists talked about stacking memory right on top of the GPU or TPU, and that's gonna be a huge innovation in the coming months or years.”
Patel: Google will buy tons of GPUs through 2027 due to TPU limits
“When we look in 26, Google would buy a lot more TPUs, but they can't ramp production fast enough, right? And so they have to buy tons of GPUs. And we go to 27, it applies again, right? Google simply cannot buy enough TPUs, and they have to buy tons of GPUs.”
O'Laughlin: TPU v7 is Google's peak TCO advantage over Nvidia
“I think Ironwood a V seven is the peak gap between on Between TCO, between NVIDIA and TPU, right?”
O'Laughlin: Google TPU business could be worth $1 trillion
“I've done the math. It could be like, it's like a trillion. It's like a trillion. It's like a trillion or something like that. Assuming it gets like 30% market share or something like that.”
Jeff Dean: ML chip design requires predicting research workloads 2-6 years out
“As a hardware designer for ML in particular, you're trying to design a chip starting today
And that design might take two years before it even lands in a data center, and then it has to sort of be a reasonable lifetime of the chip to take you three, four, or f…”
David George: Google's 7-to-8-year-old TPUs maintain 100% utilization
“Seven to eight year old TPUs, Google actually disclosed this, seven to eight year old TPUs actually have 100% utilization.”
Lemkin: Anthropic signed a deal for one million TPUs
“Anthropic signed a deal for a million TPUs, right? And is deep on the whole Amazon Tranium ecosystem.”
Patel: Google is splitting TPU production between Broadcom and MediaTek
“Yeah, so, so for the longest time Google's had, ah, one main line of TPUs, right? All made by Broadcom, and then sort of next year they've diverged it, right? Where Broadcom makes a TPU, ah, and MediaTek makes a TPU. These two TPUs are focused at different thi…”
Bourgeau: Full-stack hardware and TPU understanding is a superpower for AI research
“So being able to understand how the stack works all the way down from TPUs to research is kind of a superpower, because then you're able to kind of find these gaps in between different layers that other people weren't necessarily able to see, but also to reaso…”
Coogan: Anthropic has ordered $21 billion worth of TPUs
“Anthropic has ordered twenty-one billion dollars worth of TPUs to train to train large claws.”
Nvidia's Rubin architecture will significantly widen its lead over custom ASICs
“And then when Rubin comes out, we'll know the gap. The gap is going to expand significant versus TPUs versus TPUs and all other ASICs.”
Few custom AI ASICs beyond Amazon Trainium and Google TPU will succeed
“I will be surprised if there are a lot of ASICs other than Tranium and TPU.”
Amazon Trainium and Google TPU will eventually transition to customer-owned tooling
“And by the way, in Tranium and TPU, we'll both run on customer owned tooling at some point. We can debate when that will happen, but the economics of success that I just described mean it's inevitable. Like no matter what the companies say, just the economics …”
Jessica Lessin: Meta is in serious talks to buy Google TPUs
“One, Meta is considering, and in very serious talks, over a billions of dollars, spending billions of dollars on TPUs in a future deal”
Jason Lemkin: Seamless switching between TPUs and GPUs undercuts Nvidia's moat
“The idea that NVIDIA is unstoppable because of the software and hardware connection, because we have to have GPUs. I know there's a lot of truth to that, but literally as an end user, I went right back and forth to them today. No, no issue. TPUs, GPUs, LLMs, a…”
Chamath: AI decoding chip market will quickly become highly fragmented
“What we are quickly seeing is that there's going to be a highly fragmented layer of decoding chips that exist in the marketplace. Grok is one. TPU is one. Microsoft has a spin. Amazon has Inferentia. Facebook, I think, is Apparently spinning up their own silic…”
Coogan: Google trained Gemini 3 Pro entirely on TPUs, bypassing Nvidia
“Google trained Gemini three pro on Google's own TPUs. No mention of Nvidia chips.”
Coogan: Nvidia retains monopoly power because Google refuses to sell TPUs
“If TPUs are not for sale, Nvidia does have a monopoly. If Nvidia truly is the only seller in the market because Google is not a seller then yes, they still extract monopoly power from every other buyer because every other buyer says, yeah, I'd love to buy TPUs…”
Vahdat: Google's seven- and eight-year-old TPUs maintain 100% utilization
“Our seven and eight year old TPUs have a hundred percent utilization.”
Vahdat: Google TPUs are 10x to 100x more energy efficient than CPUs
“TPU, I'll use that example again because I know it best for certain computation, is somewhere between 10 and a hundred times more efficient per watt, and it's this watt that really matters than a CPU.”
Lessin: Gemini and Anthropic Are Trained on TPUs Rather Than GPUs
“If that's true, then the top two coding models Gemini and Anthropic are trained on TPUs, not GPUs. And I think that's a really big deal.”
Lessin: Custom AI Chips Do Not Need to Match NVIDIA's Performance
“They make a mistake thinking that a TPU or a Tranium or from Amazon or something a company develops itself has to be as good as NVIDIA. It doesn't because NVIDIA's margin is so big. It can actually be a fraction of as good as NVIDIA and still cost effective.”
Siegler: Meta lacks Google's cloud infrastructure and custom TPUs
“Unlike meta, which is also profitable and Zuck talks up like, yeah, we can fund this via profits, which is true, but they don't have the cloud infrastructure that Google does and they don't have the TPUs that Google does and they don't have all of the sort of …”
Google TPUs Are Only Scaled AI Chip Deployment Besides Nvidia
“They're a chip company with their Tensor Processing Units, or TPUs, which is the only real scale deployment of AI chips in the world besides Nvidia GPUs.”
Google Is Only Company with Both Frontier Model and AI Silicon
“Somebody put it to me in research that if you don't have a foundational frontier model, Or you don't have an AI chip, you might just be a commodity in the AI market, and Google is the only company that has both.”
Google Designed and Deployed First TPU in 15 Months
“So, the TPU was designed, verified, built, and deployed into data centers in 15 months.”
Google Operates an Estimated 2 to 3 Million Custom TPUs
“But today, Google, it's estimated, has two to three million TPUs. For reference, Nvidia shipped, people don't know for sure, somewhere around four million GPUs last year.”
Neoclouds Will Offer Google TPUs in the Coming Months
“I think within a year it might happen. There are rumors already that some NeoClouds in the coming months are gonna have TPUs.”
Patel: Google has the lowest token COGS due to TPUs
“They have the lowest cost of goods sold for any token of any company because they have their own vertical stack on TPUs.”
Ross: Google ran three parallel chip projects, but only TPU succeeded
“So, people look at the TPU as a big success, and what they don't realize is that there were about three chip efforts at Google at the same time and only one of them ended up outperforming GPUs.”
Huang: Startups should capture entire nascent markets rather than entering mature ones
“You're supposed to create a startup before the market grows. You're not supposed to come up as a startup when the market's a trillion dollars large. You know, this fallacy, and all VCs know this fallacy that a large market, if you could just take a few percent…”
Patel: Google is internally discussing selling physical TPU hardware to buyers
“I think Google's even discussing it. Internally, I think it would require a big reorg of culture and a big reorg of like how Google Cloud works and how the TPU team works and how the JAX software team and XLA software teams work.”
Google Cloud Missed Revenue From GPU Shortages but Boosted Margins
“Google Clouds revenue numbers disappointed, but this was because they didn't have enough GPUs and they were actually constrained. So they missed on top line, like they didn't bring in enough revenue, but their margins actually improved. And so what that means …”
Coogan: Google Is the Only Company With an At-Scale Nvidia ASIC Alternative
“Google is the only company, we talked about the TPU thing, they're the only company with an at-scale ASIC alternative to NVIDIA's GPUs, which should give them a meaningful cost advantage to that end, ah, to the extent that cloud compute becomes commoditized.”
Brin: Google mostly uses custom TPUs over GPUs for Gemini
“Well, we mostly, for Gemini, we mostly use our own TPUs.”
Morin: Google TPUs lack commercial success outside of Google
“They are very much successful inside of Google, but not much outside of Google, let's say, right?”
Morin: Bottom-up AI infrastructure strategies fail because developers do not care
“I think that if you are doing it bottom up, infra to applications, you will lose because nobody will care. As they don't today, right? If you look at TPUs, they're available, they're great. Nobody cares.”
Morin: Google TPUs have more mature software and compute than Nvidia
“But AMD can do training, but it's also, but in terms of maturity, the, by far the most mature software and compute is TPUs, and then it's Nvidia.”
Patel: Google built rack-scale AI systems with Broadcom in 2018 before NVIDIA
“Google actually did this alongside Broadcom you know, and they did it before Nvidia, right? You know, today everyone's freaking out about, or not freaking out, but like everyone's like very excited about Nvidia's Blackwell system, right? It is a rack Of GPUs. …”
Patel: NVIDIA is cutting Blackwell margins to compete with custom ASICs
“Like with Blackwell, not only is it way, way, way faster, anywhere from 10 to 15 times on really large models for inference, because they've optimized it for very large language models, they've also decided, hey, we're gonna cut our margin, too, somewhat, beca…”
Patel: Google TPU clusters scale up to 8,000 chips today
“Nvidia's talking about GB 200, NVL 72, TPUs go to 8000 today, right?”
Patel: Apple accounts for over 70% of Google's TPU rental revenue
“There's only one company accounts for over 70% of Google's revenue from TPUs as far as I understand, and that's Apple.”
Patel: Google TPU purchases will pause for six months over space limits
“Like in the next six months there is a bit of a slowdown in Google TPU purchases because they have no data center space. They want more. They just literally have no data center space to put them.”
Frankle: Google TPUs provide much higher bandwidth-to-compute ratios
“TPUs have a very different network bandwidth to compute ratio. They have a lot more bandwidth just objectively and TPUs per chip tend to be a little bit less compute intensive and have a little bit less memory.”