Everything Steeve Morin said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Morin: Nvidia is far from the most efficient AI hardware platform
“But it's by far not the most efficient platform. And arguably, even in terms of software, it's not the best software platform.”
Morin: The market bubble around Nvidia H100 GPUs will burst
“There's going to be a need for inference. Very hard to say whether it will be worth, you know, everybody's money to do it on H 100. That is a bubble that I think will blow some time.”
Morin: GPUs are a clever workaround, not natively built for AI
“GPUs are, you know, are a good trick for AI, but they're not built for AI.”
Morin: AI chip oversupply will lead to GPUs selling at 30% value
“I very much worry there will be an oversupply of these chips. The problem is, is that, you know, remember, the chips are the collateral. So, you know, somewhere, you know, in the US or whatever, there's going to be a data center with like a thousand GPUs that …”
Morin: GPUs cannot deliver latent space AI reasoning at scale
“Fundamentally, GPUs cannot deliver, deliver this, plain and simple at scale.”
Morin: Nvidia may lose market dominance in AI inference and training
“I think that there's a shot that they don't.”
Morin: Google TPUs have more mature software and compute than Nvidia
“But AMD can do training, but it's also, but in terms of maturity, the, by far the most mature software and compute is TPUs, and then it's Nvidia.”
Morin: Stargate data center is inefficient vertical scaling for AI
“It is a vertical scaling. And as you know, my days are spent on efficiency. So I look at these things as being like, all right, this is a bigger, you know, this is an American car of AI. It's big. It consumes a lot of gas, but ultimately, you know, it's not a …”
Morin predicts AI compute will be 95% inference within five years
“In five years, I would say 95% inference, five percent training.”
Morin: Google is the sleeping giant of the AI race
“Google has, like, you know, Android, Google Docs, Whatever, they have everything, they can sprinkle everywhere. This is the sleeping giant in my mind.”
Morin: Running standalone AI model weights will eventually become obsolete
“Models in the sense of, you know, getting, you know, weights and running them is something that is ultimately going away because you know, in favor of like full blown backends, right? You feel like you're talking to a model, but ultimately you're talking to an…”
Morin: Switching from Nvidia to AMD offers 4x spend efficiency
“A simple example is if you know, switch from Nvidia to AMD on a seven TB model, you can get four times better efficiency, right? In terms of spend.”
Morin: Nvidia H100 costs 5x A100 price for 2x inference speed
“H 100 comes along and inference is it's worth five times the price. And it may be runs twice in terms of performance on inference. That is on training. It's a lot better, but on inference, it's like maybe twice as fast when it actually, when it came out, it ra…”
Morin: AI agents and reasoning will disrupt Nvidia's chip dominance
“Ultimately, the two things that could really, very much shake the industry, the chip industry, in my opinion, is our agents and reasoning.”
Morin: Nvidia Blackwell chips suffered surface bending causing cooling issues
“For Blackwell, they assembled two chips. But the surface was so big that the chip started to, you know , wave, like, I don't know the English word, but like, you know, started to bend a bit, which further perpetuated the problem because it then didn't make con…”
Morin: Nvidia won AI training via Mellanox interconnects, not raw compute
“The reason probably Nvidia won, at least in the training space, is because of Mellanox, right? Not because of the raw compute.”
Morin details profit margins across TSMC, Nvidia, and cloud providers
“Nvidia, like a TSMC sells you at 60% margin. Nvidia sells you at, you know, 90% margin. And on top of that, there's Amazon that takes, let's say a 30% margin.”
Morin: Google TPUs lack commercial success outside of Google
“They are very much successful inside of Google, but not much outside of Google, let's say, right?”
Morin: Cloud compute hoarding creates fake AI GPU scarcity
“So in, in the case of, you know, Amazon or Google, that would be buying reserved compute, which you're not going to use because if you buy it on demand, you will get tremendously ripped off. So that creates this like face scarcity of compute because that peopl…”
Morin: Nvidia Blackwell chip shipments are delayed and orders are canceled
“Blackwell is late and orders are getting canceled.”
Morin: Nvidia will remain dominant in AI inference due to availability
“The thing is these chips are on the market. They're here. I can, you know, out tab on Chrome and get one. That is something that, you know, I don't take lightly. Availability that is right. So I think Nvidia is used to stay at least if not for the H-one hundre…”
Morin: Microsoft's AMD chip deployments made OpenAI inference profitable
“Microsoft comes along and buys it all, makes, by the way, OpenAI, or at least on the inference side, puts OpenAI in the green because of the efficiency gains.”
Morin: Being 7x better on cost won't get customers off Nvidia
“I know for a fact that being seven times better and whatever, take whatever metric you want. Whether it's spend, whether it's whatever. It's not enough to get people to switch. People will choose nothing over something.”
Morin: Doubling GPUs in AI inference yields only 10% performance gain
“If you go from one GPU to two, you don't get twice the performance. Maybe you get 10% better performance. Yeah, that's the dirty secret nobody talks about. I'm talking inference, right? So, so you go from, let's say, a hundred to a 110 by doubling the amount o…”