Everything Jonathan Ross said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Ross: Specialized AI inference chip startups benefit NVIDIA's stock
“I don't know that Nvidia will ever see it this way, but I do think that those of us focusing on inference and building stuff specifically for that are probably the best thing that's ever happened for Nvidia stock because we'll take on the low margin, high volu…”
Ross: Replace one-on-ones and private emails with open communication to eliminate politics
“And so when you are leading groups of people, if you want to reduce the amount of politics and people going off and, you know, forming side cliques and all this, stop having one-on-ones, have big meetings with everyone who you want to tell something and tell t…”
Ross: Putting excess capital into startups no longer provides an advantage
“What's changed though is that unlike the Keynesian beauty contest where the winner is the one with the most money, in reality, there's a point at which you get enough money and you don't need more. And for the first time in history, the startups are not starve…”
Ross: Splitting LLM pre-fill and generation across different chips is a mistake
“What we realized was, and this is what most people get wrong when they're trying to do this themselves. They'll, they'll take the reading of tokens or what's called pre-fill, and they'll do that on one piece of hardware, and then they'll put the generation of …”
Ross: Faster inference compute directly makes AI models smarter
“Now bring in the LPU, and you can go deeper, faster, you can search faster, and so you can actually make a model smarter by making it faster. And so the realization is, and now that you've got this ability to reflect, think deeply, and change the outcome based…”
Ross: Hiring requires vetting negatives rather than searching for positive traits
“The biggest flip in my hiring was when I went from looking for positives, which is what you do when you're trying to grow talent, to looking for negatives, which is what you do when you're trying to select talent.”
Ross: Marginal cost of software code is approaching zero
“What I'm seeing is that code is becoming almost free. It's, it, the marginal cost is approaching zero.”
Ross: Countries controlling compute will control AI
“The countries that control compute will control AI, and you cannot have compute without energy.”
Ross: OpenAI and Anthropic revenues would double with double compute
“If OpenAI were given twice the inference compute that they have today. If Anthropic was given twice the inference compute that they have today, within one month from now, their revenue would almost double.”
Ross: Microsoft withheld Azure GPUs because internal usage made more money
“Microsoft in one quarter deployed a bunch of GPUs, and then announced that they weren't going to make them available in Azure because they made more money using them themselves than renting them out.”
Ross: Hyperscaler AI spending is driven by existential survival, not ROI
“That's how the hyperscalers feel. So of course they're going to be spending like drunken sailors because the alternative is that they're completely locked out of their business. So it's not a purely economical framework that they're using. It's a, do we get to…”
Groq deployed a customer feature in four hours with zero human code
“Four hours later, it was in production. Not a single line of code was written by a human being. There was no debugging done by a human being. It was all prompting.”
Ross: Tesla's Dojo custom AI chip project was recently canceled
“And when you look around the industry, you've got a bunch of people building chips, some of them are getting canceled, like Dojo recently got canceled.”
Ross: Building custom AI chips to rival NVIDIA is like replicating Google
“Going off and saying, I'm gonna build my own AI chip to compete with NVIDIA, It's a little bit like saying, you know, that Google search is pretty nice. Let's go replicate it. It's insane. Like the level of optimization, the level of design and engineering tha…”
Ross: OpenAI, Anthropic, and all hyperscalers will build custom chips
“I have no doubt that OpenAI will be able to build its own chips. I have no doubt that eventually, Anthropic will be building their own chips, that every hyperscaler will build their own chip.”
Ross: NVIDIA effectively holds a monopsony on High Bandwidth Memory
“The thing is, Nvidia effectively has a monopsony on HBM.”
Ross: Memory Suppliers Restrict HBM Supply to Protect High Profit Margins
“There's also this situation where the margin on HBM is so high, That no one wants to actually increase the supply, because then the margin goes down.”
Ross: Hyperscaler AI CapEx of $100B Annually Is Not Overspending
“So when you hear the hyperscalers talking about that, you know, seventy five billion to a hundred billion dollar a year investment, because they're building out the capacity for data centers, they're putting a lot of money up for returns that they're expecting…”
Ross: Groq LPUs Ship in 6 Months Versus 2 Years for GPUs
“You have to write a check two years in advance to get GPUs. For us, you write us a check for a million LPUs, and the first of those LPUs starts showing up six months later.”
Ross: Hardware incumbents win because AI models are optimized for existing GPUs
“So if you are the incumbent, you have an advantage because people are designing their models for your hardware. It doesn't even matter if there's a better architecture out there. It's not going to run well, so it's not a better architecture.”
Ross: NVIDIA will keep selling chips despite AI labs building custom silicon
“NVIDIA still keeps selling chips.”
Ross: AI product quality scales directly with compute spent per query
“AI doesn't work the way SAS does. In SAS, you have a bunch of engineers who go out and build a product, and the quality of that product is determined based on what those engineers did. That's not the case in AI. In AI, I can improve the quality of my product b…”
Ross: China will build 150 nuclear reactors to power domestic AI compute
“They're going to build a 150 nuclear reactors. So they're going to have enough energy, even though their chips aren't as energy efficient.”
Ross: US will maintain AI hardware advantage over China for 2-3 years
“So my expectation is that right now for the next two to three years, the United States has a clear advantage in that away game over China.”