Assertion Supported
ChatGPT reached 100 million active users faster than any application
“First, with OpenAI's ChatGPT, which became the fastest app in history to a hundred million active users”
Assertion Partly supported
Nvidia projected a $1 trillion TAM by capturing 1% of global industry
“In one of Nvidia's earnings slides in 2021. They put up their total addressable market and they said they had a one trillion dollar TAM. And the way that they calculated this was that they were going to serve customers who Provided a hundred trillion dollars w…”
Assertion Supported
AlexNet reduced ImageNet mislabeling error rates from 25% to 15%
“I think the error rate went from mislabeling images 25% of the time to suddenly only mislabeling them 15% of the time, and that was like a huge leap over the tiny incremental progress that had been made along the way.”
Assertion Supported
AlexNet was trained on two consumer Nvidia GeForce GTX 580 GPUs
“And what these guys from Toronto did is they went out probably to their local Best Buy or equivalent in Canada. They bought two GeForce GTX-Five-Eighty's, which were the top-of-the-line cards at the time, and they wrote their algorithm, their convolutional neu…”
Insight
Parallel computing on GPUs acts as an Archimedes lever on Moore's Law
“Whatever advances are happening in Moore's law and the number of transistors on a chip, If you have an algorithm that can run in parallel, which is not all problem spaces, but many can, then you can basically lever up Moore's Law by hundreds of times, or thous…”
Assertion Supported
AI pioneer Geoffrey Hinton descends from mathematicians George and Mary Boole
“He is the great-great-grandson of George and Mary Boole.”
Opinion
AI feed recommendations made Instagram a $100B to $500B asset for Meta
“Instagram would have been a great acquisition anyway, but it was AI-powered recommendations in the feed that made that into a 102 105 hundred billion dollar asset for Facebook.”
Assertion Supported
OpenAI was founded to reach artificial general intelligence before major tech companies
“The founding of OpenAI was motivated by the desire to find AGI or Artificial General Intelligence first before the big tech companies did.”
Assertion Supported
Ilya Sutskever left Google in 2015 to co-found OpenAI
“So after the dinner, Ilya leaves Google and signs up to become, as we said, co-founder and chief scientist of a new independent AI non-profit research lab backed by Elon and Sam. OpenAI.”
Insight
Language structure and knowledge are embedded directly within raw text data
“The very structure of language and the way to interpret knowledge is actually embedded in the training data itself rather than requiring labeling.”
Insight
Ilya Sutskever argued next-token prediction inherently requires true world understanding
“The more accurately an LLM predicts that next word, i.e. The name of the criminal. Ipso facto, the greater its understanding, not only of the novel, but of all general human-level knowledge and intelligence, because you need all of your experience in the world…”
Insight
Transformer attention computational cost scales quadratically with input prompt length
“Traditionally, you'd say this is very, very inefficient, and it actually means that the larger your context window, aka token limit, aka prompt length, gets, the more computationally expensive it gets on a quadratic basis. So doubling your input means quadrupl…”
Insight
Transformers enabled parallel training of sequence models on GPUs
“So the big innovation here is you could now train sequence-based models In a parallel way. You couldn't train models of this size at all before, let alone cost effectively.”
Opinion
Google possesses strong technical AI capabilities but lacks fundamental product sense
“I mean, this is Google here. Like, the capabilities are there. The product sense. Not as much.”
Assertion Not checkable as stated
Training early large transformer models was financially untenable for anyone except Google
“GPUs and NVIDIA and the transformer made it possible, but to work with the size of models you're talking about spending an amount of money that's certainly for a nonprofit and anybody really except Google was untenable.”
Assertion Supported
GPT parameter counts grew from 120 million in GPT-1 to 1.7 trillion
“GPT-I had roughly a hundred and twenty million parameters that it was trained on. GPT-II had 1.5 billion. GPT-III had a hundred and seventy five billion, and GPT-IV, OpenAI hasn't announced, but it's rumored that it has about 1.7 trillion parameters that it wa…”
Insight
Scaling compute and data created unpredicted emergent reasoning in language models
“We don't change anything about the structure, we just give it way more data and let it Run these models for a long time and make the parameters of the model way bigger, and like, no researchers expected them to reason about the world as well as they do, but it…”
Assertion Supported
Google built TPUs to offset compute costs while still buying Nvidia hardware
“Even Google at this point in time, this is when they start building their own chips, TPUs, because, you know, they're still buying tons of hardware from NVIDIA, but they're also starting to source their own here.”
Opinion
Generative AI and cloud compute convergence is Nvidia's greatest growth tailwind
“And the combination of those three things turns out to be basically the single greatest moment that could ever happen for NVIDIA.”
Prediction Held up
Predominant enterprise AI GPU compute consumption will occur via the cloud
“It looks like the predominant way that companies are going to use that compute is gonna be in the cloud.”
Assertion Supported
Nvidia spent five years building platforms to replace Intel's x86 data centers
“NVIDIA has literally just spent the past five years working insanely hard to build a new computing platform for The data center, a GPU accelerated computing platform to, in their minds, replace the old CPU led Intel dominated x-eighty six architecture in the d…”
Insight
Nvidia bypassed the von Neumann bottleneck via massive hardware parallelism
“So the magical unlock, of course, is to make a computer that is not a von Neumann architecture. To make programs executable in parallel and massively increase the number of processors or cores. And that is exactly what NVIDIA did on the hardware side, and all …”
Insight
On-chip memory capacity is the primary hardware bottleneck in AI training
“The constraint today is actually in how much high-performance memory is available on the chip. These models need to be in memory all at the same time, and they take up hundreds of gigabytes. So while memory has scaled up, I mean, we're gonna get flashing all t…”
Assertion Supported
Semiconductor chips cannot be etched larger due to extreme ultraviolet photolithography limits
“Due to a quirk of the extreme ultraviolet photolithography that we talked about, the EUV on the TSMC episode, chips are already the full size of the reticle. It's a physics and wavelength constraint. You really can't etch chips larger without some new inventio…”
Opinion
Nvidia's $7 billion acquisition of Mellanox is one of tech's best ever
“They made one of the best acquisitions of all time back in 2020, and nobody had any idea. They bought a quirky little networking company out of Israel called Mellanox.”
Assertion Contradicted
Nvidia split consumer and data center GPU architectures in September 2022
“Up until NVIDIA's current GPU generation, the hopper generation of GPUs for the data center, there was only one GPU architecture at NVIDIA, and that same architecture and those same chips from the same wafers made at TSMC, some of them went to consumer gaming …”
Assertion Supported
Nvidia bifurcated its architecture to monopolize TSMC's CoWoS packaging capacity
“And by NVIDIA bifurcating their chip architectures into a gaming segment that does not have this latest COWAS technology, this allows them to monopolize, like, a huge amount of TSMC's capacity to make the COWAS chips, specifically for these H-one hundreds, whi…”
Assertion Open · timeframe Sep 2023
Chip-on-Wafer-on-Substrate advanced packaging accounts for 10% to 15% of TSMC's capacity
“So COOS represents right now about 10 to 15% of TSMC's capacities, and many of the facilities are custom-built for exactly these types of chips that they're producing.”
Assertion Supported
Nvidia sells DGX artificial intelligence server systems for $150,000 to $300,000
“NVIDIA sells these DGX systems for like, a 150 to 300,000 dollars a box.”
Assertion Supported
Nvidia launched the H100 GPU in September 2022 retailing at $40,000
“So they launched it in September, twenty-twenty-two. It's the successor to the A-One hundred. One GPU, one H-One hundred, cost 40,000 dollars.”
Assertion Supported
Nvidia's H100 GPU is nine times faster for AI training than A100
“So, the reason you want an H-one hundred is they're 30 times faster than an A-one hundred, which mind you is only like two and a half years older. It is nine times faster for AI training.”
Assertion Supported
Nvidia is using artificial intelligence to design its own semiconductor chips
“They're actually using AI to design the chips themselves now.”
Assertion Supported
Cloud instances cost $30 hourly for A100s and $100 hourly for H100s
“You can get access to a DGX server. That's eight A 100 for about 30 bucks an hour, or you can go over to AWS and get a P five dot 48 X large instance, which is eight H 100, which I believe is an HGX server for about a hundred dollars an hour.”
Assertion Supported
Nvidia DGX Cloud starting prices are $37,000 monthly for an A100 system
“So starting price for DGX Cloud is 37,000 dollars a month, which will get you an A-one hundred based system, not an H-one hundred based system.”
Assertion Not checkable as stated
DGX Cloud rental pricing yields a three-month capex payback on A100 hardware
“A listener helped us out and estimated that the cost to actually build an equivalent A-one hundred DGX system Would be today something like a 120 K. Remember, this is the previous generation. This is not H-one hundreds. And you can rent it for 37 K a month. So…”
Assertion Supported
Cloud service providers account for half of Nvidia's data center revenue
“The CFO Colette Kress said on their last earnings call that about half of the revenue from the data center business unit is CSPs. And then I believe after that is the consumer internet companies, and after that is enterprises.”
Opinion
Nvidia's Q2 FY24 release was one of history's greatest corporate earnings reports
“I think this was one of, if not the most incredible earnings release by any scaled public company ever. Seriously, no matter what happens going forward, last week was a historic moment.”
Assertion Supported
Global data centers contain one trillion dollars worth of installed hard assets
“There is one trillion dollars worth of hard assets sitting in data centers around the world right now.”
Assertion Supported
Global annual capital expenditure to expand and update data centers is $250B
“Annual spend on data centers to Update and add to that CapEx is two hundred and fifty billion dollars a year.”
Assertion Supported
OpenAI is reported to have exceeded a $1 billion annualized revenue run-rate
“ChatGPT made it so OpenAI is rumored to be doing over a billion dollar run rate now. Maybe multiple single digit billions, and still growing, meaningfully.”
Assertion Partly supported
Nvidia's CUDA platform reached four million registered developers by May 2023
“If you look at the number of CUDA developers over time, it was released in 2006, It took four years to get the first 100,000 people. Then by twenty-sixteen, 13 years in, they got to a million developers. Then just two years later, they got to two million. So 1…”
Opinion
Nvidia functions fundamentally as a platform company akin to Microsoft
“They do make semiconductors, and they do make data center gear, but really they are a platform company. The right analogy for NVIDIA also is Microsoft. They make the operating system. They make the programming environment. They make many of the applications.”
Assertion Supported
Nvidia released the Megatron transformer model shortly after acquiring Mellanox in 2019
“In March of 2019, NVIDIA announced they were acquiring Mellanox for seven billion dollars in cash... In August of 2019, NVIDIA released what was at the time the largest transformer-based language model called Megatron. Eight point Three billion parameters trai…”
Assertion Supported
Nvidia achieved 70% gross margins in Q2 FY2024 and forecasted 72%
“This last quarter, they had a gross margin of 70%, and they forecasted for next quarter to have a gross margin of 72%.”
Prediction Not checkable as stated
Nvidia's gross margins will not erode significantly below 65% as shortages subside
“That's gonna go away, but I don't think this very high, you know, 65% plus margin is gonna erode too much.”
Opinion
Nvidia is the only viable platform for training GPT-class frontier AI models
“If you want to train GPT or a GPT class model, there's one option. You're doing it on NVIDIA. There's one option.”
Assertion Supported
Sales to mainland China accounted for 25% of Nvidia's 2022 total revenue
“Last year, China was 25%, or sales to mainland China, was 25% of NVIDIA's revenue.”
Assertion Partly supported
Microsoft employs five times more people per market-cap dollar than Nvidia
“They have 26,000 employees, and that sounds like a big number, but for comparison, Microsoft, whose market cap is only twice as big, has 220,000. So that is five X the number of employees per dollar of market cap going on over at Microsoft”
Assertion Supported
Nvidia generates $46 million in market capitalization for every single employee
“They have forty six million dollars of market cap per employee.”
Opinion
Nvidia operates with no designated active succession plan for CEO Jensen Huang
“I don't think there's anyone else there where they're, like, getting ready for that person to take over. I think The company is a extension of Jensen's thoughts and will and drive and belief about the future, and that's kind of what happens.”
Assertion Not checkable as stated
AMD lacks Nvidia's TSMC advanced packaging capacity and CUDA developer ecosystem
“AMD doesn't have all this capacity reserved from TSMC, at least not for the 2.5 D packaging process for the high end GPUs. AMD doesn't have the developer ecosystem from CUDA.”
Insight
Competitive moats erode and margins compress as total addressable markets become massive
“Every moat only works if the castle is sufficiently small. If the prize at the end of the finish line becomes sufficiently large, you're gonna need a bigger moat, and you need to figure out a you know, how to defend the castle harder.”
Assertion Supported
Nvidia accelerated to a six-month product release cycle following the COVID-19 pandemic
“Since COVID, NVIDIA has re-accelerated to a six month shipping cycle. They've been doing two GTCs a year, most years since COVID, which is insane for the level of technology complexity that they're doing.”
Assertion Supported
Nvidia has an installed base of 500 million CUDA-capable GPUs globally
“Today there are five hundred million CUDA capable GPUs for developers to target.”
Opinion
Enterprise buyers face zero career risk choosing Nvidia hardware over unproven competitors
“Nobody is getting fired for buying Nvidia anytime soon.”
Assertion Not checkable as stated
Most AI workloads run adequately on A100s unless training GPT-4 class models
“Most workloads can be run on A-Hundreds unless you're doing model training of GPT-IV.”
Prediction Open · timeframe Sep 2028
Nvidia will eventually operate its own complete data center buildings
“They used to have to plug into other people's servers, and then they started making servers that plugged into other people's racks and rows and architectures, and then they started making their own entire rows and walls, and at some point here, they're gonna s…”
Assertion Supported
Microsoft is the exclusive cloud infrastructure provider for OpenAI
“Microsoft is the exclusive cloud infrastructure provider for OpenAI, which runs, as far as we know, solely on NVIDIA infrastructure, but they buy it all through Microsoft.”
Prediction Not checkable as stated
A cloud provider's traditional market share will dictate their AI era dominance
“I bet the way it plays out is that where you landed in cloud one point O strongly dictates where you will land in this AI cloud era.”
Assertion Contradicted
Data center GPUs represent a $30 billion market growing to $50 billion
“Right now, GPUs In the data center are like a thirty billion dollar a year market going to like a fifty billion dollar next year”
Prediction Not checkable as stated
The AI market faces an investor crisis of confidence within 18 months
“I think there's a pretty reasonable chance that there's some falter in the next 12 to 18 months where there's a crisis of confidence among investors where at some point something will come out Where we all observe, oh, maybe GPTs aren't as useful as we thought…”
Assertion Supported
TSMC expects its AI hardware revenue to grow 50% annually through 2028
“TSMC, in their last earnings, said that AI hardware currently only represents six percent of their revenue, but all indications over there is that they expect AI revenue to grow 50% per year for the next five years.”
Prediction Not checkable as stated
A large portion of machine learning inference will move to edge devices
“I suspect a lot of inference will get done on the edge. If you think about the insane amount of compute that's walking around in our pockets that is not fully leveraged right now, there's gonna be a lot of machine learning done on phones that are gonna, like, …”
Opinion
Any challenger unseating Nvidia's AI dominance requires an unforeseen flank attack
“So I think the bottom line here is it nearly impossible to compete with them head on, and if anybody's gonna unseat NVIDIA in the future of AI and accelerated computing, it's either gonna be from some unknown flank attack that they don't see, or the future wil…”