Assertion Not checkable as stated
Balaban: AI scaling laws show no signs of hitting a limit
“The part of which makes me feel so confident that there's going to continue to be demand is that we continue to see no end to the scaling laws, which are like the underlying idea that you put more compute in and you get better intelligence levels out of your m…”
Assertion Not checkable as stated
Balaban: Claims that AI GPUs have a five-year lifespan are wrong
“The usable life is longer than the accounting depreciation schedule. And what really matters is the economic usable life. And so what we're starting to see is that like the people who are the naysayers, oh, this is going to be, you're going to throw these GPUs…”
Assertion Not checkable as stated
Balaban: Only xAI and Lambda execute high-velocity AI compute deployments
“There's two people in the world that can, and two companies in the world that can do high velocity deployments, SpaceX AI and Lambda”
Opinion
Balaban: AI cloud compute is not a commodity service
“The big thing is that cloud compute is not a commodity service. It is a very complicated, highly vertically integrated type of service that spans everything from land, land entitlement, Construction, HPC, high performance computing design, software, virtualiza…”
Prediction Not checkable as stated
Balaban predicts the neocloud market will support multiple large players
“I think it's absolutely room for multiple very large players, just like the traditional cloud business has shown that there's room for multiple large winners and multiple large players.”
Assertion Not checkable as stated
Balaban: The AI industry continues to underbuild compute infrastructure
“Well, I think that we continue to be generally under building.”
Opinion
Balaban: NVIDIA's real software moat is cuDNN, not just CUDA
“One of the big moats they've got is just The QDNN stack. It's not just CUDA. It's, you know, CUDA is sure. That's like the water we all swim, but like CUDNN has got so many, you know, matrix multiplication, routine optimizations baked into it.”
Insight
Balaban: Modern AI applications are far less latency-sensitive than legacy cloud apps
“The old school traditional legacy cloud business was so latency focused because of some of the applications, but this new fleet of AI applications are far less latency sensitive.”
Assertion Not checkable as stated
Balaban: Lambda is leasing 2023-deployed H100 GPUs at higher rates today
“You actually look at the chips that we deployed in twenty-twenty-three, H-one hundreds. We're now leasing those out at a higher rate. Now than we were originally in 20, 23.”
Prediction Not checkable as stated
Balaban: xAI's 200-day data center build record can be beaten
“I think it can be matched or beat.”
Prediction Not checkable as stated
Balaban: Multimodal AI will eventually render every pixel of software directly
“I think that for a lot of the pieces of software on your computer, you might see that taking over where, you know, you can get the glimpse of the future with this ASCII art, and then eventually it'll also have a multimodal network that's generating every pixel…”
Prediction Not checkable as stated
Balaban: Everyone in the US will eventually require at least one GPU
“I believe that in the future, everybody in the United States will need the computational power of one GPU or more to just do their daily work You know, enjoy life, whether it's getting access, whether it's getting entertained, whether it's being productive, wh…”
Assertion Supported
Balaban: Most neocloud competitors cannot launch online clusters over 32 GPUs
“Most of the other NeoClouds either don't have the ability to launch a cluster from their website or max out, I'll say, 32 GPUs.”
Prediction Not checkable as stated
Balaban does not foresee model disruptions causing compute demand to decline
“So I don't really foresee a very likely outcome where we have this huge model disruption that would cause a decline in the demand for compute.”
Assertion Not checkable as stated
Balaban: Utility power commitments and MEP equipment are AI's main bottlenecks
“But broadly in the industry, the thing that is the main bottleneck is basically land-powered shell, which is basically land that is entitled to have a certain amount of megawatt commitment from a utility. And then of course the data center and the mechanical e…”
Assertion Supported
Balaban: Practically no new US data centers use evaporative cooling
“Practically no new builds in the United States are using evaporative cooling for doing the, these closed loop direct to chip liquid cooling systems.”
Assertion Supported
Balaban: Gigawatt AI clusters require up to $45B for servers alone
“If you were to talk about the capital stack, let's say you can go back down to power generation, two to three million dollars a megawatt, two to three billion dollars a gigawatt for a power plant. The data center is between 10 and fifteen billion dollars a gig…”
Prediction Not checkable as stated
Balaban: Mass adoption of neural software will begin in 10 to 15 years
“I would say that generally speaking, when I'm early on something, I tend to be about A decade to a decade and a half early. So I would say that between a decade and 15 years, we will see mass adoption beginning or otherwise happening for neural software.”
Opinion
Balaban: AI agentic workflows without clear automated feedback are overhyped
“I think a lot of the sort of agent, agentic workflows for things that are not software engineering, I think tend to be overhyped. And I'll tell you that the reason for that is because one of the ways that you get an agentic workflow working really well is that…”
Assertion Partly supported
Balaban: NVIDIA is the only chip provider in every major cloud
“They're the only server provider, the only chip provider that is available in every single major cloud platform, which is a huge platform advantage.”
Assertion Supported
Balaban: Top AI labs use multiple chip types for training and inference
“The biggest labs in the world are using multiple different Types of chips to do their inferencing and training on.”
Prediction Not checkable as stated
Balaban: Complex GPU financial securities may eventually emerge as compute matures
“I think that the, that market is starting to mature that, that, that may be an eventuality is having more complex securities that surround GPUs. But I think for right now, people are starting to realize that it's a great credit investment and that's what's cha…”
Disclosure
Balaban: Lambda's cloud revenue run rate is just under $1B
“And now it's at, you know, a little bit under a billion dollar revenue run rate. We've fully exited the hardware business.”
Assertion Supported
Balaban: Lambda alumni startup Positron is valued over $1B
“He eventually left and joined another former Lambda team member, Thomas Summers to start Positron, which is an accelerator company. And they're like now valued at over a billion dollars.”