Stephen Balaban

Co-Founder & CTO, Lambda · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderexecutiveengineer@stephenbalaban ↗LinkedIn ↗stephenbalaban.com ↗

Balaban co-founded Lambda in 2012, leading its evolution from early facial-recognition software into a multibillion-dollar AI infrastructure and GPU cloud provider. Prior to Lambda, he was the first engineering hire at Perceptio, an edge machine-learning startup that was acquired by Apple.

29statements → 21claims → 7claims resolved → 86%fully supported → 3.72/5average certainty → 2.48/5average debate potential → 4.3/5argument clarity · the sources →

6 supported 1 partly supported 0 contradicted 14 not checkable as stated how the 21 claims stand · each chip opens the sources

7 predictions · 14 assertions · 3 opinions · 2 insights · 3 disclosures · every statement was checked. The predictions and assertions are the 21 claims: statements the public record can support or contradict. 7 are resolved, and 14 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Stephen argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Balaban: Most neocloud competitors cannot launch online clusters over 32 GPUs
“Most of the other NeoClouds either don't have the ability to launch a cluster from their website or max out, I'll say, 32 GPUs.”
Stephen Balaban Jun 18, 2026 ▶ 4:49 The GPU Myth: State of AI Compute 2026 | Stephen Balaban

Expressed certainty vs assessment result

none yet certainty 1
none yet certainty 2
100% certainty 3
100% certainty 4
50% certainty 5

weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet

Argument clarity: do they answer the question? how? →

4.3 / 5 directness 4.6 · coherence 4.5 · precision 4.3 · compression 3.9

answered every one of 14 assessed questions directly

This is a score against a rubric. It is not a rank. Every host question → answer exchange is scored with names hidden on directness, coherence, precision and compression, 1–5 each, on meaning alone: disfluencies are ignored, and only raw unedited episodes count. This is the score that measures thought. Every scored exchange, scores shown → · The rubric and its checks →

How they sound: speaking style how? →

224 words/min while actually speaking · 19.6 um and uh per 1k words

Measured by listening to the audio itself: 11,018 words across 1 episode of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything Stephen Balaban said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Assertion Not checkable as stated
Balaban: AI scaling laws show no signs of hitting a limit
“The part of which makes me feel so confident that there's going to continue to be demand is that we continue to see no end to the scaling laws, which are like the underlying idea that you put more compute in and you get better intelligence levels out of your m…”
Stephen Balaban Jun 18, 2026 ▶ 8:07 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Not checkable as stated
Balaban: Claims that AI GPUs have a five-year lifespan are wrong
“The usable life is longer than the accounting depreciation schedule. And what really matters is the economic usable life. And so what we're starting to see is that like the people who are the naysayers, oh, this is going to be, you're going to throw these GPUs…”
Stephen Balaban Jun 18, 2026 ▶ 42:14 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Not checkable as stated
Balaban: Only xAI and Lambda execute high-velocity AI compute deployments
“There's two people in the world that can, and two companies in the world that can do high velocity deployments, SpaceX AI and Lambda”
Stephen Balaban Jun 18, 2026 ▶ 57:02 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Opinion
Balaban: AI cloud compute is not a commodity service
“The big thing is that cloud compute is not a commodity service. It is a very complicated, highly vertically integrated type of service that spans everything from land, land entitlement, Construction, HPC, high performance computing design, software, virtualiza…”
Stephen Balaban Jun 18, 2026 ▶ 1:50 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban predicts the neocloud market will support multiple large players
“I think it's absolutely room for multiple very large players, just like the traditional cloud business has shown that there's room for multiple large winners and multiple large players.”
Stephen Balaban Jun 18, 2026 ▶ 6:07 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Not checkable as stated
Balaban: The AI industry continues to underbuild compute infrastructure
“Well, I think that we continue to be generally under building.”
Stephen Balaban Jun 18, 2026 ▶ 7:01 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Opinion
Balaban: NVIDIA's real software moat is cuDNN, not just CUDA
“One of the big moats they've got is just The QDNN stack. It's not just CUDA. It's, you know, CUDA is sure. That's like the water we all swim, but like CUDNN has got so many, you know, matrix multiplication, routine optimizations baked into it.”
Stephen Balaban Jun 18, 2026 ▶ 27:03 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Insight
Balaban: Modern AI applications are far less latency-sensitive than legacy cloud apps
“The old school traditional legacy cloud business was so latency focused because of some of the applications, but this new fleet of AI applications are far less latency sensitive.”
Stephen Balaban Jun 18, 2026 ▶ 38:03 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Not checkable as stated
Balaban: Lambda is leasing 2023-deployed H100 GPUs at higher rates today
“You actually look at the chips that we deployed in twenty-twenty-three, H-one hundreds. We're now leasing those out at a higher rate. Now than we were originally in 20, 23.”
Stephen Balaban Jun 18, 2026 ▶ 40:40 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban: xAI's 200-day data center build record can be beaten
“I think it can be matched or beat.”
Stephen Balaban Jun 18, 2026 ▶ 57:45 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban: Multimodal AI will eventually render every pixel of software directly
“I think that for a lot of the pieces of software on your computer, you might see that taking over where, you know, you can get the glimpse of the future with this ASCII art, and then eventually it'll also have a multimodal network that's generating every pixel…”
Stephen Balaban Jun 18, 2026 ▶ 1:00:49 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban: Everyone in the US will eventually require at least one GPU
“I believe that in the future, everybody in the United States will need the computational power of one GPU or more to just do their daily work You know, enjoy life, whether it's getting access, whether it's getting entertained, whether it's being productive, wh…”
Stephen Balaban Jun 18, 2026 ▶ 1:11:18 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Supported
Balaban: Most neocloud competitors cannot launch online clusters over 32 GPUs
“Most of the other NeoClouds either don't have the ability to launch a cluster from their website or max out, I'll say, 32 GPUs.”
Stephen Balaban Jun 18, 2026 ▶ 4:49 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban does not foresee model disruptions causing compute demand to decline
“So I don't really foresee a very likely outcome where we have this huge model disruption that would cause a decline in the demand for compute.”
Stephen Balaban Jun 18, 2026 ▶ 10:33 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Not checkable as stated
Balaban: Utility power commitments and MEP equipment are AI's main bottlenecks
“But broadly in the industry, the thing that is the main bottleneck is basically land-powered shell, which is basically land that is entitled to have a certain amount of megawatt commitment from a utility. And then of course the data center and the mechanical e…”
Stephen Balaban Jun 18, 2026 ▶ 11:11 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Supported
Balaban: Practically no new US data centers use evaporative cooling
“Practically no new builds in the United States are using evaporative cooling for doing the, these closed loop direct to chip liquid cooling systems.”
Stephen Balaban Jun 18, 2026 ▶ 14:06 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Supported
Balaban: Gigawatt AI clusters require up to $45B for servers alone
“If you were to talk about the capital stack, let's say you can go back down to power generation, two to three million dollars a megawatt, two to three billion dollars a gigawatt for a power plant. The data center is between 10 and fifteen billion dollars a gig…”
Stephen Balaban Jun 18, 2026 ▶ 24:07 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban: Mass adoption of neural software will begin in 10 to 15 years
“I would say that generally speaking, when I'm early on something, I tend to be about A decade to a decade and a half early. So I would say that between a decade and 15 years, we will see mass adoption beginning or otherwise happening for neural software.”
Stephen Balaban Jun 18, 2026 ▶ 1:03:22 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Opinion
Balaban: AI agentic workflows without clear automated feedback are overhyped
“I think a lot of the sort of agent, agentic workflows for things that are not software engineering, I think tend to be overhyped. And I'll tell you that the reason for that is because one of the ways that you get an agentic workflow working really well is that…”
Stephen Balaban Jun 18, 2026 ▶ 1:12:13 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Partly supported
Balaban: NVIDIA is the only chip provider in every major cloud
“They're the only server provider, the only chip provider that is available in every single major cloud platform, which is a huge platform advantage.”
Stephen Balaban Jun 18, 2026 ▶ 25:37 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Supported
Balaban: Top AI labs use multiple chip types for training and inference
“The biggest labs in the world are using multiple different Types of chips to do their inferencing and training on.”
Stephen Balaban Jun 18, 2026 ▶ 28:53 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Prediction Not checkable as stated
Balaban: Complex GPU financial securities may eventually emerge as compute matures
“I think that the, that market is starting to mature that, that, that may be an eventuality is having more complex securities that surround GPUs. But I think for right now, people are starting to realize that it's a great credit investment and that's what's cha…”
Stephen Balaban Jun 18, 2026 ▶ 43:30 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Disclosure
Balaban: Lambda's cloud revenue run rate is just under $1B
“And now it's at, you know, a little bit under a billion dollar revenue run rate. We've fully exited the hardware business.”
Stephen Balaban Jun 18, 2026 ▶ 50:21 The GPU Myth: State of AI Compute 2026 | Stephen Balaban
Assertion Supported
Balaban: Lambda alumni startup Positron is valued over $1B
“He eventually left and joined another former Lambda team member, Thomas Summers to start Positron, which is an accelerator company. And they're like now valued at over a billion dollars.”
Stephen Balaban Jun 18, 2026 ▶ 51:27 The GPU Myth: State of AI Compute 2026 | Stephen Balaban

Show 5statements(5 left)

Appearances (1)

EpisodeDateSpeaking time
The GPU Myth: State of AI Compute 2026 | Stephen Balaban Jun 18, 2026 1h 0m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.