Everything Gavin Uberti said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Uberti: Future AI token factories will cost $40B to $100B per mega-cluster
“You could have the same kind of thing for some futuristic mega cluster.
Forty billion dollars, hundred billion dollars as a giant mega token factory serving one or a handful of models for a massive number of users to get that same economies of scale thing.”
Uberti: AI models will eventually do all kernel generation superhumanly
“And when the models keep getting smarter, they'll eventually do all of it. They will become superhuman.”
Uberti: Individual trillion-dollar data centers are inevitable
“Absolutely. It is a matter of time. It's like asking, what do you say, a billion dollar fab, or a ten billion dollar fab, or a hundred billion dollar fab? It is inevitable that the economies of scale don't stop at, oh, forty billion dollars is the magic number…”
Uberti: NVIDIA Blackwell point-to-point latency is about 4,000 nanoseconds
“For example, on Blackwell chips, it can be about 4000 nanoseconds to go point to point.”
Uberti: High-speed AI decode is bottlenecked by data movement, not math
“When you do this sort of kernel's work, what you realize is that the math is relatively easy. But to get high speed decode, the thing that matters is data movement. Almost all the work that you do is optimizing how do you move data around a single chip or acro…”
Uberti: TSMC funded an experiment on Etched's recommendation and updated its line
“TSMC customer service is way, way better than I have seen at any other company in any other industry. It's the kind of thing where if you say, hey, you can approve your yield by making this change, you can go make them a recommendation, and then we'll go run a…”
Uberti: TSMC will continue to win due to superior customer service
“It is why they're the number one and why they're going to win.”
Uberti: The best kernels are written by human-AI collaboration
“Today, it's all very hybrid, and the best kernels are still written by human AI collaborations.”
Uberti: Employees quit Etched over deemed unsolvable chip design problem
“We had people quit. That people literally were like, this problem is unsolvable, and, ah, best of luck guys.”
Uberti: Synopsys allowed Etched multi-year deferred payment for emulators
“Synopsis actually went ahead and let us get some of their emulators on extremely favorable terms where we pay over many years.”
Uberti: Compute costs will drop faster than memory costs due to DRAM limits
“Generally, loading data is very expensive, and doing math is very cheap. And as time goes on, you'll end up finding that math gets cheaper at a rate that is faster than memory gets cheaper, due to this fundamental limit on any kind of DRAM device.”
Uberti: AI Token Serving Requires Non-Linear Cluster Scaling Economics
“That the way you want this to scale is not that, oh, if I want to go serve 10 times more tokens, I buy 10 times more servers. It must be some solution where if I want to serve 10 times more tokens, then I get some economies of scale benefit with my say cluster…”
Uberti: Etched sent a dozen engineers to Bangalore to fix delays
“We went out and shipped a dozen of our top engineers across the world to Bangalore for six months. I was there as well. I lived in Bangalore for four and a half months personally”
Uberti: Etched aligned two clock signals to within 50 picoseconds
“We realized there was one and only one way to solve it. As we had to go line up two clock signals on our chip to within 50 picoseconds. That is literally 50 trillionths of a second. And we had to go get these signals aligned to this super small granularity and…”
Uberti: TSMC VP Emailed Wanting to Partner with Etched After One Dinner
“And the following day, I get an email from TSMC saying, Gavin, I want to work with Etched. Find a way to make it happen.”