The Ledger, every show

Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.

shows every show 44 of 44
every show
clear all ✕
Losing TSMC Wouldn't Be Disastrous Because Intel Is Only 2x Behind
“And my contrarian take is that it wouldn't be that bad. Supply would take a shock for sure, but The best processes that we have in the West, like Intel, not that far behind, at worst, like maybe two X worst performance per watt. And the gap is just far smaller…”
Neil Movva Aug 25, 2026 ▶ 1:17:13 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Sail Research will buy data centers with only 95% uptime
“You'd have missing zero buyers for a data center that is 95% uptime. I'm that first buyer. I will buy 95% uptime.”
Neil Movva Aug 25, 2026 ▶ 57:26 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Not checkable as stated
Nvidia's BF16 Efficiency Barely Improves Across Hopper, Blackwell, and Rubin
“If you look at, you know, Hopper to Blackwell to Rubin, and you compare like for like, what is the performance per watt of a beefload-sixteen multiply? It hasn't improved all that much.”
Neil Movva Aug 25, 2026 ▶ 1:16:40 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Not checkable as stated
TSMC Performance Per Watt Barely Changes From 5nm to 2nm
“If you look at TSMC five nanometer versus four versus three versus two, the performance per watt on these chips doesn't change like a dramatic amount.”
Neil Movva Aug 25, 2026 ▶ 1:16:56 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Not checkable as stated
Cursor Forced Baseten, Fireworks, and Together to Prioritize Low Latency
“The challenge is, all those companies, you could take your pick, Base-Ten, Fireworks together, they all focus on low latency inference. And they were pulled in that direction by one very important customer Cursor.”
Neil Movva Aug 25, 2026 ▶ 3:42 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
Future AI Agents Will Prioritize Long-Horizon Task Completion Over Speed
“You wanted more persistence, more long horizon tasks, and now it's, to me, very obvious that the future of agentic inference is long horizon tasks. You're gonna run the machine for hours or days at a time. It doesn't matter if it spits out tokens at a hundred …”
Neil Movva Aug 25, 2026 ▶ 4:02 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
AI Workloads Will Eventually Shift to 90% Background and 10% Real-Time
“So long-term, I think, you know, we're going to end this year at maybe fifty-fifty background and real-time workloads, but I see this going to ninety-ten in favor of background.”
Neil Movva Aug 25, 2026 ▶ 6:25 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Anthropic API Spend Is Now the Best Metric for Software Security
“When you want secure software, it's really a question of how many dollars did you spend on Anthropix APIs trying to break into your software. That is the best indication for how secure it is, because that's the best tool in the world.”
Neil Movva Aug 25, 2026 ▶ 8:02 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
AI Scientific Research Costs Could Soon Drop to Tens of Dollars
“And we have actually started to bring it within view, a dollar cost for these long horizon tasks that is reasonable. It's not millions, it's thousands, and maybe it could be hundreds or even tens of dollars in the near future, too. Have a definitive answer to …”
Neil Movva Aug 25, 2026 ▶ 10:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: NVLink is mandatory for low-latency AI inference
“NVLink is mandatory, I would say, for low latency inference.”
Neil Movva Aug 25, 2026 ▶ 22:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
Movva: Rival chipmakers will not replicate NVIDIA's NVLink quickly
“I think I'm not holding my breath for other companies broadly to figure out NVLink quickly. It's challenging technology to figure out. It's hard to scale. It's hard to productionize.”
Neil Movva Aug 25, 2026 ▶ 22:33 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Supported
Movva: Non-NVIDIA chips beat NVIDIA on FLOPS per dollar
“And so there are other chips that definitely rank higher than NVIDIA on flops per dollar. But they may not have as much interconnect”
Neil Movva Aug 25, 2026 ▶ 23:09 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
Movva: Cerebras and Groq will serve as accelerators alongside traditional GPUs
“Cerebris and Grok and maybe a couple others, you should think of them as accelerators. What they are really good at is being used in conjunction with a more traditional GPU-like device that critically has this off-chip memory built in.”
Neil Movva Aug 25, 2026 ▶ 30:57 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Not checkable as stated
Nvidia Strategically Rations Chips to Prevent Deep-Pocketed Buyers From Gaining Power
“NVIDIA sees the, if they just let the most deep pockets buy all the chips that Maybe hurts them in the long term if that customer ends up accruing a lot of more power. They understand that compute is power today, and so they're quite strategic about how they a…”
Neil Movva Aug 25, 2026 ▶ 45:17 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: There are no bad AI chips, only bad pricing
“I like to say there's no bad chips, there's only bad pricing. And I will make any chip work at the right price.”
Neil Movva Aug 25, 2026 ▶ 46:15 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Not checkable as stated
Building 100-Megawatt Data Centers in the US Is Now Practically Impossible
“And basically there's no way to build a gigawatt data center in the United States easily anymore. Even a hundred megawatts is, is increasingly hard. It's basically impossible unless you're a very special set of customers. 10 megawatts is probably on the edge o…”
Neil Movva Aug 25, 2026 ▶ 54:05 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Uncontrolled P99 latency yields unbeatable economics for background agents
“We tell our customers, look, our average throughput is going to be very competitive, but our P 99, our 99 percentile latency is not going to be controlled. It cannot be. And in return, I'll give you unbeatable economics. And I think that's the right fit for ba…”
Neil Movva Aug 25, 2026 ▶ 58:40 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Tolerating Data Center Outages Unlocks Cheap Power for Inexpensive AI Chips
“The trick is that it's going to give me better access to power that no one else is going to touch, because it is so annoying to deal with that kind of outage. And if my chips are cheap enough, they're probably not going to be Nvidia racks. And if my chips are …”
Neil Movva Aug 25, 2026 ▶ 59:54 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Open · timeframe Dec 2026
Movva: NVIDIA is producing 5 million Blackwell chips this year
“NVIDIA's pumping out five million Blackwell chips this year.”
Neil Movva Aug 25, 2026 ▶ 1:04:57 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Global GPU Utilization Is Actually Far Worse Than xAI's Clusters
“We all make fun of XAI for having, you know, some challenges with total flop utilization on its clusters, but the reality for the rest of the world is it's far worse. A ton of GPUs just sit in warehouses or sit In private pools allocated to a specific customer…”
Neil Movva Aug 25, 2026 ▶ 1:05:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Frontier AI labs pay huge premium for 3-6 month lead
“In a line, I would say the labs pay an immense premium to be three to six months ahead of everything else.”
Neil Movva Aug 25, 2026 ▶ 1:10:22 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
Movva doubts commercial premium for short-term AI lead will last
“I don't know that the premium for being three to six months ahead is going to last that long. I mean, if you look at, like, enterprise deployments they don't move at three to six months speed.”
Neil Movva Aug 25, 2026 ▶ 1:11:59 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Sail's Custom Chip Strategy Sidesteps HBM by Offloading to Flash
“When I talk about building custom chips, and they ask me, oh, so what's different? Basically, it's about sidestepping the HBM shortage and focusing on More extreme offload to other forms of memory, such as flash.”
Neil Movva Aug 25, 2026 ▶ 1:14:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Compute startups should primarily attack Nvidia's reliance on HBM
“You want to pick something and say, I think they've underpriced the impact of how short we're going to be on HBM. We're going to push really hard in this other direction instead, which, you know, as an aside, I do think is probably the thing to attack most.”
Neil Movva Aug 25, 2026 ▶ 1:19:42 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Open · timeframe Aug 2029
Memory Shortages Will Force Apple to Cut iPhone Memory and Raise Prices
“I think they're gonna make everything else more expensive. I think that iPhones will cut their memory, iPhones are gonna go up in price, and we're just gonna deal with it.”
Neil Movva Aug 25, 2026 ▶ 1:20:19 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Opus 4.5 Was the First Model Suitable for Long-Horizon Tasks
“Opus Four Five was the first agent that was at all suitable for longer horizon tasks.”
Neil Movva Aug 25, 2026 ▶ 5:29 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Smaller AI Models Find Software Bugs That Large Models Miss
“It's not the case that Fable finds a superset of all bugs in software. You would find some bugs with a very small model that you don't find with a large model.”
Neil Movva Aug 25, 2026 ▶ 8:16 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: AI can tackle any verifiable problem including software and math
“The long lens view to take on this is that we have a form of intelligence that can tackle any verifiable problem. Any verifiable problem means most software. It means a lot of formal, like math proofs and similar. And it could also mean scientific discovery”
Neil Movva Aug 25, 2026 ▶ 9:54 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Neil Movva: AI may fundamentally never solve human taste
“We have not solved human taste yet, and I don't know that it fundamentally can be.”
Neil Movva Aug 25, 2026 ▶ 12:34 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
The 'Original Sin' of Transformers Is Pairing Memory-Bound and Compute-Bound Layers
“And I would say the original sin of Transformers is that you've taken this extremely fundamentally memory bound layer and juxtaposed it right next to a compute bound layer. It is very difficult to have a single chip that is good at both compute operations and …”
Neil Movva Aug 25, 2026 ▶ 31:43 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Not checkable as stated
Movva: Models have trained on the whole internet and exhausted human web data
“Models have seen the entire internet many times over at this point, and there is not a whole lot more to be done on human data from the internet.”
Neil Movva Aug 25, 2026 ▶ 36:35 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Unconditioned human preference data is no longer useful for AI training
“The signal you get from random human preference, or I guess unconditioned human preference, is not actually worth anything anymore. You want expert human preference at this point.”
Neil Movva Aug 25, 2026 ▶ 37:10 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: The best route to AGI is stacking specialized intelligences
“Basically, the idea is that if you want artificial general intelligence, the best way to get there is to just keep stacking specialized intelligences until you have no more gaps to fill.”
Neil Movva Aug 25, 2026 ▶ 38:18 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
Movva: AI models will excel at kernel engineering within six months
“I'm sure in six months time we'll have much better models on kernel engineering, and I'm sure the labs would tell you that they already do a lot of their kernel engineering in a fully automated way.”
Neil Movva Aug 25, 2026 ▶ 41:26 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Supported
Nvidia's Peak FLOPs Are Impossible to Hit Due to Power Throttling
“That operation runs at, you know, 70, 80% of peak utilization, and it's limited not by software, but by power. The way NVIDIA quotes peak flops is a little optimistic. You never hit that because of power throttling”
Neil Movva Aug 25, 2026 ▶ 43:09 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: AMD chips are great, but developers cannot program them well
“AMD, I think, great chips overall. The challenge is that people don't understand how to program them very well.”
Neil Movva Aug 25, 2026 ▶ 46:24 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Supported
Movva: Meta and OpenAI purchases have consumed AMD's chip supply
“I think publicly, Meta and OpenAI have bought a ton of AMD chips, and so, we're increasingly seeing that all the AMD supply is also being allocated.”
Neil Movva Aug 25, 2026 ▶ 47:06 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: AI token consumption is non-speculative because tokens are consumed immediately
“And I think what's interesting about token consumption or AI consumption broadly is that it's no longer speculative. People buy tokens because they're immediately valuable to them. You don't hoard tokens, you use them immediately.”
Neil Movva Aug 25, 2026 ▶ 50:34 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Sail Scavenges Chips and Power to Avoid Bidding Against Anthropic
“Well, first we scavenge chips, and then we scavenge power for those chips. The idea is in both cases, I do not want to be bidding against Anthropic or of an AI for a complete capacity. I'm not going to win against them, and I don't want to. I want to be more c…”
Neil Movva Aug 25, 2026 ▶ 1:00:16 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: KV cache memory usage is inefficient by 1-2 orders of magnitude
“Specifically, the KV cache is quite uncompressed right now. I think if you look at the entropy in a KV cache, it's nowhere near, it's not earning its keep. Like, we're storing many kilobytes of data in the KV cache per token. And that's probably off by an orde…”
Neil Movva Aug 25, 2026 ▶ 1:04:16 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Screening performance engineers for CUDA experience is a red herring
“I don't look for lots of AI experience. I don't look for, you know, CUDA experience at all. That's actually a huge red herring. I mean, CUDA as a concept or GPS as a concept have evolved so much in the last five years. There's no point asking for 10 years of e…”
Neil Movva Aug 25, 2026 ▶ 1:09:54 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Nvidia avoids selling tokens to prevent competing with its customers
“NVIDIA is really smart about this. They don't compete with their customers. NVIDIA takes a long view on everything. Why don't they even start with the NeoCloud? Why don't they just sell computer out the back door? Well, NVIDIA is really good. Jensen is really …”
Neil Movva Aug 25, 2026 ▶ 1:20:30 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Nvidia creates buyer competition to neutralize defection to AMD
“Like, he wants to create a diverse community of NeoClouds and inference providers who are all jockeying to create demand for NVIDIA, such that if any one of them Decides to, I don't know, vertically integrate or go with AMD or any other option. He's got three …”
Neil Movva Aug 25, 2026 ▶ 1:20:51 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Reducing Costs Tenfold Creates a New Product Category
“We think that whenever you make something 10 times cheaper, it's a new product category, and we aspire to do that for tokens.”
Neil Movva Aug 25, 2026 ▶ 1:42 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: Tokens are not the final unit of work or intelligence
“I don't think tokens are the final unit of work or intelligence, but they are what we use today, and so it's very straightforward.”
Neil Movva Aug 25, 2026 ▶ 2:07 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Prediction Not checkable as stated
Movva: AI Agents Will Soon Self-Administer Their Own Token Budgets
“And increasingly, I think we will have agents do some unit of work, take as many shots on goal as they can, and however many tokens they use to get there is going to be kind of a dependent variable depending on the task. So you think about, like, agents that s…”
Neil Movva Aug 25, 2026 ▶ 2:30 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
INVEST LIKE THE BEST Assertion Supported
Movva: AI Agents Can Now Run Autonomously for an Hour
“Agents are capable of running for an hour at a time. I wouldn't say it's days, but definitely an hour is quite suitable today.”
Neil Movva Aug 25, 2026 ▶ 5:39 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Movva: NVIDIA Has Incredible Retention and the Best Silicon Engineers
“A lot of people I worked with in NVIDIA in 2015, 20 16 are still there today. That company has incredible retention, and these are the best engineers, frankly, on the Silicon side, at least, I've worked with them my whole career.”
Neil Movva Aug 25, 2026 ▶ 17:08 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.