Neil Movva

Co-Founder and CEO, Sail Research · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderexecutiveengineer@neilmovva ↗sailresearch.com ↗

Neil Movva is the co-founder and CEO of Sail Research, an AI infrastructure startup optimizing compute to lower inference costs. He previously worked on GPU kernels at NVIDIA and Together AI, worked on the Apple Neural Engine, and co-founded Blyss.

67statements → 32claims → 11claims resolved → 82%fully supported → 3.69/5average certainty → 2.48/5average debate potential → ≈4.5/5argument clarity, estimated →

9 supported 2 partly supported 0 contradicted 2 not yet assessed 19 not checkable as stated how the 32 claims stand · each chip opens the sources

10 predictions · 22 assertions · 13 opinions · 19 insights · 3 disclosures · every statement was checked. The predictions and assertions are the 32 claims: statements the public record can support or contradict. 11 are resolved, 2 are not yet assessed, and 19 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Neil argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Movva: Non-NVIDIA chips beat NVIDIA on FLOPS per dollar
“And so there are other chips that definitely rank higher than NVIDIA on flops per dollar. But they may not have as much interconnect”
Neil Movva Aug 25, 2026 ▶ 23:09 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper

Expressed certainty vs assessment result

none yet certainty 1
none yet certainty 2
83% certainty 3
94% certainty 4
100% certainty 5

weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet

How they sound: not measured why? →

We measure speaking style by listening to the audio itself, and a fair number needs at least 2,000 words from one person on tape we have measured. There is too little of Neil Movva on measured tape to publish a rate. This says nothing about how they speak.

Everything Neil Movva said on Invest Like the Best that made the record, most notable first. Filter by type, assessment or year in the ledger →

Opinion
Losing TSMC Wouldn't Be Disastrous Because Intel Is Only 2x Behind
“And my contrarian take is that it wouldn't be that bad. Supply would take a shock for sure, but The best processes that we have in the West, like Intel, not that far behind, at worst, like maybe two X worst performance per watt. And the gap is just far smaller…”
Neil Movva Aug 25, 2026 ▶ 1:17:13 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Disclosure
Movva: Sail Research will buy data centers with only 95% uptime
“You'd have missing zero buyers for a data center that is 95% uptime. I'm that first buyer. I will buy 95% uptime.”
Neil Movva Aug 25, 2026 ▶ 57:26 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Not checkable as stated
Nvidia's BF16 Efficiency Barely Improves Across Hopper, Blackwell, and Rubin
“If you look at, you know, Hopper to Blackwell to Rubin, and you compare like for like, what is the performance per watt of a beefload-sixteen multiply? It hasn't improved all that much.”
Neil Movva Aug 25, 2026 ▶ 1:16:40 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Not checkable as stated
TSMC Performance Per Watt Barely Changes From 5nm to 2nm
“If you look at TSMC five nanometer versus four versus three versus two, the performance per watt on these chips doesn't change like a dramatic amount.”
Neil Movva Aug 25, 2026 ▶ 1:16:56 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Not checkable as stated
Cursor Forced Baseten, Fireworks, and Together to Prioritize Low Latency
“The challenge is, all those companies, you could take your pick, Base-Ten, Fireworks together, they all focus on low latency inference. And they were pulled in that direction by one very important customer Cursor.”
Neil Movva Aug 25, 2026 ▶ 3:42 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Prediction Not checkable as stated
Future AI Agents Will Prioritize Long-Horizon Task Completion Over Speed
“You wanted more persistence, more long horizon tasks, and now it's, to me, very obvious that the future of agentic inference is long horizon tasks. You're gonna run the machine for hours or days at a time. It doesn't matter if it spits out tokens at a hundred …”
Neil Movva Aug 25, 2026 ▶ 4:02 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Prediction Not checkable as stated
AI Workloads Will Eventually Shift to 90% Background and 10% Real-Time
“So long-term, I think, you know, we're going to end this year at maybe fifty-fifty background and real-time workloads, but I see this going to ninety-ten in favor of background.”
Neil Movva Aug 25, 2026 ▶ 6:25 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Opinion
Anthropic API Spend Is Now the Best Metric for Software Security
“When you want secure software, it's really a question of how many dollars did you spend on Anthropix APIs trying to break into your software. That is the best indication for how secure it is, because that's the best tool in the world.”
Neil Movva Aug 25, 2026 ▶ 8:02 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Prediction Not checkable as stated
AI Scientific Research Costs Could Soon Drop to Tens of Dollars
“And we have actually started to bring it within view, a dollar cost for these long horizon tasks that is reasonable. It's not millions, it's thousands, and maybe it could be hundreds or even tens of dollars in the near future, too. Have a definitive answer to …”
Neil Movva Aug 25, 2026 ▶ 10:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Opinion
Movva: NVLink is mandatory for low-latency AI inference
“NVLink is mandatory, I would say, for low latency inference.”
Neil Movva Aug 25, 2026 ▶ 22:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Prediction Not checkable as stated
Movva: Rival chipmakers will not replicate NVIDIA's NVLink quickly
“I think I'm not holding my breath for other companies broadly to figure out NVLink quickly. It's challenging technology to figure out. It's hard to scale. It's hard to productionize.”
Neil Movva Aug 25, 2026 ▶ 22:33 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Supported
Movva: Non-NVIDIA chips beat NVIDIA on FLOPS per dollar
“And so there are other chips that definitely rank higher than NVIDIA on flops per dollar. But they may not have as much interconnect”
Neil Movva Aug 25, 2026 ▶ 23:09 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Prediction Not checkable as stated
Movva: Cerebras and Groq will serve as accelerators alongside traditional GPUs
“Cerebris and Grok and maybe a couple others, you should think of them as accelerators. What they are really good at is being used in conjunction with a more traditional GPU-like device that critically has this off-chip memory built in.”
Neil Movva Aug 25, 2026 ▶ 30:57 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Not checkable as stated
Nvidia Strategically Rations Chips to Prevent Deep-Pocketed Buyers From Gaining Power
“NVIDIA sees the, if they just let the most deep pockets buy all the chips that Maybe hurts them in the long term if that customer ends up accruing a lot of more power. They understand that compute is power today, and so they're quite strategic about how they a…”
Neil Movva Aug 25, 2026 ▶ 45:17 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Insight
Movva: There are no bad AI chips, only bad pricing
“I like to say there's no bad chips, there's only bad pricing. And I will make any chip work at the right price.”
Neil Movva Aug 25, 2026 ▶ 46:15 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Not checkable as stated
Building 100-Megawatt Data Centers in the US Is Now Practically Impossible
“And basically there's no way to build a gigawatt data center in the United States easily anymore. Even a hundred megawatts is, is increasingly hard. It's basically impossible unless you're a very special set of customers. 10 megawatts is probably on the edge o…”
Neil Movva Aug 25, 2026 ▶ 54:05 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Insight
Movva: Uncontrolled P99 latency yields unbeatable economics for background agents
“We tell our customers, look, our average throughput is going to be very competitive, but our P 99, our 99 percentile latency is not going to be controlled. It cannot be. And in return, I'll give you unbeatable economics. And I think that's the right fit for ba…”
Neil Movva Aug 25, 2026 ▶ 58:40 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Insight
Tolerating Data Center Outages Unlocks Cheap Power for Inexpensive AI Chips
“The trick is that it's going to give me better access to power that no one else is going to touch, because it is so annoying to deal with that kind of outage. And if my chips are cheap enough, they're probably not going to be Nvidia racks. And if my chips are …”
Neil Movva Aug 25, 2026 ▶ 59:54 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Assertion Open · timeframe Dec 2026
Movva: NVIDIA is producing 5 million Blackwell chips this year
“NVIDIA's pumping out five million Blackwell chips this year.”
Neil Movva Aug 25, 2026 ▶ 1:04:57 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Opinion
Global GPU Utilization Is Actually Far Worse Than xAI's Clusters
“We all make fun of XAI for having, you know, some challenges with total flop utilization on its clusters, but the reality for the rest of the world is it's far worse. A ton of GPUs just sit in warehouses or sit In private pools allocated to a specific customer…”
Neil Movva Aug 25, 2026 ▶ 1:05:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Opinion
Movva: Frontier AI labs pay huge premium for 3-6 month lead
“In a line, I would say the labs pay an immense premium to be three to six months ahead of everything else.”
Neil Movva Aug 25, 2026 ▶ 1:10:22 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Prediction Not checkable as stated
Movva doubts commercial premium for short-term AI lead will last
“I don't know that the premium for being three to six months ahead is going to last that long. I mean, if you look at, like, enterprise deployments they don't move at three to six months speed.”
Neil Movva Aug 25, 2026 ▶ 1:11:59 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Disclosure
Movva: Sail's Custom Chip Strategy Sidesteps HBM by Offloading to Flash
“When I talk about building custom chips, and they ask me, oh, so what's different? Basically, it's about sidestepping the HBM shortage and focusing on More extreme offload to other forms of memory, such as flash.”
Neil Movva Aug 25, 2026 ▶ 1:14:23 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper
Insight
Movva: Compute startups should primarily attack Nvidia's reliance on HBM
“You want to pick something and say, I think they've underpriced the impact of how short we're going to be on HBM. We're going to push really hard in this other direction instead, which, you know, as an aside, I do think is probably the thing to attack most.”
Neil Movva Aug 25, 2026 ▶ 1:19:42 Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper

Show 24statements(43 left)

Appearances (1)

EpisodeDateSpeaking time
Ex-NVIDIA Engineer: Why AI Is About to Get 1000x Cheaper Aug 25, 2026 1h 0m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 60 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.