The Ledger, every show

Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.

shows every show 44 of 44
every show
clear all ✕
MAD Opinion
Fu: Current LLMs meet the definition of AGI from 5-10 years ago
“By almost any definition anyone could have written down, let's say five years ago or 10 years ago, certainly when, you know, Tim, you and I started our PhD. We basically have the vision of AGI that, that we had back then. We have things that can write code. Th…”
Dan Fu Jan 22, 2026 ▶ 4:14 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Prediction Not checkable as stated
Fu: Next-generation models currently in training will achieve AGI
“You know, we maybe already have AGI or like some form of AGI. And if not, then certainly the next generation of models, the models that today are training already. If they're at all better than what we have today, then we're, we we've already hit something tha…”
Dan Fu Jan 22, 2026 ▶ 5:17 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
LATENT SPACE Assertion Not checkable as stated
Fu: Embedding model quality barely matters for final RAG performance
“We had this experience over and over again where you could have any, an embedding model of any quality, so you could have a really, really bad embedding model, or you could have a really, really good one by, and by any measure of good, and for the final RAG ap…”
Dan Fu Dec 24, 2024 ▶ 33:00 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
MAD Assertion Not checkable as stated
Fu: AI coding tools enable expert programmers to move 10x faster
“But if you give an expert programmer This set of tools, they can go 10, 10 times faster than they were able to go before.”
Dan Fu Jan 22, 2026 ▶ 34:41 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Fu: Modern GPU compute primitives should be matrices, not floats
“We basically built a whole library just around this basic idea that all your basic compute primitives should not be a float, but it should be a matrix and everything should just be matrix compute.”
Dan Fu Dec 24, 2024 ▶ 30:29 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
LATENT SPACE Prediction Open · timeframe Dec 2029
Fu: Real-time long-context video generation cannot use quadratic attention
“You're certainly not going to do a giant quadratic attention computation to try to run that.”
Dan Fu Dec 24, 2024 ▶ 31:33 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
MAD Assertion Partly supported
Dan Fu: DeepSeek-V3 was trained on ~2,000 H800s with 20% MFU
“If you look at the deep seek model, for instance, this is one of the best open source models we have out there today. It was trained at the end of 2024. On last generation, kind of nerfed GPUs, H 800 instead of H 100, the 800 is nerfed by all sorts of ways fro…”
Dan Fu Jan 22, 2026 ▶ 17:26 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Insight
Dan Fu: Deployed AI models lag cluster infrastructure by 1–2 years
“The models that we see today that we can play with today have been pre-trained on clusters that were built out a year or two ago. Because, you know, you need enough time to get the cluster running. You need enough time to do the large pre-training run. And the…”
Dan Fu Jan 22, 2026 ▶ 21:26 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Assertion Not checkable as stated
Fu: Hardware utilization during AI inference is under 5%
“At inference time, when the, when you have the model, when it's already been trained, already been post-trained, the hardware utilization is like less than five percent.”
Dan Fu Jan 22, 2026 ▶ 55:13 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Assertion Not checkable as stated
Dan Fu: Chinese AI labs take more architectural risks
“I think you see a lot more risk taking out of the Chinese labs where you're trying to differentiate the next model of your next open source model.”
Dan Fu Jan 22, 2026 ▶ 1:03:19 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
Fu: Efficient AI Architectures Are Dead on Arrival Without Hardware Co-Design
“Even if your model is theoretically more efficient, if somebody goes and runs it and it's two times slower one of the things that, that we've learned is that if you're in that situation, it's just going to be dead on arrival. So you want to be designing your a…”
Dan Fu Dec 24, 2024 ▶ 15:51 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
Dan Fu: Nobody is actually submitting 2M token prompts into LLMs
“Nobody is actually putting in a two million context prompt into these models.”
Dan Fu Dec 24, 2024 ▶ 37:33 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
MAD Prediction Not checkable as stated
Fu predicts increasing hardware diversity, particularly for AI model inference
“I'm sure NVIDIA will still do great and still grow beyond their five trillion dollar company or whatever it is at the time of recording. But I think you're going to see a lot more diversity especially around, I think inference of the model.”
Dan Fu Jan 22, 2026 ▶ 31:39 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Insight
Dan Fu: Junior engineers using AI agents communicate better and level up faster
“When they are really gung ho about understanding and being able to use the AI agents, there's, they're able to communicate so much better than in the olden days. They're able to level up their level of understanding a lot faster.”
Dan Fu Jan 22, 2026 ▶ 49:34 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Insight
Dan Fu: Proper speculative decoding yields 2x to 3x model speedups
“So if you do the speculative decoding right, you can get, again, two X, three X speed ups over, over, you know, just running a vanilla model.”
Dan Fu Jan 22, 2026 ▶ 57:34 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
MAD Assertion Not checkable as stated
Dan Fu: Some top audio models use state space architectures
“So some of the best audio models in the world are at least partially based on state space models.”
Dan Fu Jan 22, 2026 ▶ 1:02:29 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
LATENT SPACE Assertion Not checkable as stated
Fu: AI21's Jamba Is the State of the Art Non-Transformer Model
“AI-II trained this hybrid MOE called Jamba that, that, that seems, that is currently the state of the art for these non-transformer architectures.”
Dan Fu Dec 24, 2024 ▶ 16:50 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
Fu: Changing one PyTorch line requires a week of CUDA development
“If we decided to change one thing in PyTorch, like one line of PyTorch code is like a week of CUDA code at least.”
Dan Fu Dec 24, 2024 ▶ 29:38 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
MAD Assertion Supported
Poolside and Reflection are building clusters with massive B200 GPU deployments
“They're companies like Poolside. They're building out tens of thousands of B-two hundred, GB-two hundred chips. You know, there's other folks like Reflection who are who are building out. Tens of thousands of B 200 chips.”
Dan Fu Jan 22, 2026 ▶ 19:13 The End of GPU Scaling? Compute & The Agent Era — Tim Dettmers (Ai2) & Dan Fu (Together AI)
LATENT SPACE Assertion Supported
Fu: Stanford and Arc Institute's DNA SSM Made the Cover of Science
“One of those gated, SSM gated states-based models ended up on the cover of science because a great group of folks went and trained some DNA models. So that's Michael Polley, Eric Yuen from Stanford and the Arc Institute.”
Dan Fu Dec 24, 2024 ▶ 17:31 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.