Jeff Schmidt is co-founder of Nous Research. He explains how distributed AI training models could alter silicon design trade-offs between memory and compute density.
“What might happen sooner would be a redesign of the types of chips that NVIDIA or someone would make. Okay, under this model, we can dedicate more VRAM versus, there's like this question of how much VRAM versus how much processing power is on a die, and that, that could like change sort of the dynamics of what that optimal looks like”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Jeff Schmidt
Insight
Schmidt: AI industry relies on outdated 1990s architecture assumptions ripe for disruption
“It turns out that most assumptions in the AI space right now are a product of that's just how things had been done when there was not nearly as much energy and attention to it. So someone made an assumption maybe in the early nineties that everyone just kind o…”
Jeff SchmidtOct 1, 2024▶ 15:48The Quest for Community-Trained Open Source AI Models
Insight
Schmidt: Relying solely on fine-tuning is an existential threat to open-source AI
“Because that for us, if we're just fine tuning models. Right. That's like an existential threat, right? That's like an actual existential threat because the closed providers will continue to get better and we would be like dead in the water in a lot of sense.”
Jeff SchmidtOct 1, 2024▶ 20:04The Quest for Community-Trained Open Source AI Models
Insight
Schmidt: Sharing only key signals in distributed training yields equivalent model learning
“We know that like what we, what needs to be communicated between these things, the two, the different nodes are just these few key pieces of information. And that is necessary. That is a necessary condition or rather a sufficient condition To get the equivalen…”
Jeff SchmidtOct 1, 2024▶ 32:42The Quest for Community-Trained Open Source AI Models
AssertionPartly supported
Schmidt: Compute chips in NVIDIA's RTX 4090 and H100 are almost identical
“I think people don't actually realize that like a forty-ninety and like an H 100 are in a lot of ways the same card. For the non-gamers in the room, explain the forty-ninety. The chip that's inside of them is almost identical. The chip, the actual compute chip…”
Jeff SchmidtOct 1, 2024▶ 44:54The Quest for Community-Trained Open Source AI Models
PredictionNot checkable as stated
Schmidt: Consumer gaming GPUs will become the sweet spot for distributed training
“Because you're able to distribute it so wide, I think the gaming GPU angle is really going to be like the sweet spot. If, as long as there's continued to be sort of like higher end gaming GPUs, and those are on comparison with the high end training GPUs, even …”
Jeff SchmidtOct 1, 2024▶ 45:39The Quest for Community-Trained Open Source AI Models
PredictionDidn’t hold up
Schmidt: Decentralized training of 400B parameter AI models is solvable by 2025
“I think it still is, it would still be, you know, like a next year sort of environment thing that we would have to do. There are some scaling problems, or not scaling problems, but technical things about how you shard the model, because at that point you get t…”
Jeff SchmidtOct 1, 2024▶ 54:41The Quest for Community-Trained Open Source AI Models
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 1,000 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.