multimodal models

also referred to as: multimodal model

2 statements across 1 episodes · 2 bullish · 0 bearish · 2 people on the record · first statement Jun 27, 2024 by Karan Goel · across every show →

Everything said about multimodal models, oldest first

Jun 27, 2024 bullish
Disclosure
Goel: Cartesia intends to train an on-device multimodal model
“Yes. But, you know, we have our own sort of set of techniques that we're developing in order to be able to do that effectively. I think that I will maybe leave for another podcast. But I think, yeah, I think that is the intention at the end of the day is build…”
Karan Goel Jun 27, 2024 ▶ 27:54 No Priors Ep. 70 | With Cartesia Co-Founders Karan Goel & Albert Gu
Jun 27, 2024 positive
Insight
Gu: High-quality speech synthesis requires multimodal foundation models
“And so actually to really get, like, perfect even just TTS or, like, speech-to-speech you actually really need to have, like, a model that has, More understanding, like at least of the language, but kind of like, it's not really an isolated component anymore. …”
Albert Gu Jun 27, 2024 ▶ 24:27 No Priors Ep. 70 | With Cartesia Co-Founders Karan Goel & Albert Gu
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.