text-to-speech

also referred to as: tts

2 statements across 1 episodes · 1 bullish · 1 bearish · 2 people on the record · first statement Jun 27, 2024 by Albert Gu · across every show →

Everything said about text-to-speech, oldest first

Jun 27, 2024 positive
Insight
Gu: High-quality speech synthesis requires multimodal foundation models
“And so actually to really get, like, perfect even just TTS or, like, speech-to-speech you actually really need to have, like, a model that has, More understanding, like at least of the language, but kind of like, it's not really an isolated component anymore. …”
Albert Gu Jun 27, 2024 ▶ 24:27 No Priors Ep. 70 | With Cartesia Co-Founders Karan Goel & Albert Gu
Jun 27, 2024 negative
Opinion
Goel: Current AI speech models fail to capture profession-specific vocal nuances
“And that's sort of the nuance that I don't think any models really capture well, which is like, you know, if you're a nurse, you need to talk in a different way than if you're a lawyer, or if you're a judge, or if you're a venture capitalist, you know, very di…”
Karan Goel Jun 27, 2024 ▶ 23:42 No Priors Ep. 70 | With Cartesia Co-Founders Karan Goel & Albert Gu
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.