text-to-video
1 statements across 1 episodes · 0 bullish · 0 bearish · 1 people on the record · first statement Jun 1, 2026 by Ethan He · across every show →
Everything said about text-to-video, oldest first
Jun 1, 2026 neutral
He: Audio-video models require precise time alignment unlike text-to-video
“So one important thing is like the alignment. So the model has to know like the video and audio, the it has to have a time based alignment, like at which time step the video and the audio token correspond to each other. We actually don't have these kind of ali…”