native multimodal models

1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement Feb 12, 2026 by Sherwin Wu · across every show →

Everything said about native multimodal models, oldest first

Feb 12, 2026 bullish
Prediction Not checkable as stated
Wu: Native speech-to-speech multimodal models will improve dramatically in 6-12 months
“I think they're gonna get a lot better at audio over the next six to 12 months, especially the likes, you know, the native multimodal models, the speech-to-speech ones.”
Sherwin Wu Feb 12, 2026 ▶ 52:30 OpenAI’s head of platform engineering on the next 12-24 months of AI | Sherwin Wu
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.