why aren't all 7 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Reddy: Voxtral speech model is much stronger than Whisper
“And I think a big people, I think there's a big rich ecosystem of people finding whisper and people want the same thing with Voxer. It's much stronger than whisper.”
Assertion Not checkable as stated
Hsu: Whisper outperformed human listeners on accented Korean English clips
“There were four of us in the room, we all closed our eyes, and none of us had any idea, and the model got it right. So, I mean, superhuman.”
Assertion Supported
Ben-Smith: Multimodal LLM audio transcription remains vastly costlier than self-hosted pipelines
“The big difference right now is still, like, the cost difference of doing speaker diarization this way, or doing transcription this way, is a huge difference to the pipeline that we've built up.”
Assertion Not checkable as stated
Whisper v2 outperforms v3 at certain tasks, delaying its API rollout
“And so whisper V two is better at some things than whisper V three. And so it didn't seem that worthwhile to ship whisper V three compared to like the other things in our priorities. I think we still will at some point, but yeah, it's just, you know, there's a…”
Disclosure
Hsu: Speak runs custom ASR for core loops alongside Whisper for tutoring
“There's many other sort of product surfaces within the app today that are more LM powered, where it's more open-ended, real tutoring, where we actually give you feedback on what you said in the semantics and so on. So that stuff is more like whisper powered, m…”
Disclosure
Sutin: Local model setup friction killed Owl AI's open-source developer adoption.
“I learned, like, we did not make the developer experience very good. It was very complicated like, because we were using, like, local whisper, local models, and, like, getting it to work on CUDA, Mac, Windows. We didn't do a good job, so it was very difficult …”
Disclosure
Duffy: Every is releasing Monologue, a local Whisper-style voice tool
“Naveen's actually releasing something called monologue. You know, it's kind of like, I think a better whisper flow for the things that we do can run locally.”