ElevenLabs co-founder and CEO Mati Staniszewski explains their strategy of building proprietary foundational speech models independently of OpenAI.
Opinion
Staniszewski: No AI bubble brewing in Silicon Valley because revenue is real
“I don't think that the bubble is brewing. I think there's, like, what I think is a very different part, as I compare it to other cycles, I think there's a clear value, clear revenue numbers for a lot of the companies.”
Prediction Not checkable as stated
Staniszewski: Headphones will likely be the primary voice AI form factor
“I do think it most likely will be the most natural form factor given it's already one of the key devices.”
Opinion
Staniszewski: Skeptical of valuations for GPU and inference providers built on NVIDIA
“It's all the, like, GPU providers that are built on top of NVIDIA that, like, help you get Effectively inference capacity. Like I'm a little bit skeptic on some evaluations there”
Prediction Not checkable as stated
Staniszewski: Only a few global companies will pre-train foundational AI models
“Like let's take building foundational AI models. I think the kind of, the base will be pre-trained by few companies in the world, just given the resources required. They'll be open source of course offerings in that bucket. But then Probably, as I think about …”
Insight
Staniszewski: Deep domain integrations create defensibility for AI startups
“If you know domain extremely well, then The AI model, this, you know, gives you a good starting point, but then integrations with the system, the legacy systems of how you deliver that, telephony how you bring the knowledge, that will take so much time. Then l…”
Prediction Not checkable as stated
Staniszewski: Users will carry pendants or wristbands for AI context capture
“I do think there will be some version of a pendant or a wristband that you will carry to just capture more.”