Disclosure
Lample: Mistral plans a phased approach to full-duplex audio models
“Ultimately what we want to do is to be this, Full duplex model, but we are not going to start this, start there directly. I think it's some approach that people are doing, but. Just to confirm, full duplex means it can speak while I'm speaking, or? Okay. Yeah,…”
Insight
Lample: Specialized AI models are more cost-effective than monolithic models
“That's why we can actually use models audio, but also like OCRs that are like really, really good at that, and that will be much more cost effective than a general model. That will contain a lot of capabilities you don't really need.”
Insight
Lample: 1B to 3B parameter models are optimal for speech transcription
“For instance, for audio here, if you want to do transcription, I think it makes no sense to use a model as this large. If you just want to transcribe tech, it would be very inefficient. Like if you want to do audio, you probably just want to do the one B or a …”
Assertion Not checkable as stated
Lample: Mistral is far from reaching pre-training saturation
“We are still working a lot on the pre-training side. We are very, very far from any sort of situation on the pre-training.”
Insight
Lample: Long-horizon RL trajectories require new algorithms beyond GRPO
“GRPO, for instance, it doesn't really work with any bit of policy, which was okay initially, because you are solving math problems that can be solved in like a few thousand tokens, so the model can actually generate them pretty quickly, so when you do your upd…”
Assertion Not checkable as stated
Lample: Voxtral TTS matches leading models at a fraction of cost
“So we support nine languages and this is a pretty small model a three-dimensional model, so very fast, and also state-ups, yeah, very equal. Performed at the same level of the best model, but it's Much more efficient in terms of cost, and also much, in terms o…”
What-if
Lample: Post-training breakthroughs like DPO were impossible without open LLaMA
“And if you look at many of the techniques that were developed after, for instance, Temma was open source, like all these post-training approaches like even DPOD, like performance optimization, all of this were done by people that had access to this model, and …”
Opinion
Lample: Closed models prevent enterprises from leveraging proprietary domain data
“When customers use this off-the-shelf closed model, what's very sad is that they are not leveraging, you know, these data that they have been collecting for four years, or sometimes for decades. So much data, sometimes it's trillions of tokens or data in a ver…”
Disclosure
Lample: Mistral AI is releasing Voxtral TTS, its first speech generation model
“So we are releasing Vokstral TTS. So it's our first audio model that generates speech.”
Assertion Not checkable as stated
Lample: Mistral can build fine-tuned models 10x cheaper than closed models
“On here we can sometimes build something 10 X cheaper by just fine tuning a model, and it would be better on prem, on their own server, and also much cheaper as well”
Disclosure
Lample: Mistral develops specialized single-capability models before merging them
“The way we kind of do things internally, that we have like one team, focus on one capability, build one model, and then when it's mature enough, we decide to merge this into the Mixture. So, but hey, here's, it was the first time we basically merged all of thi…”
Insight
Lample: Frontier models underperform in legal and CAD due to missing benchmarks
“Things around, like, legal, finance, computer-aided design, all of these things that it's, these models out of the box are never too good at that, because people really don't prioritize this, there is no, like, too many benchmark on that but it's not hard to m…”
Prediction Not checkable as stated
Lample: AI coding agents will drastically expand the software verification industry
“Now with coding agents that are there, It's going to be very different. We're going to see much more of this. So I think, yes, industry there is going to be much larger in the future that we have these models.”
Disclosure
Lample: Mistral initially underestimated enterprise model deployment complexity
“What we underestimated initially is the complexity of deploying this model and connecting them to everything to be sure it has access to the company knowledge.”