Quen
product on 6 shows · 7 statements across 7 episodes
More or Less
BG2 Pod
Latent Space
Lenny's Podcast
the Startup Ideas Podcast
Big Technology
7 statements about Quen, every show
Isenberg: Alibaba's Qwen is very strong for coding, multilingual, and agentic tasks
“Quen is Alibaba's model family, and has become very strong, especially around coding, multilingual work, long contacts, and agentic tasks.”
AI models differ by cognitive disposition rather than raw intelligence
“It's not that one is ahead of another one is more intelligent. It's rather One is, sort of, has a mind that's shaped in one direction, perhaps creativity and openness for Quen, and others that are shaped in other directions, like, you know, neuroticism and pre…”
Sam Lessin predicts AI spend will shift to cheap open-source models
“And so the vast majority of tokens and spend in the market, the good market, the good morning America audience is going to be going for cheap Quen effectively over time. And it's going to be because they care about the costs, right? Like that's going to be the…”
Ramaswamy: Alibaba's new Qwen model is shockingly close to Anthropic's Claude Sonnet
“A new Quen model came out yesterday that is shockingly close to the best Sonnet model that there is from Anthropic.”
Morcos: Qwen is much easier to align than Llama due to pre-training
“It's much easier to RL Quen than it is to do Lama. Likely that has to do with the fact that Quen put a lot of synthetic reasoning traces into their training data.”
Madra: Alibaba's 30-billion parameter Qwen model matches GPT-4o performance
“There was a release today of a Quen You know, thirty billion parameter model, which is performing as good as GPT-IV-O.”
Soldani: Llama and Qwen models fail OSI open source AI definition
“Under this definition, for example, Lama or some of the Quen models are not open source because the license says you can, you can't use this model for this, or it says if you use this model, you have to name the output this way or derivative needs to be named …”