Qwen, every mention
2 scenes · ← back to Qwen
tap a year for its mentions
every year anyone Ankit Gupta 1
Verbatim, from the transcripts: the passages where Qwen comes up
Multi-GPU Kernels, Intelligence per Watt, Heterogeneous Inference, and More | YC Paper Club · Y Combinator
- ▶ 25:43 unnamed speaker We did about twenty-plus different state-of-the-art local models, um, across GEMMA, GBD OSS, QUAN, IBM Granite, all in the range of one to two hundred billion parameters, some MOE, some dense.
Anthropic Head of Pretraining on Scaling Laws, Compute, and the Future of AI · Y Combinator
- ▶ 34:09 Ankit Gupta Quen smaller reasoning models distilled off of the larger Quen models, for example, and similar with Deep Seek, for example.