Atallah: GLM 5.2 Was a Major Step for Open-Weight Models
“GLM 5.2 was a really big, big step for open weight models. Kimmy was kind of like moonshot getting up to that step. That's a little bit how I see it.”
Modifying base LLM weights for vision degrades original text performance
“You don't want to mess with the model weights because you run a chance of making the model dumber at something else for the purpose of giving it vision.”
Baseten's vision-retrofitted GLM-5.2 scored 56% on MMLU Pro without text degradation
“It's not, you know, it got to a 56% on MMLU Pro, I think, so not, not quite Frontier, but if you're running this model, you haven't suffered any loss on your GLM-Five-II quality.”
Unoptimized GLM-5.2 delivers a baseline 30 to 40 tokens per second
“So let's say you have, as a reasonable baseline, 30 or 40 tokens per second. You can achieve 10 X that. So like on GLM 5.2 if you want to get unquantized perhaps on hoppers even and you're just using an off the shelf inference engine with no particular optimiz…”
GLM-5.2 autonomously wrote and guided production GPU kernels for Baseten's inference engine
“Some of the GPU kernels that were on GLM-Five-two within our inference engine is written by GLM-Five-two. And the trace and the kernels were guided by GLM-Five-two as the driver.”
Calacanis: Enterprise customer moved over $100M from frontier AI to GLM 5.2
“He said he has got a customer who just moved like nine figures off of the frontier labs to put it on GLM five, two.”
Hugging Face defended against AI breach using Chinese open model GLM-5.2
“So hugging face had to turn to open model, specifically GLM 5.2, which is deeply ironic, a Chinese open weight model that they run on their own infrastructure.”
Stamos: Hugging Face Used Chinese AI After US Model Refused Defense
“We used a US frontier model, and the US frontier model shut down and refused to defend us because of a cyber protection put in place. Those are the cyber protections that were required by the Trump administration. So we had to switch to a Chinese model to defe…”
Gerstner: Z.ai's GLM-5.2 contains watermarks showing distillation from Mythos
“GLM 5.2 has watermarks from mythos all over it, right? So we know they were distilling, etc.”
China's open-weight GLM 5.2 rivals top Anthropic and OpenAI models
“It is within percentage points of the top closed models from anthropic and open AI in a bunch of things a bunch of evals. We don't know how good it is at bug finding yet, so there hasn't been any good testing here, but encoding a bunch of other intelligence ta…”
Stamos: US AI labs are not years ahead of Chinese rivals
“This whole conversation is predicated on the idea that like the American labs are years ahead of our adversaries. And that is just not true, right? So while Fable was shut down on Tuesday of this week, GLM 5.2 was shipped right from Zeta AI, which is a Chinese…”
Morin: GLM 5.2 on B200s achieves 150 tokens per second
“And we got this thing up and running and out of the box, it's doing a 150 tokens per second, which is like 10 times what you get out of frontier models.”
Morin: GLM 5.2 is as good or better than frontier models at coding
“I was testing it, doing the exact same coding tasks that I'm doing with frontier models. It's like as good or better.”