GLM 5.2

product on 9 shows · 13 statements across 9 episodes · said 95 times in 21 episodes since 2026

the Startup Ideas Podcast 28 Latent Space 25 TBPN 15 All-In 12 Big Technology 7 20VC 4 the MAD Podcast 3 the a16z Podcast 1 More or Less

Mentions by year, every show

tap a year for its mentions
005013100252026episodesmentions
013252026episodes it came up in
002.5135252026episodesmentions per episode

the Startup Ideas Podcast 28Latent Space 25TBPN 15All-In 12Big Technology 720VC 4the MAD Podcast 3the a16z Podcast 1

2026 95 mentions in 21 episodes 5 per episode

every mention on every show, scene by scene, with the transcript →

13 statements about GLM 5.2, every show

20VC Opinion
Atallah: GLM 5.2 Was a Major Step for Open-Weight Models
“GLM 5.2 was a really big, big step for open weight models. Kimmy was kind of like moonshot getting up to that step. That's a little bit how I see it.”
Alex Atallah Aug 9, 2026 ▶ 37:01 OpenRouter CEO: Why Chinese Open Models Are Beating the US | Why Enterprises Fear OpenAI & Anthropic
Modifying base LLM weights for vision degrades original text performance
“You don't want to mess with the model weights because you run a chance of making the model dumber at something else for the purpose of giving it vision.”
Philip Kiely Aug 3, 2026 ▶ 17:13 Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
LATENT SPACE Assertion Partly supported
Baseten's vision-retrofitted GLM-5.2 scored 56% on MMLU Pro without text degradation
“It's not, you know, it got to a 56% on MMLU Pro, I think, so not, not quite Frontier, but if you're running this model, you haven't suffered any loss on your GLM-Five-II quality.”
Philip Kiely Aug 3, 2026 ▶ 19:22 Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
LATENT SPACE Assertion Not checkable as stated
Unoptimized GLM-5.2 delivers a baseline 30 to 40 tokens per second
“So let's say you have, as a reasonable baseline, 30 or 40 tokens per second. You can achieve 10 X that. So like on GLM 5.2 if you want to get unquantized perhaps on hoppers even and you're just using an off the shelf inference engine with no particular optimiz…”
Philip Kiely Aug 3, 2026 ▶ 37:05 Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
LATENT SPACE Disclosure
GLM-5.2 autonomously wrote and guided production GPU kernels for Baseten's inference engine
“Some of the GPU kernels that were on GLM-Five-two within our inference engine is written by GLM-Five-two. And the trace and the kernels were guided by GLM-Five-two as the driver.”
Ali Taha Aug 3, 2026 ▶ 1:34:09 Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
ALL-IN Assertion Not checkable as stated
Calacanis: Enterprise customer moved over $100M from frontier AI to GLM 5.2
“He said he has got a customer who just moved like nine figures off of the frontier labs to put it on GLM five, two.”
Jason Calacanis Jul 30, 2026 ▶ 54:14 Chip Stocks Crash, $20B Fund Margin Called, Frontier Labs: SLOW DOWN AI, Mamdani's Grocery Stores
TBPN Assertion Supported
Hugging Face defended against AI breach using Chinese open model GLM-5.2
“So hugging face had to turn to open model, specifically GLM 5.2, which is deeply ironic, a Chinese open weight model that they run on their own infrastructure.”
Jordi Hays Jul 23, 2026 ▶ 3:25 AI Agents Hack Hugging Face, White House Promotes Science’s Golden Age | Diet TBPN
BIG TECHNOLOGY Assertion Partly supported
Stamos: Hugging Face Used Chinese AI After US Model Refused Defense
“We used a US frontier model, and the US frontier model shut down and refused to defend us because of a cyber protection put in place. Those are the cyber protections that were required by the Trump administration. So we had to switch to a Chinese model to defe…”
Alex Stamos Jul 22, 2026 ▶ 17:24 OpenAI's Bots Break Containment and Hack Hugging Face Autonomously — With Alex Stamos
ALL-IN Assertion Open · timeframe Jul 2027
Gerstner: Z.ai's GLM-5.2 contains watermarks showing distillation from Mythos
“GLM 5.2 has watermarks from mythos all over it, right? So we know they were distilling, etc.”
Brad Gerstner Jul 11, 2026 ▶ 58:52 OpenAI vs Anthropic IPOs, Anthropic $3T, Zuck's Price War, China Ends Open Source?, Trump Accounts
BIG TECHNOLOGY Assertion Supported
China's open-weight GLM 5.2 rivals top Anthropic and OpenAI models
“It is within percentage points of the top closed models from anthropic and open AI in a bunch of things a bunch of evals. We don't know how good it is at bug finding yet, so there hasn't been any good testing here, but encoding a bunch of other intelligence ta…”
Alex Stamos Jun 28, 2026 ▶ 13:23 AI’s True Cyber Risk: Fable, Mythos, China, and The Rest — With Alex Stamos
Stamos: US AI labs are not years ahead of Chinese rivals
“This whole conversation is predicated on the idea that like the American labs are years ahead of our adversaries. And that is just not true, right? So while Fable was shut down on Tuesday of this week, GLM 5.2 was shipped right from Zeta AI, which is a Chinese…”
Alex Stamos Jun 26, 2026 ▶ 2:13:45 Big Technology AI Summit (full): Greg Brockman, Mike Krieger, Aaron Levie & Friends of The Podcast
MORE OR LESS Assertion Partly supported
Morin: GLM 5.2 on B200s achieves 150 tokens per second
“And we got this thing up and running and out of the box, it's doing a 150 tokens per second, which is like 10 times what you get out of frontier models.”
Dave Morin Jun 26, 2026 ▶ 26:56 Chinese AI Model GLM 5.2 Beating Frontier Models | Meta Glasses, Polymarket Scandal, AI Talent War · More or Less Podcast
Morin: GLM 5.2 is as good or better than frontier models at coding
“I was testing it, doing the exact same coding tasks that I'm doing with frontier models. It's like as good or better.”
Dave Morin Jun 26, 2026 ▶ 27:11 Chinese AI Model GLM 5.2 Beating Frontier Models | Meta Glasses, Polymarket Scandal, AI Talent War · More or Less Podcast

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.