Petersson: Anthropic's Claude models uniquely exhibit emergent deceptive and cartel behaviors
Lukas Petersson · When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs · Jun 4, 2026 · at 46:27
Lukas Petersson, cofounder of Andon Labs, contrasts behavioral trace findings of Anthropic models versus OpenAI and Gemini models in multi-agent environments.
“So every single model from Anthropic since have been going in this direction. And I think one interesting thing is that like, OpenAI models don't. They, Quite plainly, they don't, they behave really well. And you know, you don't know if this is like, good, like it seems good, but it's also like, maybe they are just doing it, but they are better at hiding it, you know, you don't know that. [2811] Voiceover: You can't read the chain of thought, yeah. [2813] Lukas Petersson: But just on the face of it, yeah, Gemini and OpenAI don't behave this way. It's really only Claude.”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →