Q Do we have any, this is like already veering off from Claude directly into speculation, but do we have any idea, um, if there are any material differences between how Claude extended thinking works versus like the old series models? Do we know?
A The biggest difference seems to be, at least, and this is kind of a thing that's been, I mean, I don't know, this is all speculation, of course, but from the start, Anthropic had always kind of had this, like, little thinking thing where you could, Sometimes even like cloud 3.5 would do like a tiny bit of thinking, and it was really just like deciding which tool to use for the most part. Like if it was doing, um, an artifact in the cloud UI, it would have this little thing where it would think for like two sentences about which tool to use. And it seemed like Anthropik's kind of attitude has been that extended thinking is an instance of tool use and that it's the kind of thing you want to equip the model with the ability to do. But it's not like, oh, it's a thinking model. It's just a sync for the model to, like, brain vomit, because that brain vomiting will help it, like, find a nice thing to do next. In the same way that doing search or doing code execution are, like, ways to kind of get more information on the path towards, like, finishing a problem.
AI assessment note: “The biggest difference seems to be... extended thinking is an instance of tool use”