Mike Krieger, Anthropic CPO, compares UI-based computer use against protocol-based integrations (MCP) for enabling AI agency.
Assertion Not checkable as stated
Mike Krieger estimates Claude Code is over 95% written by AI
“At this point, I would be shocked if it wasn't 95% plus.”
Disclosure
Anthropic's Claude Code team uses Claude to write and review itself
“The team that works in the most futuristic way is the Claude code team, because they're using Claude code to build Claude code in a very self-improving kind of way. And, you know, early on in that project, they would do very line by line pull request reviews, …”
Prediction Not checkable as stated
Mike Krieger expects AI to autonomously resolve user feedback by 2025
“Hey, I'm in the discord, the, you know, the, Cloud anthropic discord. I'm in the user for I'm on X and I'm reading things and like, here's what's emergent. That's step one. Models can do that today. Step two, which the models probably can do today. We just hav…”
Insight
Krieger: AI benchmark evals often contain flawed ground-truth answers
“And some of these model these evals we've seen, like, even the golden answer, I'm like, I'm not sure a human would say it, or, like, I think that math is actually a little wrong, like, getting a hundred percent is gonna be really hard, because even just gradin…”
Disclosure
Mike Krieger uses Claude Opus 4 as his primary strategy partner
“Where my go-to product strategy partner is Claude, and it has been basically for that full year where I'll write an initial strategy. I'll share it with Claude basically, and I'll have it, you know, look at it. And in the past, it's pretty anodyne kind of comm…”
Assertion Supported
Anthropic models hit 72% on SWE-bench, crushing earlier timeline predictions
“He's like, I think we'll be at 90% by the end of 20, 25 or something like that. And sure enough, we're at about 72 now with the new models and. We're at 50% when you made that prediction and it's like continued to scale pretty much like as predicted.”