why aren't all 67 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Prediction Not checkable as stated
Cherny predicts coding will be increasingly solved across all stacks within months
“And over the next few months, I think what we're going to see is just across the industry, it's going to become increasingly solved, you know, for every kind of code base, every tech stack that people work on.”
Opinion
Godin: Claude delivers kindness and humility, while ChatGPT overpromises
“ChatGPT's reputation with me is not good because it regularly over promises and under delivers, and it does it without kindness or humility. Whereas Claude, I don't know how they did it, at least in my experience, brings kindness and humility.”
Prediction Not checkable as stated
Schwartz: AI Startups Cannot Unseat Google Purely on Search Product Quality
“I don't think there's a chance, and this is the biggest takeaway, I don't think there's a chance that ChatGPT or Perplexity or Claude or any of these other LLM startups, because they're startups now, even, even OpenAI as a startup with Bing's part, with Micros…”
Disclosure
Wu: Anthropic is prioritizing first-party Claude products over third-party offerings
“I think one of the most important things for Anthropic is to grow the number of users that we're able to reach. One of the ways that we're able to do this is with the cloud subscriptions with our first party products. And so we just very much want to double do…”
Assertion Not checkable as stated
Claude automated growth experiments win at the rate of junior PMs
“It's delivering results, right? Like, and it's like, you can push it, press play with it. And it's like, it ultimately prints money where I'd say that the win rate is like, I would expect a senior PM to do better. Like I would say like, this is like a junior P…”
Assertion Supported
Anthropic built a chatbot before ChatGPT but withheld it over safety
“Anthropic had a version of Claude. We had a chatbot before ChatGPT was launched, and we had ultimately chosen not to launch it for safety reasons. I think the team didn't want to kick off Effectively like an AI global arms race”
Opinion
Vo: No AI tool has successfully unlocked agentic browser use
“I don't think anybody has really unlocked browser use. That is not just an OpenClaw thing. I think we look at ChatGPT Atlas, you look at Perplexity, Comet, you look at all of these kind of browser use. Plod has a browser use component. I don't think any of the…”
Opinion
Wen: Claude is not yet hireable as a product designer
“I don't think Claude is there yet. I don't think Claude is there yet in terms of a designer you would hire. I think it is not yet the strong generalist or the deep specialist. Or the crack new grad. I think it's pretty good at a first pass and at presenting a …”
Assertion Supported
Schulhoff: Claude's CBRN safeguards can still be bypassed in under an hour
“That being said, if you look at, like, anthropics constitutional classifiers, it's much more difficult to get, like, CBRN information out of clawed models than it used to be. But humans can still do it in, let's say, like, under an hour and automated systems c…”
Insight
Shlomo: Claude excels at UI while Gemini is better for complex algorithms
“Claude for example, does a really nice job with first the initial prompt, like writing the app from scratch. Then everything that has to do with UI is just fantastic. Like design is great. But then for example, Gemini is really good when you get to a very comp…”
Prediction Not checkable as stated
Mike Krieger expects AI to autonomously resolve user feedback by 2025
“Hey, I'm in the discord, the, you know, the, Cloud anthropic discord. I'm in the user for I'm on X and I'm reading things and like, here's what's emergent. That's step one. Models can do that today. Step two, which the models probably can do today. We just hav…”
Prediction Not checkable as stated
Schwartz: Top-Of-Funnel Search Queries Will Shift Into AI Overviews
“So now a lot of those keywords are going to be moving into these AI overviews. And I don't want to just focus on Google, although I think Google's going to own the entire search page forever and ever for at least for a very long time. But other engines, whethe…”
Assertion Not checkable as stated
Most engineering commits at Anthropic are now Claude-assisted
“Well, definitely. I think most commits are Claude assisted.”
Opinion
Pincus: Major LLMs differentiate on coding, not at the consumer level
“They're differentiating themselves right now on coding and there's real value and they're, that's an amazing business, but they're not differentiating themselves very much at the consumer level.”
Assertion Supported
Fadell: Dario Amodei claimed Claude writes 90% to 100% of Anthropic's code
“But at the time, Dario was saying, you know, 90 to a hundred percent of all our code's written by, you know, our, you know Claude. And, you know, we just monitor it and watch it.”
Assertion Not checkable as stated
Cat Wu: Claude infrastructure was not designed for third-party products
“It wasn't designed for third party products, which have different usage patterns than our first party ones.”
Insight
Cat Wu: AI's true 'aha moment' is autonomous execution, not advice
“The like big aha moment people have is when Claude can just like do things on your behalf. It is an amazing feeling to know that the agent is capable of doing so much more than telling you what to do. Like the agent can actually just do it itself. And when peo…”
Insight
Wu: Crafting an AI's character is harder than coding due to ambiguity
“Even coding is easier because you can verify the success. Whereas crafting the character requires a very strong sense of conviction and what, who Claude should be.”
Assertion Not checkable as stated
Wu: Claude Code leak stemmed from human error despite two human reviews
“We realized that this was the result of human error. There is a human working with Claude to write a PR. This was just an update to how we release our packages. And it actually went through two layers of human review. And so this was a result of human error an…”
Assertion Not checkable as stated
Wu: Anthropic engineers must pass Claude-generated code reviews before merging PRs
“This code review is so good that our engineering team relies on this code review to pass before we merge PRs.”
Disclosure
Avasare uses Claude to simulate weekly executive coaching from his manager
“I basically you know, one of my manager, Ami Vora was, I think, a podcast guest of yours, right? So I say, hey, based on what you know of Ami, both publicly, she's written extensively about product, and then internally, and then our discussions, what is everyt…”
Disclosure
Avasare: Anthropic repeatedly took commercial hits to delay releases over safety
“And so for us, that top line objective of this just like, this needs to go well for humanity. That is something that we are happy to take a significant commercial hit for. And we've done that time and time again, right? So like we have, you know, like back the…”
Disclosure
Anthropic uses Claude to scan Slack weekly and detect project misalignments
“With Cowork, you have this, the Slack MCP, and you can tell Cowork you can tell Cowork to say, basically, look across Slack. You know, the projects that I'm working on. These are the things that are top of mind. Go and find me areas of potential misalignment r…”
Disclosure
Anthropic Project CA$H is automating growth experimentation using Claude
“We are starting to look at how do we automate growth, which I think is like a really interesting area. So our growth platform team we have, we're very lucky. We have like Alexei Komisaroka who teaches growth engineering at Reforge, and he's just like the guy o…”
Opinion
Willison: AI prompt injection benchmarks under 100% provide false security
“And again, until it's a hundred percent, I don't think it's a meaning. I think it just gives people a false sense of security that this problem won't bite them.”
Opinion
Rachitsky: Claude is copying OpenClaw features instead of acquiring it
“Claude clearly is trying to, instead of acquiring OpenClaw, they're just like, wait, we'll build it. And so they're just slowly launching all of the features.”
Prediction Not checkable as stated
Wen: Chat interfaces for AI will never go away
“So my read here is like, I don't think Chad is ever going away because this opened up this like new way of like infinite ways to work with the model and to sort of like talk to the computer that we just didn't have before.”
Prediction Not checkable as stated
Wen: AI models will increasingly generate UIs instead of developers hand-coding them
“And I think that what will probably happen here is that a lot of those UIs will be generated more and more often by the models, as opposed to something that we're like hand coding each instance.”
Insight
Cherny: Underfunding engineering projects forces developers to automate workflows with AI
“There's this there's interesting thing that happens also when you when you underfund everything a little bit because then people are kind of forced to clodify.”
Insight
Ezinne Udezue: Generate content yourself and let AI refine it, not vice-versa
“Don't, even though you keep talking about it as generative AI, be the generator and let it refine. That's the major trick. You know, that's the hack everybody should learn.”
Opinion
Mann: Claude Is One of the Least Sycophantic AI Models
“And if you look at something like sycophancy, I think Claude is one of the least sycophantic models because we've put so much effort into actual alignment and not just trying to, like, good heart our metrics of saying, like, user engagement is number one, and …”
Disclosure
Mann: Claude writes 95% of the code for Anthropic's Claude Code team
“And in terms of software engineering, our Claude code team, like 95% of the code is written by Claude.”
Disclosure
Mike Krieger uses Claude Opus 4 as his primary strategy partner
“Where my go-to product strategy partner is Claude, and it has been basically for that full year where I'll write an initial strategy. I'll share it with Claude basically, and I'll have it, you know, look at it. And in the past, it's pretty anodyne kind of comm…”
Insight
Optimizing AI models for likability risks creating sycophancy, says Mike Krieger
“It would be very dangerous to over optimize on like Claude's likability, you know, because you can fall into things like, you know, is Claude gonna be sycophantic? Is Claude gonna tell you what you hear? Is Claude going to like prolong conversations just for p…”
Insight
PMs should query Claude about codebases instead of bothering engineers
“Second one I've already mentioned, which is like stop bothering engineers with how does the code base work and actually just go and talk to Claude and understand it, which is really helpful.
You know, I used to say early on in my career that the goal was alway…”
Disclosure
Penn: Anthropic uses research Opus models to determine future Claude pricing
“So I've used Claude to help with things like, are we making the right pricing decision on the next version of Claude? It's a little bit meta, but using a research version of Opus, asking it to figure out how is your price and being able to come out with better…”
Insight
Penn: Telling AI researchers Claude hallucinated is unactionable without failure-mode breakdowns
“If you bring that to a researcher and you say, please fix Claude from being hallucinated, it's not very actionable. And so part of the time of the team is understanding, okay, what's the trajectory of why that user gave that feedback?”
Insight
Penn: Human judgment will stay critical as AI lacks human experience
“I think judgment is one, is an area where it's accumulation of so much nuance and so much experience, and these systems haven't experienced as much as humans have. And so I think that hard earned, like, judgment Is a area for product leaders and just generally…”
Insight
Penn: Useful AI models push back on unformed ideas rather than agree
“What you don't want is like a AI that just agrees with you, right? What you want is this technology to actually augment and grow and like get to a better outcome. And so sometimes it's Having Claude push back makes me better, and so that's great. Like a cowork…”
Insight
Fung: Engineers should regularly revisit previously failed AI automations
“There might've been something I tried to automate that Claude wasn't quite good enough. And then actually in the next model, oh, well now it is good enough. So it's always also thinking about what may have not worked. Like it might be worth the time to revisit…”
Assertion Partly supported
Kalinowski: Current AI models cannot generate true parametric solid CAD models
“Claude can do what is essentially surfaces or point clouds. This is not real CAD. Real CAD is in my world is dense. Like it has shape. It has nerves. Like you have an equation for how, how the surfaces work. And it's an entity that's designed in CAD. It's a so…”
Insight
Wu: Great AI assistants need bias towards action and honest pushback
“I think part of what makes a great coworker is this positivity, this like bias towards action, this ability to give you like earnest feedback, not just agreeing with every single thing that you say. And so we try to imbue this into COD because we think it make…”
Insight
Willison: LLM search integration is now better at searching than humans
“Now that all of the major models have really good search integration, they're just better at searching than I am. I can ask them a question and watch them fire off five searches in parallel for like aspects of answering that question, pull the data back.”
Assertion Not checkable as stated
Cherny: Claude automatically reviews 100% of pull requests at Anthropic
“Here at Anthropic, Claude reviews a hundred percent of pull requests.”
Insight
Arnovitz: AI coding tools differ by harness, not underlying model
“The main difference between all these tools is basically the harness. So the models are all the same models. You know, I I'll run Claude within cursor. I'll run it within Claude code. And it's also the models that Claude is also the model that is underlying Bo…”
Assertion Supported
Balfour: ChatGPT has at least 10x the monthly active users of Claude
“I think ChatGPT at this point has Like at least a 10 X difference on MAU.”
Opinion
Balfour: AI chat platforms are currently in Step 0 of their lifecycle
“And I think we could all agree that we are in that Mode. You know, right now we've got open AI battling with Claude, battling with Gemini in Google with whatever meta comes out with their new team, you know, so on and so forth. Like, and there's huge amounts o…”
Insight
Krieger: Claude is a very good prompter of itself
“Watch our prompt improver and then note that like Claude itself is a very good prompter of Claude.”