why aren't all 11 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Insight
Building hyper-specialized AI products is risky as models dynamically build infrastructure
“As the models get more and more capable, what I'm noticing inside my products and inside my work is that we're sort of like pulling back the edge cases we account for. And I mentioned earlier that memory is just a text file. If Claude needs a database, it will…”
Assertion Not checkable as stated
Patel: Engineer built an RTS game using $10K of Claude API
“He used, like, 10,000 dollars of Claude in one week and built an entire RTS from scratch about, like, but instead of, like, being a standard RTS where it's like, oh, Age of Empires where you advance through ages or Starcraft, it is an RTS where it's China vers…”
Insight
Evans: Consumer chat interfaces are thin wrappers, not vertical SaaS
“The only thin, thin GPT wrappers are what you get when you go to chatgpt.com and claude.com and grok and all these others. That's a thin wrapper on a model. Whereas you know, name your vertical enterprise SaaS company. That's not a thin wrapper.”
Opinion
Wolf: $20 ChatGPT and Claude Subscriptions Are Subsidized Below Cost
“The closest model may be in a way subsidized right now, like the amount, the number of token you get for your 20 dollar chat GPT or cloud subscription might not be the full price that they actually pay for your token.”
Opinion
Balaban: AI agentic workflows without clear automated feedback are overhyped
“I think a lot of the sort of agent, agentic workflows for things that are not software engineering, I think tend to be overhyped. And I'll tell you that the reason for that is because one of the ways that you get an agentic workflow working really well is that…”
Disclosure
Rieseberg: Anthropic Is Prioritizing Local Execution for Claude Cowork
“So in the short term, I want to make it very possible for Claude to meet you where you're working. If you're working on your local computer, that's where Claude should be.”
Opinion
Leading generalist models like ChatGPT, Gemini, and Claude show functional parity
“Like if you use or compare ChatGPT, Gemini Claude, Grock. I think they are all pretty much on the same level. Like, and I think that's because they're trying to do everything. Like the generalist models for a general person to do a lot of things. I mean, Claud…”
Insight
Polu: Enterprise AI needs frontier models rather than complex query routing
“Most of the tasks are pretty general, right? Most of the tasks are pretty like a human would do. And so you just want the best models. And as it happens today, the best models are before enclosed. So that's what you want.”
Insight
Zeghidour: LLMs cannot generate massive diverse synthetic voice scripts without collapsing
“I mean, you cannot ask Claude or ChatGPT write 100,000 hours of scripts and make them as diverse as possible. So it doesn't work. It's going just to be In a loop and collapse on a few topics, you know.”
Disclosure
Misra: Captions switched its script generation LLM provider to Claude
“And then, but we ended up switching to Claude. So now we're on Claude.”
Opinion
Socher: Anthropic's Claude models excel at legal reasoning tasks
“One very concrete example is Anthropics, quite good for legal types of reasoning and like just kind of connected to some of their morals and ethics and so on. And so that those questions are often better routed to a Claude-like model from Anthropics.”