why aren't all 38 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 1 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Partly supported
Evans: OpenAI has 900M weekly active users, but only 5% pay
“You've got nine hundred million weekly active users, but most of them are not using it every day and can't think of anything to do with it. And only five percent of them are paying for it.”
Assertion Partly supported
Patel: ChatGPT has roughly one billion users
“ChatGPT has a billion users roughly.”
Assertion Not checkable as stated
Google Search survives because ChatGPT relies heavily on referencing it
“People say, oh, Google search is dead. I think that's like probably completely wrong because ChatGPT references Google a ton.”
Insight
Evans: Consumer chat interfaces are thin wrappers, not vertical SaaS
“The only thin, thin GPT wrappers are what you get when you go to chatgpt.com and claude.com and grok and all these others. That's a thin wrapper on a model. Whereas you know, name your vertical enterprise SaaS company. That's not a thin wrapper.”
What-if
Levie: The AI wave would not have happened without ChatGPT
“Without chat GPT, none of this would have happened.”
Prediction Not checkable as stated
Ramaswamy: Google Gemini won't become a hit unless 10x better than ChatGPT
“They can't make Gemini into a consumer hit unless it's 10 times better than ChatGPT, and that's a really tall order.”
Prediction Open · timeframe Nov 2029
Azeem Azhar: 75% of people will use AI daily by 2029
“I think you have to be making quite a bold and, you know, heterodox claim to say that we won't all, by which I mean, 75% of us, be using these things daily in less than seven years from when ChatGPT launched, which was November the 22. And, ah, and that means …”
Disclosure
Benedict Evans still hasn't found a daily use case for ChatGPT
“In my actual job, I can't work out something that I would actually use ChatGPT for.”
Assertion Partly supported
Ratner: $200 of ChatGPT calls can clone closed models into open ones
“If you take a couple hundred bucks of API calls to, say, ChatGPT, and you graph that onto a model that is substantially smaller, say a seven billion parameter model like LLAMA, or now increasingly fully open for commercial use ones, like Red Pajama is one that…”
Opinion
Wolf: $20 ChatGPT and Claude Subscriptions Are Subsidized Below Cost
“The closest model may be in a way subsidized right now, like the amount, the number of token you get for your 20 dollar chat GPT or cloud subscription might not be the full price that they actually pay for your token.”
Assertion Supported
ChatGPT disproved an Erdős conjecture using cross-disciplinary mathematical reasoning
“The big result was that this conjecture of this lower bound for the number of pairs that you can make is, is false. Not only is it false, it was false due to a really interesting connection to another field of mathematics.”
Assertion Supported
Evans: Submitting 1,000 prompts placed ChatGPT users in top 20%
“Turns out that if you did a thousand posts, if you did a thousand prompts last year, you're in the top 20%.”
Prediction Not checkable as stated
Evans: No one will vibe code their own ERP software
“No, no one will vibe code their own ERP or their own frame.io, but they may ask Anthropic or Gemini or ChatGPT, can you do this thing for me?”
Insight
Enterprises should transition from generalist LLMs to cheaper task-specific models
“If you have a business problem, you are maybe manufacturing something, maybe you can start with a generalist model, but then once you know exactly what the task is and you want to hone in on it, maybe it makes sense to replace that expensive thing by something…”
Opinion
Leading generalist models like ChatGPT, Gemini, and Claude show functional parity
“Like if you use or compare ChatGPT, Gemini Claude, Grock. I think they are all pretty much on the same level. Like, and I think that's because they're trying to do everything. Like the generalist models for a general person to do a lot of things. I mean, Claud…”
Assertion Not publicly verifiable
Kaiser: ChatGPT uses a secondary model to summarize raw reasoning steps
“So in the current chat GPT, you will see a summary of the chain of thought on the side. So there is another model that takes the full chain of thought and shows you a summary because the full ones are usually not very nice to read.”
Disclosure
Tworek: OpenAI does not currently train models online via live interactions
“This is not what I am aware, at least, like, not, not, not what OpenAI is doing at the moment”
Opinion
Tworek: Online RL shouldn't be used at ChatGPT's scale without robust safeguards
“So I, at least until, until we have a really good safeguards, I don't think we should try to do that in anything like as complex and large scale as ChatGPT.”
Disclosure
Tworek: OpenAI's high and low reasoning modes use the exact same model
“Where you can have like a high rezoning model and the low rezoning models. And this is like in the end, the same model. You just, we just tweak the parameter, which says we want you to think longer or shorter.”
Disclosure
OpenAI VP Jerry Tworek pays $200 per month for ChatGPT Pro
“I think I am pretty heavy user of ChatGPT right now, happily paying like 200 dollars a month for it”
Opinion
Rauch: ChatGPT's UI simplicity gives it a fundamental advantage over Google
“If you look at ChatGPT, one fundamental advantage they have over Google is the simplicity of the UI.”
Insight
Knoop: LLM failure rates break unsupervised server-based automation workflows
“You know, it fails randomly two out of 10 times, which might be fine in a supervised setting like ChatGPT. You know, where you're talking with these sort of assistants but it doesn't really work in an automation case where it's hands-off keyboard running on a …”
Assertion Supported
A large cohort of internet users has replaced search engines with ChatGPT
“There's a whole cohort of humans that have already largely replaced a very, very large proportion of their traditional search behavior on search engines from looking through these search engine results pages to just trusting the answers that that come out of C…”
Insight
Zeghidour: LLMs cannot generate massive diverse synthetic voice scripts without collapsing
“I mean, you cannot ask Claude or ChatGPT write 100,000 hours of scripts and make them as diverse as possible. So it doesn't work. It's going just to be In a loop and collapse on a few topics, you know.”
What-if
Tworek: Someone from 10 years ago would view today's ChatGPT as AGI
“If you talk to someone from 10 years ago and show them chat GPD from today, they would probably call it AGI”
Assertion Partly supported
39% of Americans use ChatGPT and 28% use it at work
“39% of Americans say that they use ChatGPT, including 28% that say that they use it at work, and 11% use it every day.”
Assertion Not checkable as stated
Moody's Found OpenAI Outperformed Competing Models in Internal Benchmarks
“ChatGPT was the one, OpenAI was the one that was performing better in a, in over Different parameters, right?”
Assertion Supported
Zhou: Fine-tuning transformed GPT-3 into ChatGPT
“Fine tuning is the technology that got from a research project in 2020 called GPT-III and turned that into ChatGPT, a billion dollar app, right?”
Insight
Zhou: ChatGPT guardrails are difficult due to broad scope
“The reason why ChatGPT is hard to put guardrails on is because they're trying to go after every possible use case, right?”
Insight
Shah: Pre-generative AI functioned as an idiot savant in narrow tasks
“It's been a bit of what I call an idiot savant till now. It could do one narrow thing well, but if you took it anything outside of that, it just like didn't work. You know, at least traditional classifier AI.”
Insight
Liu: Generating long-form content over custom data remains a hard problem
“Generating something that's like a paragraph is pretty easy for ChatGPT to do. Generating like an entire blog post or essay, especially over your data is a pretty challenging problem.”
Assertion Not checkable as stated
Sternberg: Notion began developing Notion AI before ChatGPT released
“While no one saw notion AI capabilities before ChatGPT, I'm not going to give us too much credit. It wasn't way, way, way before, but we were definitely working on this a little bit before that, and there was a pretty big focus at that point.”
Disclosure
Liberty: Pinecone did not foresee the massive ChatGPT-driven AI surge
“We didn't, by the way, foresee any of this ChatGPT thing happening. We knew it was, it would keep growing, but that, I think, completely took everybody by surprise, including us.”
Assertion Supported
Hierarchical Reasoning Model matches larger LLMs on ARC using Transformer architecture
“Hierarchical reasoning model, it became like popular because it performed relatively well on that benchmark compared to very expensive models like Gemini, Chachupiti, and so forth. And it is a transformer architecture.”
Assertion Supported
Evans: ChatGPT usage on Google Trends drops sharply in summer and Christmas
“I mean, it's funny if you look at Google Trends. There's a big sag in the summer and then a big sag in the Christmas week.”
Assertion Supported
Many people use ChatGPT as an informal therapy tool
“A lot of people use ChatGPT to just, like, sort of vent and talk.”
Disclosure
Awadallah: Vectara provides ChatGPT for proprietary enterprise data
“So what we do is ChatGPT for your own data.”
Assertion Supported
Kambadur: Stack Overflow banned ChatGPT for generating plausible but incorrect answers
“ChatGPT got banned from stack overflow because it sounded convincing, but wasn't quite right all the time.”