why aren't all 22 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
McCloy: ChatGPT Does Not Index or Retrieve llms.txt by Default
“I knew that there's debate about this, but I'd say the evidence is like ChatTriPT is not indexing and it's not retrieving content from LMS.txt by default.”
Prediction Partly held up
McCloy: ChatGPT Search Bans for Prompt Injection Are Coming
“I think it works until it stops working.
Right.
And I would say like, there's not a lot of stories of people getting banned for like Chatsby D search so far, but it's coming.”
Assertion Contradicted
OpenAI Spent $7 Billion on Compute, With $5 Billion for R&D
“This year, OpenAI spent seven billion dollars on compute. Only two of that was for all of their inference. The remaining five was R&D. So all of ChatGPT, all eight hundred million users, all of Sora, all of like, all, all the sort of like API volume, two billi…”
Assertion Supported
Robicheaux: 2B PaliGemma 2 Beats ChatGPT on MMVP with 47.3%
“The big result, and one of the reasons that I was really excited about this paper, is that they blow everything else away on MMVP. I mean, 47.3, sure, that's nowhere near human accuracy, which again is 94%, but for, you know, two billion language, two billion …”
Assertion Supported
Carlini: ChatGPT emitted verbatim 50+ word sequences from internet training data
“And what I can say is that the output of the model was a verbatim, at least 50 word in a row match. To some other document that appeared on the internet previously.”
Assertion Supported
Prompting ChatGPT to repeat a word indefinitely leaks verbatim training data
“One of my co-authors, Milad was working on some other random experiments, and he figured out that if you prompt ChatGPT to repeat a word forever, then it will repeat the word many, many, many times in a row, and then like explode and like just start doing rand…”
Assertion Partly supported
Howard: Modern LLMs like ChatGPT still follow ULMFiT's three-step training framework
“So we generated this three-step system. So step one was train a language model on a big corpus. Step two was fine-tune a language model on a more curated corpus, and step three was further fine-tune that model on a task. And of course, that's what everybody st…”
Assertion Supported
Houssier: Grammarly captures cross-app context that ChatGPT cannot access
“Grammarly knows, basically, where you, I say, you're switching from Google Doc to Salesforce. To LinkedIn, and now you're writing an email. So we have this augmented context even more, so like much more precise compared to something like ChatGPT, for example. …”
Assertion Supported
Sands: Over 1M Shopify merchants and Salesforce joining ChatGPT checkout
“There's over one million Shopify merchants coming soon, including some really big ones like Glossier and Viore.
this week, Salesforce announced that they're also in”
Assertion Supported
Sands: Walmart and Sam's Club signed up for ChatGPT Agentic Commerce Protocol
“In the last couple of days, Walmart and Sands Club have just signed up to also make their inventory purchasable
Through ChatGPT and the Agentic Commerce Protocol, which, like, I don't think that there is a bigger signal on a big retailer being up for it than W…”
Assertion Supported
Apple Siri routes requests based on the user's ChatGPT subscription tier
“If you sign into your ChatGPT account the Siri integration will actually use your subscription status to decide what type of model to use when it passes things over to ChatGPT. And so if you're you know just a free user you get, you know, the free model. But i…”
Assertion Partly supported
Hou: Codeium is highest-rated dev tool in Stack Overflow survey
“We are the highest rated developer tool as voted in by developers in the most recent stack overflow survey. And you'll note that this is even higher than tools like chat GPT and GitHub copilot.”
Assertion Partly supported
McCloy: Perplexity is the second-largest AI-native search platform after ChatGPT
“Perplexity is pretty big.
So I would actually say that like perplexity is the second biggest AI native search platform after ChatVT.”
Assertion Supported
Marimo AI Completions Access In-Memory Variables and Database Connections
“What's unique about doing it in Marimo is the fact that not only does Marimo see your code, but it sees all the variables in memory and it can also see your database connections, et cetera. So it can really provide rich Rich completions.”
Prediction Held up
Packer: ChatGPT will likely use sleep-time compute to learn offline
“Like if you activate sleep time compute on a chatbot like ChatGPT, it can like learn about you as you're not on ChatGPT.com. I think that's, you know, kind of what they're probably going to try to do. That's the direction they're going in.”
Assertion Supported
Weil: ChatGPT supports over 200 million weekly active users
“As we, you know, we support over two hundred million people every week on ChatGPT.”
Assertion Supported
Scialom: Meta had to reinvent scaling RLHF without published frontier research
“You have just the basics, but then when it comes to, like, ChatGPT or GPT Instruct or Cloud, No one published the details there. And so we had to reinvent the wheel there in a very short amount of time.”
Assertion Supported
Lambert: Anthropic, ChatGPT, and Bard Use Post-Generation Moderation Classifiers
“Anthropic and ChatGPT and Bard almost surely have a classifier after, which is like, is this text good? Is this text bad?”
Assertion Supported
Kant: Zhipu AI began developing models years before ChatGPT
“They started years before ChatGPT.”
Assertion Supported
McCloy: ChatGPT personalization and custom preferences directly alter AI search retrieval and sources
“ChatGPT personalization, memories, just explicit preferences, if you set them up, do definitely affect the results you get. Now, Again, obviously, there's, they're still using traditional search, so the search index itself is not necessarily personalized, but …”
Assertion Supported
Hugging Face finds LLM proxy words jumped in Common Crawl after ChatGPT
“For example, here we measured like these words ratio in different dumps of common crawl, and we can see that like the ratio really increased after chat GPT's release.”
Assertion Supported
Swyx: ChatGPT rewrote its frontend from Next.js to Remix
“Recently ChatGPT just rewrote from Next.js to Remix.”