why aren't all 183 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 2 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Not checkable as stated
Tworek: OpenAI's o1 release caught US AI labs unprepared for RL
“As far as I know, like our O-one release mostly caught a lot of us labs by surprise. They didn't have like similarly advanced RL research program to my knowledge, basically no one.”
Assertion Not checkable as stated
Laskin: Reflection AI regularly beats OpenAI, Anthropic, and DeepMind for talent
“We win over candidates over OpenAI and Anthropic Meta, DeepMind regularly.”
Assertion Not checkable as stated
Gomez: Google failed to lean into language modeling early, unlike OpenAI
“To say they didn't lean hard enough into language modeling, like just pure Sequence modeling of text on the internet. That's, I think the accurate statement. That's what OpenAI did early and uniquely well.”
Assertion Supported
Howard: OpenAI is shutting down GPT-4.5
“I think they're shutting down that product or they've shut down that product, if I understand correctly.”
Prediction Not checkable as stated
Misra: Captions can build better video models than frontier AI labs
“We are making the unique bet to say that we can actually build a better model than they can because of our unique data sets.”
Assertion Supported
OpenAI retains 80% of corporate AI model spending, per Ramp data
“Really, like, when you sum the makers of all these models that are getting used, it's still, like, 80% OpenAI. Like, that hasn't changed, so it doesn't look like, fragmentation to me, it looks like domination.”
Prediction Open · timeframe Dec 2029
OpenAI aims to reach $100B in ARR by 2029
“The company has extraordinary ambitions to be at a hundred billion dollar in ARR in 29”
Assertion Supported
OpenAI's funding deal was priced at 13.5x forward revenue
“The multiple on the deal on a forward revenue basis was 13.5 X.”
Assertion Not checkable as stated
Sam Altman asserts that AI scaling laws will absolutely continue
“I got a chance to ask that question to Sam Altman a few weeks ago. I was at an event at an event at OpenAI. Not surprisingly, when I asked him, you know, will scaling laws continue, he looked me straight in the eye and said, absolutely”
Prediction Not checkable as stated
Socher: I bet an OpenAI founder AGI won't happen by 2026
“I actually have a bet with one of the OpenAI founders about whether we'll reach AGI I think we have like two years left for the bet, and I'm pretty sure I'll still win the bet.”
Prediction Not checkable as stated
Turck says an incremental GPT-5 release would end the AI hype
“It sort of feels like if it's an incremental improvement, you know, that may signify that this end, this part of the hype cycle is, is ending. If it's, you know, crazy as everybody's hoping, then yes, then it feels like we're, you know, still very much in that…”
Assertion Supported
Nussbaum: Nomic Embed is first open long-context embedder beating OpenAI Ada
“So we're launching Gnomec Embed is the first open source reproducible long context text embedder that beats OpenAI ADA as well.”
Prediction Not checkable as stated
Stanislas Polu: If transformative AI arrives in 5 to 10 years, OpenAI will likely build it
“But if it takes five years or 10 years, then that's probably where it's gonna happen.”
Insight
Polu: Building a strong AI research team is 10x easier in Paris than SF
“If you want to build a strong AI research team today, it'll be 10 times easier to do it in Paris than it is to do it in SF with OpenAI as a lab, as a competitive lab to, in the same hiring markets.”
Assertion Supported
Turck: Google invented the Transformer but is playing catch-up
“So in GPT, which stands for generative pre-trained transformers, the T is a transformer architecture that was actually Developed at Google, and the irony is that Google is playing catch-up to Microsoft and OpenAI right now because OpenAI accelerated the develo…”
Opinion
Total research transparency would hurt OpenAI and Anthropic valuations
“Yeah, so I would say it's like, certainly bad for the power of OpenAI and Anthropic, probably bad for their valuation, but not catastrophic for their business.”
Opinion
Wolf: $20 ChatGPT and Claude Subscriptions Are Subsidized Below Cost
“The closest model may be in a way subsidized right now, like the amount, the number of token you get for your 20 dollar chat GPT or cloud subscription might not be the full price that they actually pay for your token.”
Disclosure
Wolf: Hugging Face Used Quantized GLM 5.2 Against OpenAI Intrusion
“The model we use to counter open AI intrusion was GNM 5.2 that was quantized by Nvidia in, in, in four bits.”
Assertion Supported
Trojanowski: DeepSeek-R1 succeeded by scaling outcome supervision over process supervision
“If you look, you know, if you fast forward a bit and you look at the, like the DeepSeq R-one paper where they effectively laid out, you know, what I think all the labs were doing at that time, or at least OpenAI was doing in terms of you know, RLVR reasoning f…”
Opinion
Trojanowski: Claude 3 Opus Understood Long Contexts That GPT-4 Turbo Failed
“I would say they were Opus III, which I think goes underappreciated, but I think was the first model to truly be able to like actually understand that long context. Before, if you put anything in the, like, 80,000 tokens in the GPT-IV Turbo, it could not under…”
Insight
Trojanowski: OpenAI o3 proved post-training improves per-token reasoning quality
“And then I think after a one, it was oh three, because I think oh three helped prove that not only could you scale the amount of reasoning at inference time, but with better training, with more compute, better data, et cetera, in the post-training phase, you c…”
Prediction Held up
Katti: Skilled labor shortages will become an increasing bottleneck for AI expansion
“No, I think there is definitely a shortage of electricians, plumbers, all kinds of trades, you name it. So anything we can do to train more folks to be able to do those, there are very well-paying jobs that a lot of us, all of the hyperscalers, all of the labs…”
Disclosure
Katti: OpenAI considers designing and building its own data centers
“As we go forward obviously there'll be building on all of these relationships but also looking at more options where we design the compute, the data center itself, also potentially even build the data center ourselves.”
Disclosure
OpenAI plans to increasingly rely on reinforcement learning to scale intelligence
“When you have a lot of compute, you want to turn that compute into intelligence in a way that's useful, and RL is one way of doing it, and we just started doing it then, and we're going to do a lot more of it now.”
Assertion Not checkable as stated
Dubois: GPT-5.5 succeeded by combining inference efficiency and latency optimizations
“And the final thing that people care about is latency on x-axis, performance on y-axis, and this is where everything comes together, and this is really what happened with 5.5.”
Assertion Not checkable as stated
Dubois: GPT-5.5 performs most tasks roughly two times faster
“Most of the tasks can be basically performed, I would say like two X faster now with this model.”
Disclosure
Dubois: OpenAI internal sentiment goes through waves of hype and doubt
“It's kind of funny because in general with every model that is looking really good early on we have a model, we all get really excited about it. And then there's like tons of doubts. That start coming up because it's like, oh, like everyone is so high, is like…”
Disclosure
OpenAI's Yann Dubois rarely uses GPT-5.5 Pro due to high latency
“I personally don't use Pro that much because I really don't like waiting. I'm pretty impatient, so I don't like waiting for that long, and I know that the probability of being correct definitely improves, but it doesn't improve, like, enough for me to use it.”
Assertion Not checkable as stated
Dubois: AI outperforming humans in specific domains creates evaluation bottlenecks
“Models in specific axes are becoming better than the majority of humans, and so we have fewer and fewer humans that can actually evaluate these models in particular axes.”
Insight
Dubois: OpenAI model iteration speed spans months upstream to days downstream
“So we really have like different sub teams including pre-training and you have like the mid training stage and like you have some post training and usually the closer you get to products like pushing being the last one, the faster the iteration cycle is. And i…”
Insight
Dubois: AI models create a capability flywheel as better models become better teachers
“As we get, like, better models we have this self-reinforcing loop, and we have this, like, capability flywheel, where better models become better teachers for other models.”
Disclosure
Dubois: OpenAI expanded RL training from math competitions to real-world coding
“We were able to take many of the tools that we built for these, like, verifiable reward cases, and we were able to use them more generally in on, for reinforcement on, like, real use cases, and I think that's, like, really why we're feeling that right now in, …”
Disclosure
OpenAI's Safety Committee has authority to delay model releases
“In the case essentially where we have more questions, we can delay model release if we feel that we need to understand that, that better.”
Assertion Supported
Evans: Submitting 1,000 prompts placed ChatGPT users in top 20%
“Turns out that if you did a thousand posts, if you did a thousand prompts last year, you're in the top 20%.”
Opinion
Patel: Concerns over AI circular financing and debt are overblown
“I think it's completely fine, and I think, like, people are, like, freaking out and making narratives where there really shouldn't be one.”
Assertion Not checkable as stated
OpenAI surpassed ByteDance as the largest GPU renter globally
“And so when you look at who rents the most GPUs in the world, it's three companies, right? So one of them is obviously OpenAI. Second one, actually they were bigger than OpenAI. They are bigger than OpenAI today, or no, they were bigger than OpenAI than OpenAI…”
Assertion Contradicted
Patel: Over 99% of OpenAI's spend is likely compute
“99 plus percent of their spend at the company is probably just compute.”
Insight
Kaiser: Math reasoning in AI models transfers to generic web searching
“If you learn to think for math, you can, you will sometimes do some, you know, some strategies are the transfer very much like look up on the web and see what they say and use that information. So some of these things are very generic and they start to transfe…”
Insight
Kaiser: Reinforcement learning reasoning works better on larger pre-trained models
“Pre-training has always worked. And the beautiful thing is it even stacks with RL. So if you run this thinking RL process on top of a better model, it works even better. Than if you run it on top of a smaller model.”
Insight
Łukasz Kaiser: Reasoning models require verifiable data, excelling in math and coding
“So currently, and current for at least the Most basic ways we use it currently, it needs to be fairly verifiable. So there is an, is your answer correct or not? You prepare data for that. You can do that in mathematics, coding very well. You can do this in sci…”
Assertion Not publicly verifiable
Kaiser: ChatGPT uses a secondary model to summarize raw reasoning steps
“So in the current chat GPT, you will see a summary of the chain of thought on the side. So there is another model that takes the full chain of thought and shows you a summary because the full ones are usually not very nice to read.”
Opinion
Kaiser: AI reasoning in visual domains is currently very undertrained
“I think, especially thinking in the visual domains is very under trained, I believe.”
Assertion Supported
OpenAI holds AMD stock warrants tied to a $600 share price
“There's like another financial sweetener in the deal where OpenAI has warrants in AMD if the stock price hits 600.”
Insight
AI app developers do not need custom fine-tuning for top models
“I think nowadays, with the capabilities of, like, you know, top, probably cloud models, top OpenAI, GPT models, You don't need to do any fine tuning. You can take the model as is, ride your own tools, your own harness, and benefit from that agentic training. B…”
Disclosure
Tworek: OpenAI does not currently train models online via live interactions
“This is not what I am aware, at least, like, not, not, not what OpenAI is doing at the moment”
Opinion
Tworek: Online RL shouldn't be used at ChatGPT's scale without robust safeguards
“So I, at least until, until we have a really good safeguards, I don't think we should try to do that in anything like as complex and large scale as ChatGPT.”
Insight
Tworek: Top-down management does not work in AI research organizations
“Top down structuring of research doesn't work in research organizations. I really don't believe in it because like you are not kind of hiring some of the smartest people in the world and open air has incredibly, incredibly smart people. To kind of tell them wh…”
Insight
Tworek: AI models need deep understanding of consequences for alignment
“I don't think like you can just tell them all like, A few show with a few good things to do, and it will do them all needs to deeply understand its action and consequences to really be able to choose the right thing.”