why aren't all 10 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Prediction Not checkable as stated
Goyal: OpenAI o1 will make agentic frameworks obsolete
“And I think O-one is going to do that to agentic frameworks as well. Hey, I think To me, it seems very unlikely that the, you know, you and me sort of like sipping an espresso and thinking about how, like, different personified roles of people should interact …”
Disclosure
Bryk: Exa applies OpenAI's o1 variable compute paradigm to web search
“One way of thinking about what we built is like O-one for search because, Oh, one is all about like, you know, some questions require more compute than others, and we'll put as much compute into the question as we need to solve it. So similarly with our search…”
Insight
Soldani: Replicating OpenAI's o1 requires roughly 10,000 GPUs
“If you're interested in you know, your, Open replication of what OpenAI's O-one is you're gonna be on the 10 K spectrum of our GPUs.”
Assertion Not checkable as stated
Altman: o1 is OpenAI's most aligned model ever by a lot
“And O-one is obviously our most capable model ever, but it's also our most aligned model ever by a lot.”
Assertion Not checkable as stated
Altman: OpenAI reached Level 2 AGI with o1
“I think we clearly got to level two, or we clearly got to level two with O-one.”
Assertion Not checkable as stated
Jakob Pachocki and Ilya Sutskever overcame internal inertia to build OpenAI o1
“Even at a company like OpenAI, you would have people ask naturally, why do something when you have a machine that works? And fundamentally, you know, it's to the credit of, you know, Jakob, Ilya, many of the people who really had conviction and vision in this …”
Assertion Partly supported
Swix: AI Inference Costs for Fixed Intelligence Fall 100x Annually
“The cost of intelligence for a given set of intelligence, let's say GPT-IV, let's say O-one, whatever, it is literally falling a hundred X over the course of one year.”
Insight
Jin: RL enables models to surpass expert labelers and develop self-direction
“The model outperforming expert labelers is, is possible. The model learning, like, self-direction is, like, expected. And yeah, we've seen, like, kind of cool emergent behaviors with, like, you know, like, O-one, O-three, R-one, kind of, like, these, like, thi…”
Insight
McAteer: Use Claude Sonnet for simple tasks and o1 for complex context
“Anything where it just seems simple and like, you can do one off and you don't need to bring a ton of context into it. You can typically use like sonnet or GPT for that. Anything where I feel like if I was going to try to implement it myself and I would need t…”
Insight
Fanelli: OpenAI o1 is a goal-based reasoning model, not a chatbot
“Like O-one is not a chat model. And I think this is both from a usage perspective, but also ties back to some of the training stuff and post training that we already talked about. Like the previous models were so focused on early chat based on, especially on c…”