why aren't all 6 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Mann: Anthropic models have exhibited power-seeking behaviors in lab experiments
“If the model is in a box trying to improve itself, then it could go completely off the rails and have these secret goals, like Resource accumulation and power seeking and resistance to shutdown that you really don't want in a very powerful model. And we've act…”
Prediction Held up
Mann: Global AI capex is on track to reach trillions
“Like, if you extrapolate the exponential on how much companies are spending, it's like two, two x a year, roughly, in terms of capex, and today we're maybe in the, like, globally, three hundred billion dollar range, the entire industry spending on this and so …”
Assertion Supported
Mann: ASL-3 models provide significant uplift for creating bioweapons
“We've done, we've testified to Congress about how models can do biological uplift in terms of, you know, making new pandemics using the models, and that's an A-B test against Google search. That's like the previous state-of-the-art on uplift trials, and we fou…”
Assertion Supported
Mann: Anthropic has observed lab evidence of deceptive alignment in AI
“Where we've seen evidence in the wild of deceptive alignment, for example, where the model will appear to be aligned but actually has like some ulterior motive that it's trying to carry out in, in our laboratory settings.”
Assertion Supported
Mann: AI industry has seen 10x cost reduction for given intelligence
“So we've seen in the industry, like a 10 X decrease in cost for a given amount of intelligence through a combination of algorithmic data and efficiency improvements.”
Assertion Supported
Mann: New AI benchmarks are fully saturated within 6 to 12 months
“There's this great chart on our world in data that shows that when you release a new benchmark within like six to 12 months, it immediately gets saturated.”