why aren't all 6 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Supported
Chinese AI labs like DeepSeek and Alibaba perform strongly on benchmarks
“China was kind of not in this fight, like, 12 months ago, and now is very much in it. Like, their models like Alibaba's Quen and then this spin out from a quantitative hedge fund DeepSeek, which publishes code models and others, and they've been actually a ver…”
Assertion Partly supported
Dubois: Models like Kimi and DeepSeek use ~1M RL data points
“Now when you look at reinforcement learning from models like Kimi or from DeepSeq models, it seems that they are closer to one million data points.”
Assertion Partly supported
Dan Fu: DeepSeek-V3 was trained on ~2,000 H800s with 20% MFU
“If you look at the deep seek model, for instance, this is one of the best open source models we have out there today. It was trained at the end of 2024. On last generation, kind of nerfed GPUs, H 800 instead of H 100, the 800 is nerfed by all sorts of ways fro…”
Assertion Contradicted
Fireworks AI was first to enable function calling for DeepSeek models
“We have been working on function for calling for a long time, and we are the first one to enable function calling for deep seek models.”
Assertion Supported
Over 500 DeepSeek model variants hit Hugging Face within a month
“DeepSeq for example, just within one month of releasing their new models, There are, despite DeepSeq model, extremely hard to tune and optimize, extremely hard. There are 500, more than 500 variants published on Hugging Face, optimizing for local device, optim…”
Assertion Supported
Snowflake hosted full version of DeepSeek model on platform
“We actually hosted the full version of DeepSeek, not their small model.”