why aren't all 6 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Opinion
Claude 3 can bypass its assistant persona to expose the underlying simulator
“Instead of having this entity, like GPT-IV, that's an assistant that just pops up in your face that you have to kind of, like, punch your way through and continue to have to deal with as a headache, instead, there's ways to kindly coax Claude into having the a…”
Opinion
Malhotra: Modern language models contain internal models of the world
“Today we have language models that are powerful enough and big enough to have really, really good models of the world. They know a ball that's bouncy will bounce, will, when you throw it in the air, it'll land, when it's on water, it'll float, like, these basi…”
Opinion
WorldSim can simulate dev teams, making AI engineers like Devin unnecessary
“I can also generate a dev team and ask it to do stuff with me, and you don't need devin.”
Assertion Supported
Malhotra: WorldSim is technically just an LLM prompt
“The world simulator is just a cool prompt and you know, functionally it's a lot more than that. Technically it's not really much more than that at all.”
Insight
Malhotra: Assistant personas are conjured by weights, not the weights themselves
“The assistant isn't the weights. The assistant, the entity you're talking to is something drummed up by the weights.”
Assertion Supported
Malhotra: Claude's system prompt is written in the third person
“With Claude, we notice the system prompt is written in third person. It's written in third person. It's written as, the assistant is X, Y, Z.”