why aren't all 14 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Not checkable as stated
Parakhin: CLI AI tools outpace IDEs like Cursor at Shopify
“The other thing I would claim you could see is that CLI-based tools and tools that don't require you to look at the code becoming more popular, and you could see, yeah, various versions of Cloud Code and Codex and Pi and internal development tools taking off e…”
Assertion Not checkable as stated
Mohan: Codeium quality matches Copilot and drives user churn
“The product is actually one of those products where even use Copilot and use us, it's hard to tell the difference actually. And a lot of our users have actually churned off of Copilot.”
Assertion Contradicted
Friedman: GitHub Copilot user retention in enterprise is 38% to 50%
“Between 38 to 50%
Retention for users using Copilot and Enterprise.”
Opinion
Chintala: LMSYS leaderboard is biased and misses major use cases
“The LLSS leaderboard is the best thing we have right now to understand whether a model is better or not versus another model, but it's also biased and only having a sliver of view into how people actually use these models. Like, the people who actually end up …”
Assertion Not checkable as stated
Liu: Cody matches GitHub Copilot completion acceptance rates using open-source StarCoder
“Like today, Cody uses StarCoder for inline completions, and with the benefit of the context that we provide, we actually show, like, comparable completion acceptance rate metrics. It's kind of like the standard metric that folks use to evaluate inline completi…”
Opinion
Daigle: The real value of AI agents is full software lifecycle automation
“What we think is that it's not solely about the code generation. It's really about having the ability to use these you know, coding agent brained harnesses or runtimes across, not just. The coding experience where I'm going to like send a bunch of tasks out, o…”
Assertion Supported
Brockman: OpenAI's robotics team pivoted to build GitHub Copilot
“And we've been through times where, for example, robotics was one in 2018, where we had a great result, but we kind of realized that actually, like, that we can move so much faster in a different domain, right? That, that actually, you know, we had this great …”
Insight
Mohan: AI coding tools create a self-fulfilling loop by changing developer behavior
“Once you start using products like this, where in the beginning there's like skepticism, like how, how valuable can it be? And suddenly now like user behavior fundamentally changes so that now when I need to write a function, I'm like documenting my code more …”
Opinion
Copilot and Cursor Still Require Human Engineers to Drive Most Work
“Github Copilot or Cursor, which are incredible products. We use them internally. We're very happy with them. But again, this is, these are products where the engineer is driving most of the work. Like, even in agent mode, really the engineer is, like, driving …”
Assertion Not publicly verifiable
Swyx: GitHub Copilot Has at Least $1 Billion in ARR
“Copilot has a billion in ARR, I think, at least.”
Opinion
Scott Wu: AI is shifting from text completion tools to autonomous agents
“The first wave of generative AI is what I generally call these text completion products, right? And, you know, that makes a lot of natural sense if you think about it that obviously the interface of a language model is text completion, right? You give it a pre…”
Assertion Partly supported
Hou: Codeium is highest-rated dev tool in Stack Overflow survey
“We are the highest rated developer tool as voted in by developers in the most recent stack overflow survey. And you'll note that this is even higher than tools like chat GPT and GitHub copilot.”
Assertion Contradicted
Swix: Copilot report estimates 60-70% of AI-generated code is checked in
“There's a report this morning from Copilot where they were estimating the key tabs on amount of code generated by a Copilot that is then left in code repos and checked in. And it's something like 60 to 70%.”
Disclosure
Daigle: GitHub Copilot tools and agents share a single unified SDK
“Also have now had, now have like a single underlying SDK and harness for our coding agent, you know, co-pilot ultimately the new CLI, the new desktop app, cloud agents that use the same SDK.”