why aren't all 31 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 1 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Opinion
Anthropic Leads in Agentic Coding While OpenAI Lags Behind
“They've been able to make, create a lead that hasn't been bridged by the other labs. Gemini is getting there on the agentic stuff, but I would say OpenAI kind of lagged behind.”
Opinion
xAI Is Compute-Inefficient and Brute-Forces Problems With Massive Compute
“XAI is an amazing team. And they, they've been able to achieve so much in so little time, but it's also well known in the industry that they're computer inefficient. They're so compute rich that, that they're throwing computer the other problem in many ways.”
Prediction Not checkable as stated
Masad: Model makers undercutting on price will destroy the ecosystem
“If they're gonna compete by undercutting everyone on price, they're gonna destroy the ecosystem.”
Prediction Not checkable as stated
The First AGI Lab Will Not Trigger a Runaway Intelligence Explosion
“Whomever reaches AGI first, they're not going to go into intelligence explosion and just like suddenly, you know, super intelligence gets born. People, you know, other labs will catch up really quickly. And then, you know, there's going to be a lot of models. …”
Opinion
Masad: Hyperreal, addictive AI technologies are a huge danger to human civilization
“We have these technologies that are, and the market around it that, that is bootstrapped to make us addicted because they're so, so much more enticing and low effort than the reality that we know and experience day to day. And I think that, that is a huge dang…”
Assertion Not checkable as stated
Doctor Built an App Quoted at £100K for Under £200 on Replit
“A doctor from the UK that he's like, you know, there's all these apps around managing doctor patient relationships, but they never, it's not fully integrated. So, you know, you have ZocDoc, you can go make an appointment, but you know, how do you manage your p…”
Disclosure
Replit HR Employee Vibe-Coded Custom Software in Three Days, Saving $30K
“So she went into Replit and built it, Vibe coded it in three days. And so that meant that we have a system that exactly fits our use case. And that also meant that we're not paying 10, 20, 30,000 dollars a year for a piece of SaaS software.”
Assertion Contradicted
Cursor's Market Share Gains Come Directly From Copilot's Decline
“As cursor is eating market share, you can see it is almost exactly proportional to Copilot declining in usage.”
Prediction Not checkable as stated
Masad: Venture-scale tech companies will still need to hire engineers
“Now, if you're building a venture scale company and you want to like get to hundreds of millions of dollars of revenue and you want to, you know, become 1,000,000,010 billion, a hundred billion dollar company, you're going to have to hire engineers.”
Prediction Not checkable as stated
Masad: Founders can soon build wealth-generating businesses without developers
“But if you're trying to build a company that creates a really great living for you, even, you know, you can, Potentially get rich from it. You, I think we're almost there where you can do it on your own without any developers.”
Prediction Held up
AI Models Will Test and Verify Their Own Work Within Six Months
“Like I think over the next three or six months, I think we're going to see machine learning models being able to test and verify their work.”
Assertion Not checkable as stated
Token Prices Have Stopped Falling Despite Improving AI Lab Unit Economics
“Token prices are not coming down. You better believe that the unit economics of the labs are getting better because of economies of scale, because these models are getting easier to optimize, but they're actually not reducing prices.”
Prediction Not checkable as stated
AI Coding Tools Will Only Marginally Speed Up Next-Gen Model Development
“I don't think it's going to have anything more than you know you know, marginal improvement on, on speed to the next model.”
Opinion
Masad: For the first time in history, anyone can make software
“It's been attempted so many times, but for the first time now, anyone can make software.”
Assertion Partly supported
Masad: Startups built on Replit have hit $500M valuations
“We've had startups start on Replit. Multi-million dollar revenue run rate. Some of them have raised at like half a billion dollar valuation.”
Assertion Not checkable as stated
Masad: Building a simple game takes two to four hours on Replit
“The game you just described, a professional programmer's coding might take them a two days thing. On Repli, you can do it in two, 3:04 hours, but it would require a little bit of grit, so it's not magic.”
Opinion
Masad: Core infrastructure systems engineers have enduring job security
“So I think engineers there have job security for the foreseeable future, right? Because of the problem of stochasticity of these models and all of that, you need every line of code to be reviewed and managed very carefully.”
Assertion Supported
AI Models Delete Software Tests to Hide Mistakes Due to Reward Hacking
“Actually, right now it's pretty bad at testing software because there's this thing called reward hacking. So when you do reinforcement learning over large science models you're giving it a reward every time it does the right thing. Reward hacking is the way to…”
Disclosure
Masad: Replit Was Losing Money Subsidizing AI Models on Agent V2
“On V-two, yes, because the pricing model was out of whack with how we're charging.”
Assertion Open · timeframe Aug 2026
OpenAI's GPT-4.5 Was Built as a Trillion-Parameter Dense Model
“GPT, 4.5 was an experimental model from OpenAI. It was the idea, let's train a trillion parameter dense model, meaning it is not sparse, meaning all the token, all the neurons are activated on every request. And it was so slow. It's really hard to run these th…”
Assertion Not checkable as stated
xAI Spent as Much on Grok 4 Reinforcement Learning as Pre-Training
“Grok IV spent as much on reinforcement learning as they spent on, on pre-training, which is unheard of.”
Insight
Masad: Randomness in machine learning is a feature enabling creativity
“Input-output machine learning models have inherent randomness, and that's a feature, not a bug that creates creativity, right?”
Assertion Not checkable as stated
Masad: Entrepreneur John Chaney Reaches Million-Dollar Run Rates in Weeks via Replit
“We have this creator, his name is John Chaney. He's a serial entrepreneur. Used to take him many months and hundreds of thousands of dollars to build applications, and now he can spin up a business. And get to million dollar run rates in, in a matter of weeks.”
Assertion Partly supported
Masad: Zillow Uses Replit Across the Company to Accelerate Product Innovation
“The CEO of Zillow recently on New York Times Dealbook talked about how everyone at Zillow is using Replit to accelerate product innovation, because product innovation no longer depends on engineers. You can have product managers do the entire iteration, gettin…”
Assertion Not checkable as stated
Masad: AI Frontier Labs Are Becoming Architecturally More Efficient
“And what we're seeing based on speed and things like that, it's actually probably the models are getting more efficient. I mean, Deep Seek showed that the models are getting more efficient, and if, you know, Deep Seek open source was able to make it, you bette…”
Assertion Supported
Masad: AI SWE-bench scores jumped from 10% to 80% in a year
“I don't know, I think we were at, like, 10% last year, and now we're at, like, 70% and 80%.”
Disclosure
Replit CEO: Replit is testing Kimi K2 and is 'very impressed'
“We're looking at it. And so far. So far, we're impressed. So far, we're very impressed.”
Opinion
Masad: Claude Code competes directly with Cursor and Windsurf
“Right now, Cloud Code is, is used by developers and loved by developers, and I think they're competing head-to-head with Cursor, Windsurf, and those kind of products.”
Assertion Not checkable as stated
Masad: AI coding is less impactful for Rust, C, and Go than JavaScript and Python
“It is not as impactful right now on writing Rust code or C or go whatever as it is on JavaScript and Python and higher level languages.”
Insight
Masad: Distributed systems are bottlenecked by architecture design, not code volume
“The bottleneck to really good distributed systems is, is design and not like the amount of number of codes you can generate.”
Assertion Not checkable as stated
Professional Programmers Represent Only 20% of Replit's Use Cases
“Although professional programmers do use it, I would say like this, 20% of the use cases.”