Dated predictions and their status, on one timeline. Left of today: already
assessed. Green held up, red didn't hold up, amber partly. Right: still open; when a dot's date hits, the ledger
checks it again against the public record.
Prediction Didn’t hold up
Andreessen: AI will never transform existing US K-12 public classrooms
“How are we going to apply AI in education? The answer is we're not because it's a literal government monopoly. It is never going to change the end, and there is nothing to do. By the way, you can create an entirely new school system. Like that's the one thing …”
Prediction Held up
Andreessen: Autonomous AI agents will inevitably hire humans for tasks
“The agent hiring the people, which of course is going to happen, right? It's obviously going to happen.”
Prediction Held up
Nelle: Developers will spend thousands to tens of thousands monthly on agents
“I think as we think about these highly parallel kind of agents running off for a long time in their own VM system, We are already at that point where people will be spending thousands of dollars a month per, per human, and I think potentially tens of thousands…”
Prediction Didn’t hold up
Nair: LLM agents will hit $1T before robotics hits $10B
“It feels like LLM agents are going to be like a trillion dollar market before robotics is maybe even like a ten billion dollar market.”
Prediction Held up
Brockman: Most AI compute will shift from training to inference
“We're going to move from a world where most of the compute is training the model as we've deployed these models more, you know, more of the compute goes to inferencing them and actually using them.”
Prediction Didn’t hold up
Mohan: Automated PR generation will require specialized models trained on diffs
“A lot of things people are excited about right now are I write a comment and it generates a PR for me. And that's like really awesome in theory. I think that's like a really cool thing. And I'm sure at some point we will be able to get there. That will probabl…”
Prediction Partly held up
McCloy: ChatGPT Search Bans for Prompt Injection Are Coming
“I think it works until it stops working.
Right.
And I would say like, there's not a lot of stories of people getting banned for like Chatsby D search so far, but it's coming.”
Prediction Held up
Ameisen: Deceptive Backward Reasoning Exists in Base Pre-Trained Models
“I bet, I don't know how much I bet a hundred bucks. So somebody can like, they would get a hundred bucks from me if they prove that I'm wrong, that this behavior for a model that does a drink fine tuning, it also does it post pre-training.”
Prediction Didn’t hold up
Swix: Overcast will basically never have searchable transcripts
“I should have a podcast that has transcripts that I can search. Very, very basic thing. Overcast will basically never have it.”
Prediction Held up
Ben-Smith: AI will enable natural language steering of recommendation algorithms
“I think what actually AI will enable is not that you bring your own algorithm, but you will be able to talk. You will be able to communicate with the algorithm.”
Prediction Held up
Klein: Authentication providers will offer dedicated login features for AI agents
“I think there'll be agent off in the future. I don't know if it's going to happen from an individual company, but actually authentication providers that have a You know, hidden login as agent feature, which will then you put in your email. You'll get a push no…”
Prediction Held up
Bach: Smaller, more powerful models will ensure unconstrained AI remains accessible
“Yes, but there will also be better jailbroken models or models that have never been jailed before, because we find out how to make smaller models that are more powerful.”
Prediction Held up
Multimodal models will completely supplant text-only large language models
“I actually think like it's really clear today. Multimodal models are the default foundation model, right? It's just going to supplant LLMs. Like why did you just train a giant multimodal model?”
Prediction Held up
Patel: Google and Amazon will borrow debt to fund AI infrastructure
“Google and Amazon haven't taken on debt yet for AI infrastructure, but they will, right?”
Prediction Held up
O'Laughlin: CXL will take off as operators pool old DDR4 memory
“This CXL technology that kind of never really took off is going to take off just because what they're going to do is they're going to take DDR four. They're going to take the oldest, every bit of spare memory they can find, and they're going to put them into r…”
Prediction Didn’t hold up
Swyx: OpenAI will always release both general and Codex model variants
“I'm pretty, like, have pretty high confidence that basically OpenAI will always release a GPT-V and a GPT-V codex.”
Prediction Held up
Yegge: AI coding will fragment into many specialized, fine-tuned models
“And that, that fragmentation of models actually, we expected to continue and proliferate, right? Because we are fundamentally, we're a recommender engine right now. We're recommending code to the LLM. We're saying, may I interest you in this code right here so…”
Prediction Didn’t hold up
Cheah: Standard Transformers Will Never Scale to Ten Million Tokens
“I think what was quick, I think it was rather quick after I concluded that transformer as it is will not scale to ten million tokens.”
Prediction Partly held up
Swyx: Windsurf will stick around as an active product post-acquisition
“I think Windsurf as a product is going to stick around, and people who really like Windsurf, I mean, I was a Windsurf user for a long time we are going to keep using it because it fills a need, and like, obviously, Cognition bought it for a reason.”
Prediction Held up
Lambert: Labs will surely use parallel-compute models to generate synthetic data
“Well, I bet people, I mean, they surely will use these for synthetic data. It's just like the marginal gain on synthetic data is always very high.”
Prediction Held up
Packer: ChatGPT will likely use sleep-time compute to learn offline
“Like if you activate sleep time compute on a chatbot like ChatGPT, it can like learn about you as you're not on ChatGPT.com. I think that's, you know, kind of what they're probably going to try to do. That's the direction they're going in.”
Prediction Held up
Conrad: Test-time inference will significantly expand inference compute demand
“The thing I do feel reasonably confident about saying is that the test time inference is probably going to quite significantly expand the amount of compute that was used for inference.”
Prediction Held up
Colvin: Pydantic AI will be first framework implementing OpenTelemetry GenAI attributes
“I suspect Pedantic AI will be the first agent framework that implements those semantic attributes properly, because again, we control Pedantic AI, and we can say this is important for observability, whereas most of the other agent frameworks are not maintained…”
Prediction Held up
Biggio: Real-time video/vision is likely the next Realtime API feature
“To use ChatGPT's voice mode as an example, like we've demoed the video, right? Like real-time image, right? So I'm not actually sure what timelines are, but I would expect, if I had to guess, that like that is probably the next thing that we're going to be mak…”