Jeff Dean discusses the feasibility and future of recursive self-improvement in AI systems with Y Combinator's Diana Hu.
Opinion
Jeff Dean: AI models have reached junior engineer capability level
“The models have been getting a lot better at sort of agent-based, longer-running coding tasks, and it seems pretty clear that they are now actually pretty capable, and depending on exactly your definition of junior engineer, it seems pretty spot on, I would sa…”
Prediction Not checkable as stated
Dean: Specialized inference hardware will surpass general GPUs and TPUs
“I think, ah, you're gonna see more and more, ah high performance and low energy inference hardware systems, because I think everyone is now realizing that inference is the key to making, you know, these agent-based systems be available to more and more people,…”
Insight
Jeff Dean: AI agents can run autonomously for days or weeks
“Probably one thing is people don't quite realize how possible it is to have, you know, agent based systems that can run not just for an hour or two hours on a problem you care about, but for some problem domains and with highly capable models underlying them, …”
Insight
Dean: AI startups should target tasks with 1% model success, not 20%
“If they're completely failing, that's probably a good sign. If they're kind of able to do some of it, but not very well, That's maybe not a great sign because that's probably a sign that the capability is starting to be present in those models and with more tr…”
Prediction Not checkable as stated
Jeff Dean: AI models will not have good taste in problem selection
“I think models are not necessarily going to be that good at it. So you're going to have people steering a lot of AI assisted computation in order to accomplish great things and more quickly. But that essence of what it is you want your models to do is the key …”
Insight
Jeff Dean: Inference-time compute search improves reliability in agent workflows
“Inference time compute to perform search over plausible ways of solving the problem that can get much, much higher performance or much more reliability in, Long-running agent flows.”