Google Chief Scientist Jeff Dean speaks with Y Combinator's Diana Hu about common founder misconceptions, arguing people underestimate multi-day autonomous agent systems.
Opinion
Jeff Dean: AI models have reached junior engineer capability level
“The models have been getting a lot better at sort of agent-based, longer-running coding tasks, and it seems pretty clear that they are now actually pretty capable, and depending on exactly your definition of junior engineer, it seems pretty spot on, I would sa…”
Prediction Not checkable as stated
Dean: Specialized inference hardware will surpass general GPUs and TPUs
“I think, ah, you're gonna see more and more, ah high performance and low energy inference hardware systems, because I think everyone is now realizing that inference is the key to making, you know, these agent-based systems be available to more and more people,…”
Insight
Dean: AI startups should target tasks with 1% model success, not 20%
“If they're completely failing, that's probably a good sign. If they're kind of able to do some of it, but not very well, That's maybe not a great sign because that's probably a sign that the capability is starting to be present in those models and with more tr…”
Prediction Not checkable as stated
Jeff Dean: AI models will not have good taste in problem selection
“I think models are not necessarily going to be that good at it. So you're going to have people steering a lot of AI assisted computation in order to accomplish great things and more quickly. But that essence of what it is you want your models to do is the key …”
Prediction Not checkable as stated
Jeff Dean: There is no impediment to fully automating ML model research
“There's no, ah, you know, real impediment to making that be a much more automated loop, where the model itself decides it's going to explore, or maybe with a nudge from some people, ah, at the various highest level, like, oh, why don't you try some new ideas a…”
Insight
Jeff Dean: Inference-time compute search improves reliability in agent workflows
“Inference time compute to perform search over plausible ways of solving the problem that can get much, much higher performance or much more reliability in, Long-running agent flows.”