Evaluator

topic on 2 shows · 2 statements across 2 episodes

Latent Space No Priors

2 statements about Evaluator, every show

NO PRIORS Prediction Not checkable as stated
Kohli: AI self-improvement for cognitive tasks will work with proper evaluators
“It should work. But as we sort of mentioned that having good evaluators is an important element, right? And so having a sort of evaluator, which can say this proposal that you have just suggested for me to improve the training process will yield A good result.…”
Pushmeet Kohli Jun 26, 2025 ▶ 26:40 No Priors Ep. 120 | With Google DeepMind’s Pushmeet Kohli and Matej Balog
Yao: Evaluator quality is the key bottleneck for agent self-reflection
“I think a key bottleneck is the evaluator, right? Basically you need to have a good sense of the signal. So for example, like if you are trying to do a very hard reasoning task, say mathematics, For example, and you don't have any tools, right? It's operating …”
Shunyu Yao Sep 27, 2024 ▶ 17:56 Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.