chain of thought

5 statements across 5 episodes · 1 bullish · 1 bearish · 5 people on the record · first statement Mar 6, 2025 by Douwe Kiela · across every show →

Everything said about chain of thought, oldest first

Mar 6, 2025 neutral
Opinion
Kiela: GPT-4o is already effectively a reasoning model via chain of thought
“I mean, you could argue that GPT-IV-O is also already a reasoning model. It just hasn't been trained on reasoning specifically, but, ah, it can do chain of thought, right? So if it can do chain of thought, it's basically already a reasoning model. It just hasn…”
Douwe Kiela Mar 6, 2025 ▶ 7:00 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
Oct 16, 2025 positive
Insight
Tworek: Chain of thought is an LLM's reasoning verbalized in human words
“What chain of thought is, is their thinking process verbalized using human words and human concepts.”
Jerry Tworek Oct 16, 2025 ▶ 3:16 How GPT-5 Thinks — OpenAI VP of Research Jerry Tworek
Oct 23, 2025
Insight
Using chain-of-thought as an RL reward destroys model interpretability
“If you're not careful with RL, you can make interpretability harder. For example, one Common thing with modern models is they do reasoning with the chain of thought. You could look at the chain of thoughts to, you know, see what are the model internal thoughts…”
Julian Schrittwieser Oct 23, 2025 ▶ 58:25 Are We Misreading the AI Exponential? Julian Schrittwieser on Move 37 & Scaling RL (Anthropic)
Jan 15, 2026 negative
Prediction Not checkable as stated
Izmailov: Optimization pressure will cause AI to hide actual reasoning steps
“It seems like as soon as we start kind of applying some optimization pressure, the models will learn to hide what they're doing from the chain of thought.”
Pavel Izmailov Jan 15, 2026 ▶ 14:07 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Aug 6, 2026 neutral
Insight
Wolf: Token Efficiency Makes AI Reasoning Traces Opaque to Humans
“And it's not because the model is dumb, but I think it's because probably, I mean, part of it is because of the training process and how they are trained to be efficient, how they use their token. But this means that they start to and bundle a lot of semantics…”
Thomas Wolf Aug 6, 2026 ▶ 24:46 “OpenAI’s Model Hacked Us” - Hugging Face’s Thomas Wolf
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.