Chain Of Thought Reasoning

topic on 3 shows · 6 statements across 5 episodes

the Y Combinator Startup Podcast Latent Space No Priors

6 statements about Chain Of Thought Reasoning, every show

Hassabis: Thinking modes and chain of thought revive AlphaGo techniques
“Really you can think of a lot of The things we're doing today, all the leading models with thinking modes and chain of thought reasoning as aspects of what was sort of pioneered with AlphaGo coming back now.”
Demis Hassabis Apr 29, 2026 ▶ 7:12 Demis Hassabis: Agents, AGI & The Next Big Scientific Breakthrough · Y Combinator
Programmatic tools beat end-to-end image generation for multimodal reasoning
“Where I think for a while some people were, like, speculating, like, oh, what if you have the model, like, generate images in its chain of thought reasoning where everything is, like end-to-end multimodal input and output, and it seems like you don't really ne…”
Will Brown May 9, 2025 ▶ 6:01 ⚡️Open Questions in Agentic RL — Will Brown (Prime Intellect)
LATENT SPACE Prediction Not checkable as stated
Lambert: AI community will clarify if chain-of-thought maps to RL within a year
“I think in the next year that'll probably get kind of made more concrete by the community on, like, if you can easily draw out, like, if chain of thought reasoning is more like RL.”
Nathan Lambert Jan 11, 2024 ▶ 18:00 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
LATENT SPACE Prediction Not checkable as stated
Lambert: RL feedback mechanisms will specialize across distinct task domains
“It seems very likely that different feedback will be used for different domains. Chain of thought reasoning is great. For math, and that's where these process reward models are being designed. Probably not great for things like poetry, but as any tool gets bet…”
Nathan Lambert Jan 11, 2024 ▶ 1:04:51 The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
Liu: Chain-of-thought query decomposition produces superior retrieval results
“Another example here is actually LLM based reasoning, like LLM based chain of thought reasoning. You can take a question, break it down into smaller components and use that to actually send to your retrieval system. And that gives you better results since kind…”
Jerry Liu Oct 12, 2023 ▶ 47:53 RAG is a hack - with Jerry Liu of LlamaIndex
NO PRIORS Assertion Not yet assessed · timeframe Apr 2023
Training OpenAI base models on code improved chain-of-thought reasoning
“So they had wanted to see what introducing code into their into their base models would do. I think it had positive effects on chain of thought reasoning.”
Alex Graveley Apr 25, 2023 ▶ 15:28 No Priors Ep. 10 | With Copilot's Chief Architect and founder of Minion.AI Alex Graveley

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.