Chain Of Thought Reasoning
topic on 3 shows · 6 statements across 5 episodes
the Y Combinator Startup Podcast
Latent Space
No Priors
6 statements about Chain Of Thought Reasoning, every show
Hassabis: Thinking modes and chain of thought revive AlphaGo techniques
“Really you can think of a lot of The things we're doing today, all the leading models with thinking modes and chain of thought reasoning as aspects of what was sort of pioneered with AlphaGo coming back now.”
Programmatic tools beat end-to-end image generation for multimodal reasoning
“Where I think for a while some people were, like, speculating, like, oh, what if you have the model, like, generate images in its chain of thought reasoning where everything is, like end-to-end multimodal input and output, and it seems like you don't really ne…”
Lambert: AI community will clarify if chain-of-thought maps to RL within a year
“I think in the next year that'll probably get kind of made more concrete by the community on, like, if you can easily draw out, like, if chain of thought reasoning is more like RL.”
Lambert: RL feedback mechanisms will specialize across distinct task domains
“It seems very likely that different feedback will be used for different domains.
Chain of thought reasoning is great.
For math, and that's where these process reward models are being designed.
Probably not great for things like poetry, but as any tool gets bet…”
Liu: Chain-of-thought query decomposition produces superior retrieval results
“Another example here is actually LLM based reasoning, like LLM based chain of thought reasoning. You can take a question, break it down into smaller components and use that to actually send to your retrieval system. And that gives you better results since kind…”