Multimodal Language Models

topic on 2 shows · 2 statements across 2 episodes

Invest Like the Best All-In

2 statements about Multimodal Language Models, every show

Levine: Multimodal LLMs hold broad knowledge but lack physical grounding
“Multimodal language models are really good at pulling in knowledge and trying to articulate that knowledge. They're not very good at, like, grounding that knowledge in physical situations, but they know stuff.”
Sergey Levine Mar 31, 2026 ▶ 12:00 World's Top Researcher on AI, LLMs, and Robot Intelligence · Invest Like The Best
ALL-IN Insight
Sergey Brin: Pre-multimodal robotics efforts feel 'silly' in hindsight
“It, yeah, it just feels kind of silly having done all of that work and seeing now how capable these general language models are that include, for example, vision and image, and they're multimodal, and they can understand The scene and everything, and not havin…”
Sergey Brin Sep 10, 2024 ▶ 12:26 Sergey Brin | All-In Summit 2024

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.