Molmo

product on 1 show · 5 statements across 1 episodes · said 15 times in 3 episodes since 2024

Latent Space 15

Mentions by year, every show

tap a year for its mentions
00821532024episodesmentions
0232024episodes it came up in
002.51.5532024episodesmentions per episode

Latent Space 15

2024 15 mentions in 3 episodes 5 per episode

every mention on every show, scene by scene, with the transcript →

5 statements about Molmo, every show

LATENT SPACE Assertion Partly supported
Molmo 1B Matches GPT-4V Across Academic Benchmarks and Elo
“The most efficient model, the one B is based on their one B MOE. That one matches performance of four V on most academic benchmarks and their ELO ranking.”
Vibhu Sapra Oct 13, 2024 ▶ 45:06 [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz
LATENT SPACE Assertion Supported
Molmo Reads Clocks but Fails to Generalize to Dials
“The model didn't work on clocks and then the lead was really on clocks and no models work on clocks. So they're like, we've got to make it work on clocks. One of the interesting things is that it doesn't work on dials, even though it works on clocks.”
Nathan Lambert Oct 13, 2024 ▶ 28:34 [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz
LATENT SPACE Assertion Supported
Molmo Uses Base Model Without Instruction Tuning or Chat Template
“This is just, like, straight base model, no real instruction tuning. There's literally, like, no chat template for multi-turn. It just concatenates the messages together and, like, there's, like, go, look, good luck.”
Nathan Lambert Oct 13, 2024 ▶ 14:17 [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz
LATENT SPACE Prediction Not checkable as stated
Major Foundation Model Companies Will Train on AI2's Vision Data
“The things that this model is good at are things that all the foundation companies, like they're just going to take our data and train on it.”
Nathan Lambert Oct 13, 2024 ▶ 12:15 [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz
LATENT SPACE Assertion Supported
Molmo Outperforms Gemini 1.5 and Claude 3.5 Sonnet With 1M Samples
“They can get better than Gemini, 1.5, better than Claude, 3.5 sonnet, better than GPT for V at a much smaller size with about a million samples of data, which is very impressive, right?”
Vibhu Sapra Oct 13, 2024 ▶ 4:23 [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.