Gemma

includes Gemma 4, Gemma 2, Gemma 2B, Gemma 27B, Gemma 3, Gemma 3n, Gemma 9B, Gemma 1, Gemma 2 2B, Gemma Scope, Gemma 3.1, Gemma 7B and 3 more

8 statements across 5 episodes · 4 bullish · 1 bearish · 6 people on the record · first statement Mar 23, 2025 by Rishabh Agarwal · said 109 times in 20 episodes since 2024 · across every show →

Mentions by year, the whole family

brought up most by Omar Sanseviero (32), Alessio Fanelli (17), Shawn Wang (12), Rishabh Agarwal (11), Vibhu (Viboo) (8), Emmanuel Ameisen (4), Peter Robicheaux (2), Jack Morris (2)

tap a year for its mentions
003056010202420252026episodesmentions
0510202420252026episodes it came up in
0045810202420252026episodesmentions per episode
2026 58 mentions in 8 episodes 7 per episode
2025 37 mentions in 9 episodes 4 per episode
2024 14 mentions in 3 episodes 5 per episode

every mention, scene by scene, with the transcript →

Everything said about Gemma, oldest first

Mar 23, 2025
Assertion Supported
Agarwal: Filtered 9B Synthetic Data Outperforms 27B Self-Generated Data
“One thing we found consistently, so here what we had two models, nine Gemma, nine B and Gemma, 27 B, and we found consistently that actually generating data from nine B in a compute match setting is always better, even better for distilling or actually improvi…”
Rishabh Agarwal Mar 23, 2025 ▶ 17:41 The Magic of LLM Distillation — Rishabh Agarwal, Google DeepMind
Jun 6, 2025 positive
Assertion Supported
Ameisen: Multi-Hop Reasoning Circuits Are Extremely Similar Across Small and Large Models
“The way the circuit looks in Gemma, like a really small model is extremely similar to the way that it looks like a huge model, which that in itself is, I think like a pretty novel discovery. It's like, oh, you have these models that are like super different. Y…”
Emmanuel Ameisen Jun 6, 2025 ▶ 3:36 The Utility of Interpretability — Emmanuel Amiesen
Jun 6, 2025 positive
Assertion Open · timeframe Jun 2026
Vibhu: Gemma Activates Abstract Behavioral Traits Over Simple Token Completion
“It also shows internally that there's more than just token completion of, you know, this plus this equals this. No, it has some under understanding of characteristics, right? Like this is a pretty stubborn dog. It has a stubborn feature. Pretty high up that ac…”
Vibhu (Viboo) Jun 6, 2025 ▶ 21:47 The Utility of Interpretability — Emmanuel Amiesen
Jul 2, 2025 negative
Opinion
Morris: Open model creators do not use differential privacy or anonymization
“I would be extremely surprised if they do any type of like private training. Like there are these mechanisms for doing like differentially private language model training, or even just anonymization in the pre-training pipeline. I bet they don't do any of that…”
Jack Morris Jul 2, 2025 ▶ 1:02:27 Information Theory for Language Models: Jack Morris
Aug 29, 2025 positive
Assertion Not publicly verifiable
Morcos: Arcee 4.5B beat Gemma before reaching one trillion tokens
“It was beating Gemma pretty consistently before the one trillion mark, which was pretty cool to see.”
Ari Morcos Aug 29, 2025 ▶ 1:07:33 Better Data is All You Need — Ari Morcos, Datology
May 24, 2026 positive
Assertion Supported
Sanseviero: Smaller Gemma 4 models process audio and 30-60 second videos
“Multimodal wise, the smaller models can understand audio Images and short videos, so, 30 to 62nd videos and audios.”
Omar Sanseviero May 24, 2026 ▶ 6:43 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
May 24, 2026 neutral
Assertion Contradicted
Sanseviero: 31B is the largest quantized model fitting consumer GPUs
“The 31 is really like the largest model size that quantize would fit in a consumer GPU.”
Omar Sanseviero May 24, 2026 ▶ 17:17 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
May 24, 2026 neutral
Assertion Supported
Sanseviero: Gemini Nano on Pixel and Samsung phones is built on Gemma
“If you buy a Pixel phone or a high-end Samsung, they come with a Gemini Nano, and Gemini Nano is packed into the operating system, and Gemini Nano is really built on top of Gemma.”
Omar Sanseviero May 24, 2026 ▶ 2:13 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.