Gemma, every mention

69 scenes across 13 shows · ← back to Gemma

tap a year for its mentions
0040138025202420252026episodesmentions
01325202420252026episodes it came up in
00213425202420252026episodesmentions per episode

Latent Space 61the Startup Ideas Podcast 25the MAD Podcast 7the Y Combinator Startup Podcast 7TBPN 5Big Technology 420VC 4All-In 25 more shows

every year every show Latent Space 61 the Startup Ideas Podcast 25 the Y Combinator Startup Podcast 7 the MAD Podcast 7 TBPN 5 Big Technology 4 20VC 4 All-In 2 We Live to Build 1 American Optimist 1 the Neon Show 1 Sourcery 1 the a16z Podcast 1

Verbatim, from the transcripts: passages where Gemma comes up on Latent Space, the Startup Ideas Podcast, the Y Combinator Startup Podcast, the MAD Podcast, TBPN

loading…

I'm Obsessed With Local AI. Here's Why Sep 8, 2026 · 22 mentions

  • ▶ 0:29 Greg Isenberg By the end of today's episode, you're going to understand what local AI is, when it matters, how to run open models at work, where Hugging Face fits in here, which Gemma model I'd start with, how I'd run a model locally with LM Studio or… 2 times in the scene
  • ▶ 3:17 Greg Isenberg That could be something like Gemma, Llama, or Mistral.
  • ▶ 3:46 Greg Isenberg Gemma is a model family. 2 times in the scene
  • ▶ 9:56 Greg Isenberg Gemma is Google's open model family, and Google's a trusted brand.
  • ▶ 10:44 Greg Isenberg So, Gemma is Google's family of open models, and Gemma IV is built for the efficient, local, and on-device use. 6 times in the scene
  • ▶ 14:21 Greg Isenberg Beyond Google Gemma, I'll give you a quick primer on the other
  • ▶ 18:20 Greg Isenberg Like, if you actually want to run Gemma, here's how I would do it. 4 times in the scene
  • ▶ 23:22 Greg Isenberg Then I would run a local model like Gemma and ask it to create a file called what customers are telling us .md, the markdown file. 3 times in the scene
  • ▶ 35:07 Greg Isenberg It could be anything from sales calls or meeting transcripts, old tweets, ideas that you have, then run Gemma, whatever model you choose, 2 times in the scene

The Inference Frontier: from 100 to 10,000 tokens per second — Sean Lie, Cerebras CTO Sep 2, 2026 · 2 mentions

  • ▶ 11:25 Sean Lie And what this ultimately means is you'll be able to run, you know, medium sized models like GPT-OSS or JAMA at speeds up to 10,000 TPS.
  • ▶ 12:15 unnamed speaker Right now, sure, there's a Gemma OSS.

More Details on OpenAI-Hugging Face, SpaceX SPV Investors Get Rugged, Demis and DeepMind | Diet TBPN Aug 6, 2026 · 2 mentions

Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten Aug 3, 2026 · 3 mentions

Arena CEO: There Will be a $100BN US Open-Source Model & Data is a Trillion Dollar Market Aug 2, 2026 · 3 mentions

Multi-GPU Kernels, Intelligence per Watt, Heterogeneous Inference, and More | YC Paper Club · Y Combinator Jul 29, 2026 · 1 mention

  • ▶ 25:43 unnamed speaker We did about twenty-plus different state-of-the-art local models, um, across GEMMA, GBD OSS, QUAN, IBM Granite, all in the range of one to two hundred billion parameters, some MOE, some dense.

The Biggest AI Deployment Nobody Talks About | Samsara CEO Sanjit Biswas Jul 29, 2026 · 2 mentions

  • ▶ 44:45 Sanjit Biswas They have, uh, an inference model that you can run on their chip, which is basically Gemma four, but like hyper accelerated. 2 times in the scene

RSI Is Closer Than People Think, Per Tae Kim Jul 29, 2026 · 2 mentions

  • ▶ 9:31 unnamed speaker Yeah, it seems very reasonable that he would have no problem with, like, Gemma, or Lama, or any of the open source from, like, American hyperscalers, where if you find out that they're distilling, you just walk across the street and sue… 2 times in the scene

Cerebras CEO: Why GPUs Can't Do Fast Inference Jul 23, 2026 · 1 mention

  • ▶ 59:43 Matt Turck So if I want to run to me, like, I know you have like incredible stats for, uh, QE and Gemma in terms of speed, I forgot to mention them.

The AI Frontier: from open weights to open research — Eiso Kant, Poolside AI Jul 22, 2026 · 1 mention

  • ▶ 24:42 unnamed speaker We're very much like, ah, okay, look, it's like, you know, on par with Kimmy, DeepSeek, whatnot, the small ones, Gemma level.

Inside Nemotron & NVIDIA’s AI Lab | Bryan Catanzaro Jul 2, 2026 · 1 mention

  • ▶ 9:58 Bryan Catanzaro You know, I was really excited when, uh, OpenAI released the GPT OSS models, um, uh, a while back, and then of course Google's been doing great work with Gemma.

How AI Agents Are Creating an Infinitely Scalable Workforce · Joe Lonsdale Jun 18, 2026 · 1 mention

  • ▶ 31:37 Nima Ghamsari Yeah, and every new, I want to put an update to it, every new model that comes out, so that a new model comes out, let's say Gemma, you know, there's, there's this rumor that Google is going to release something at Google I.O. next year,…

Claude Fable 5 is BANNED. What to do? Jun 13, 2026 · 2 mentions

Inside xAI: Building Grok Imagine in 3 Months, Videogen vs World Models, and Video Agents— Ethan He Jun 1, 2026 · 1 mention

  • ▶ 33:51 unnamed speaker It's a Gemma level model trained on roughly 40 trillion tokens at this many H 200 over this much time, right?

⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind May 24, 2026 · 21 mentions

  • ▶ 2:13 Omar Sanseviero Yeah, so actually, if you install, like, if you buy a Pixel phone or a high-end Samsung, they come with a Gemini Nano, and Gemini Nano is packed into the operating system, and Gemini Nano is really built on top of Gemma.
  • ▶ 3:21 Omar Sanseviero The Gemma team is actually relatively small. 6 times in the scene
  • ▶ 6:35 Omar Sanseviero Yeah, so Gemma four was built on the same research as Gemini three, which pretty much means that we benefited from all of the improvements that happened with Gemini three. 6 times in the scene
  • ▶ 14:01 Omar Sanseviero So as I was saying, like for Gemma four, we had 50 5 times in the scene
  • ▶ 16:29 Alessio Fanelli Yeah, I have a question about the bigger Gemma models. 2 times in the scene
  • ▶ 29:18 Omar Sanseviero I mean, the way we are doing Gemma, Gemini, and all of our tools is really like based on the feedback from the startups, the community, the developers, that's why you see like Logan, Paige, everyone in the team talking with the community…

Demis Hassabis: Agents, AGI & The Next Big Scientific Breakthrough · Y Combinator Apr 29, 2026 · 6 mentions

  • ▶ 10:18 Demis Hassabis And you also see some of that goodness in our Gemma models, which hopefully you're all enjoying our Gemma four models, which I think are really amazing power for their sizes. 2 times in the scene
  • ▶ 20:24 Garry Tan I mean, the recent release of Gemma, you're making highly capable, open and accessible ones that can actually run locally. 4 times in the scene

Demis Hassabis: Why AGI is Bigger than the Industrial Revolution & Where Are The Bottlenecks in AI · 20VC with Harry Stebbings Apr 7, 2026 · 1 mention

  • ▶ 11:01 Demis Hassabis Um, but we are also, uh, pushing hard on a kind of suite of open source models called Gemma, which are, you know, we're determined to kind of make best in class for their sizes.

AI is Already Building AI — Google DeepMind’s Mostafa Dehghani Apr 2, 2026 · 1 mention

  • ▶ 33:08 Mostafa Dehghani Given that, like, I've, I've seen, like, very impressive progress on this, uh, of the Sweden Gemma and GDM.

Two Legendary Founders: Travis Kalanick & Michael Dell Live from Austin, Texas Mar 17, 2026 · 1 mention

  • ▶ 56:09 Michael Dell You know, Google has these Gemma models, G-E-M-M-A, and they work really, really well on small machines.

The AI Frontier: from Gemini 3 Deep Think distilling to Flash — Jeff Dean Feb 12, 2026 · 2 mentions

  • ▶ 50:07 Shawn Wang But, uh, the Gemma models, for example, right?
  • ▶ 54:51 Shawn Wang And for listeners, I think I will highlight the Gemma three end paper where they, there was a little bit of that, I think.

Goodfire AI’s Bet: Interpretability as the Next Frontier of Model Design — Myra Deng & Mark Bissell Feb 5, 2026 · 2 mentions

  • ▶ 46:10 unnamed speaker Um, DeepMind has opened a lot of essays on, um, Gemma.
  • ▶ 49:43 Mark Bissell Robotics, I know, like, a lot of the companies just use Gemma as, like, the, the, like, backbone, and then they, like, make it into a VLA that, like, takes these actions.

Erotic ChatGPT, Zuck’s Apple Assault, AI’s Sameness Problem Oct 20, 2025 · 1 mention

  • ▶ 29:54 Alex Kantrowitz It's built on Google's open source gamma family of models, and the model was able to generate a novel hypothesis about cancer cellular behavior

⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF Oct 20, 2025 · 1 mention

Google Makes MAJOR Cancer Research Breakthrough | Diet TBPN Oct 17, 2025 · 1 mention

  • ▶ 1:16 unnamed speaker He said, an exciting milestone for AI and science, our C to S scale, 27 B foundation model built with Yale and based on Gemma generated a novel hypothesis about cancer cellular behavior, which scientists experimentally validated in living…

A Technical History of Generative Media Sep 8, 2025 · 1 mention

  • ▶ 38:20 unnamed speaker Then you got Gemma, two hundred seventy million.

Better Data is All You Need — Ari Morcos, Datology Aug 29, 2025 · 2 mentions

  • ▶ 1:07:06 Alessio Fanelli And this is a 4.5 B model, which is par with Gemma four B and a little worse than Quan three, but roughly the same. 2 times in the scene

Dylan Patel on GPT-5’s Router Moment, GPUs vs TPUs, Monetization Aug 18, 2025 · 1 mention

Information Theory for Language Models: Jack Morris Jul 2, 2025 · 4 mentions

  • ▶ 48:55 Shawn Wang Yeah, I would say, uh, okay, I pulled out something very current, uh, which is Gemma three N, which launched, which, uh, sort of was, uh, generally available yesterday. 2 times in the scene
  • ▶ 1:01:58 Jack Morris So like you were just mentioning Gemma three B came out yesterday and you can download it and it takes up a certain amount of space on disk 2 times in the scene

The Utility of Interpretability — Emmanuel Amiesen Jun 6, 2025 · 13 mentions

  • ▶ 2:37 Vibhu (Viboo) What, why should we probe Gemma, Lama? 4 times in the scene
  • ▶ 7:54 Emmanuel Ameisen And so if you say like, thanks for having me on the whatever, like Gemma seems to have pretty consistently guessed that you're like on a podcast, uh, which makes sense, right? 2 times in the scene
  • ▶ 22:35 Vibhu (Viboo) So if I have a base Gemma and I have a chat model, what are differences in their attributions, right? 4 times in the scene
  • ▶ 43:10 unnamed speaker He did Neuronpedia and released a bunch of SAEs for, I think, the Llama models and the Gemma models. 2 times in the scene
  • ▶ 1:27:29 Emmanuel Ameisen There's some of the JAMA models.

Google DeepMind CTO: Advancing AI Frontier, New Reasoning Methods, Video Generation’s Potential May 24, 2025 · 2 mentions

  • ▶ 28:34 Koray Kavukcuoglu Nowadays, I think, like, the other thing that we always, like, need to remember is, actually we have, at Google, we have our Gemma models, right? 2 times in the scene

Sergey Brin, Google Co-Founder | All-In Live from Miami May 20, 2025 · 1 mention

Jeremy Howard on Building 5,000 AI Products with 14 People (Answer AI Deep-Dive) May 15, 2025 · 2 mentions

  • ▶ 4:50 Matt Turck But for open source Gemma, like you found that interesting. 2 times in the scene

Google Cloud CEO Thomas Kurian on AI Competition, Agents, And Tariffs Apr 9, 2025 · 1 mention

  • ▶ 27:25 Thomas Kurian You know, an example of our own model, we put out an open source model called Gemma, which is getting a lot of adoption among the developer community for people wanting to build certain class of applications.

It Was Printing Money But Every Buyer Walked Away Apr 8, 2025 · 1 mention

How Accel Deploys $650M in India & Manages $5B in Growth + Late Stage | Prayank Swaroop Apr 4, 2025 · 1 mention

The Magic of LLM Distillation — Rishabh Agarwal, Google DeepMind Mar 23, 2025 · 1 mention

Gemini 2.0 Flash and Flash Thinking: the new SOTA models for the agentic era Feb 28, 2025 · 1 mention

  • ▶ 27:51 unnamed speaker But also, I mean, Google came through with, like, some other presenters, and you also had Kathleen from the Gemma team, and I think people are very excited about open models still.

Why is everyone cloning Deep Research? Feb 18, 2025 · 1 mention

Google AI studio replaces your AI tech stack (full demo) Feb 15, 2025 · 1 mention

  • ▶ 7:20 Logan Kilpatrick Uh, there's some Gemma, which is our open source version of the models for folks who need open source models.

Inside Theory Ventures: Tomasz Tunguz’s $688M Thesis on Go-To-Market Disruption · Sourcery with Molly O'Shea Jan 23, 2025 · 1 mention

  • ▶ 11:45 Tomasz Tunguz There's a big push now to make the smaller models much more accurate, which they are like the five, five, four, five, four model from Microsoft and the Gemma models from Google are incredibly small, incredibly potent.

2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents Jan 1, 2025 · 1 mention

Best of 2024 in Vision [LS Live @ NeurIPS] Dec 22, 2024 · 1 mention

The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap) Aug 2, 2024 · 2 mentions

  • ▶ 15:51 unnamed speaker I would say also similar for Gemma, Gemma one and two, uh, Gemma two
  • ▶ 49:31 unnamed speaker The last piece I had, which I kind of deleted, was, uh, there's a special mention, honorable mention of Gemma again, with PolyGemma, which is one of the smaller releases from Google I.O.

LLM Asia Paper Club Survey Round May 22, 2024 · 1 mention

  • ▶ 15:06 unnamed speaker So before we go into like the paper itself, like actually, why does this matter?
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.