Gemini Flash, every mention
19 scenes, the whole family · ← back to Gemini Flash
tap a year for its mentions
every year anyone Shawn Wang 10Logan Kilpatrick 4Jeff Dean 4Alessio Fanelli 4Rishabh Agarwal 3Pratik Bhavsar 1Olivia Watkins 1Dhravya Shah 1Dharmesh Shah 1Ben Holmes 1
Verbatim, from the transcripts: the passages where Gemini Flash comes up
⚡️ OpenClaw's Memory Sucks and the fix is simple — Dhravya Shah, Supermemory
- ▶ 18:16 Dhravya Shah And like in Gemini Flash or whatever, like you can just literally, the, the questions, the sessions themselves are short enough that you can
The End of SWE-Bench Verified — Mia Glaese & Olivia Watkins, OpenAI Frontier Evals
- ▶ 11:54 Olivia Watkins And in SweetBenchVerified, we found many instances of contamination across, um, like, across OpenEye models, across, like, Quad, uh, Opus, 4.5, Gemini Flash, and, and all of these, we saw things like regurgitating the ground truth…
The AI Frontier: from Gemini 3 Deep Think distilling to Flash — Jeff Dean
- ▶ 6:35 Jeff Dean Close, very close to your largest model performance with distillation approaches, and that, that seems to be, you know, a nice sweet spot for a lot of people because it enables us to kind of, for multiple Gemini generations now, we've been… 4 times in the scene
- ▶ 7:37 Shawn Wang And then obviously I think the economy of flash is what's led to the total dominance. 6 times in the scene
- ▶ 9:21 Alessio Fanelli Does it feel like there's some breaking point for like the proto flash distillation, kind of like one generation delayed? 4 times in the scene
- ▶ 1:19:49 Shawn Wang And the multi-turn taking with a flash model is enough.
⚡ Open Model Pretraining Masterclass — Elie Bakouch, HuggingFace SmolLM 3, FineWeb, FinePDF
- ▶ 33:49 Shawn Wang I would say, from what I've heard, Gemini Flash is also pretty, pretty strong around, around this.
Long Live Context Engineering - with Jeff Huber of Chroma
- ▶ 17:57 Shawn Wang And then GPT four one and Gemini flash are, are, uh, degrade a lot quicker in terms of the context length.
⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
- ▶ 7:35 Pratik Bhavsar and so it became kind of a no-brainer that if anybody wants top performance, they can just directly go and go with the Gemini models, and they just not only released like the pro model, they also released the flash version later on and…
⚡️Warp 2.0: the Agentic Development Environment - Zach Lloyd and Ben Holmes
- ▶ 39:38 Ben Holmes And we also have the benefit of having access to every model, so I know you can fall back to Gemini Flash and use that for simple asks, and you don't, uh, have to worry about overages in those cases.
The #1 SWE-Bench Verified Agent
- ▶ 29:38 Shawn Wang You already have the, one of the fastest, cheapest models with Flash.
The Agent Network — Dharmesh Shah, Agent.ai + CTO of HubSpot
- ▶ 1:33:39 Dharmesh Shah By the way, in terms of like the coolest thing AI wise, recently, I'll say last a week to 10 days has been the new, um, image model, Gemini flash experimental, whatever they call it.
The Magic of LLM Distillation — Rishabh Agarwal, Google DeepMind
- ▶ 18:27 Rishabh Agarwal So if you go down to the slide, uh, it can be, yeah, if you do cost match, I was, so far I was talking about compute match, but if you think about API pricing, like, let's say we can, we have two models, Gemini Pro and Gemini Flash, the… 3 times in the scene
Gemini 2.0 Flash and Flash Thinking: the new SOTA models for the agentic era
- ▶ 1:44 unnamed speaker I think one of the interesting things that people are trying to figure out is the pricing strategy, where pro is, and then what flash thinking is and why there, there's like flashlight now. 2 times in the scene
- ▶ 2:04 Logan Kilpatrick The sort of general principle is no flash. 4 times in the scene
- ▶ 12:43 unnamed speaker And the moment Gemini Flash came out, I ran it against the Mini as well. 4 times in the scene
- ▶ 15:12 unnamed speaker I know there's like not a ton that's public about flash thinking, but like, you know, is deep seek the right path or they're like, are you seeing multiple paths going on? 3 times in the scene
OpenAI o1 isn’t a chat model (and that’s the point)
- ▶ 20:55 Ben Hillock I think flash sort of like became so much faster that I would then start using kind of both, but it feels like with
In the Arena: How LMSys changed LLM Benchmarking Forever
- ▶ 28:03 Anastasios Angelopoulos Absolutely, but I mean, to be clear, none of the leaderboard currently is apples to apples, because you have, like, Gemini Flash, you have, you know, all sorts of tiny models, like Llama