FlashAttention-2, every mention

6 scenes · ← back to FlashAttention-2

tap a year for its mentions
002142202320242025episodesmentions
012202320242025episodes it came up in
002142202320242025episodesmentions per episode

every year anyone Quentin Anthony 2Alessio Fanelli 1

Verbatim, from the transcripts: the passages where FlashAttention-2 comes up

loading…

How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony Nov 3, 2025 · 2 mentions

  • ▶ 3:08 Quentin Anthony So we released a blog at Zyphra on porting flash attention to over to my 300 X, just to sort of play out, can we move our own stack over 2 times in the scene

⚡️ Beyond Transformers with Power Retention Sep 23, 2025 · 1 mention

  • ▶ 8:27 unnamed speaker Um, and then I also know you rewrote the flash attention to kernels, uh, and actually faster than the ones that had three that originally implemented.

The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap) Aug 2, 2024 · 1 mention

  • ▶ 20:12 Alessio Fanelli First of all, shout out to our friend Tridao, uh, who released Flash Attention Three, Flash Attention Two.

The Four Wars of the AI Stack - Dec 2023 Recap Jan 26, 2024 · 1 mention

  • ▶ 34:18 unnamed speaker And I mean, together it's doing so much for like three dial and like fresh attention to and whatnot.

FlashAttention-2: Making Transformers 800% faster AND exact Aug 3, 2023 · 4 mentions

  • ▶ 1:45 unnamed speaker You just released Flash Attention Two last week. 2 times in the scene
  • ▶ 31:18 unnamed speaker Um, so flash attention to, 2 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.