Tri Dao, every mention

13 scenes · ← back to Tri Dao

tap a year for its mentions
0043852023202420252026episodesmentions
0352023202420252026episodes it came up in
0012.5252023202420252026episodesmentions per episode

every year anyone Alessio Fanelli 6Ali Taha 2Quentin Anthony 1Diego Bachman 1Chris Lattner 1

Verbatim, from the transcripts: the passages where Tri Dao comes up

loading…

Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten Aug 3, 2026 · 2 mentions

  • ▶ 45:21 Ali Taha It's, it's, it's a paper by Tridao, and it's like, it's basically doing speculative decoding. 2 times in the scene

How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony Nov 3, 2025 · 1 mention

  • ▶ 14:16 Quentin Anthony You know, tree dial is like one example of making very public kernels, but how many more do exist?

Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21 Oct 11, 2025 · 1 mention

  • ▶ 4:55 unnamed speaker Like, I think like researching hybrid LLMs is something that basically only Tridao and Albert Gu were kind of like looking into.

⚡️ Beyond Transformers with Power Retention Sep 23, 2025 · 1 mention

  • ▶ 27:20 Diego Bachman There's things like flash attention from TreeDAO, which produced these massive, massive improvements in performance.

The Shape of Compute (Chris Lattner of Modular) Jun 13, 2025 · 1 mention

  • ▶ 30:59 Chris Lattner We're beating the tree DAO reference implementation that everybody uses, for example, right?

2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents Jan 1, 2025 · 1 mention

Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph Sep 27, 2024 · 1 mention

  • ▶ 6:19 Alessio Fanelli We had Trida on the podcast that you mentioned, uh, he was really inspired working with like systems people to think about flash attention.

Answer.ai & AI Magic with Jeremy Howard Aug 17, 2024 · 1 mention

  • ▶ 43:11 unnamed speaker We had a three DAO on the podcast and, uh, we talked about how a lot of it is like systems work to make some of these things work.

The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap) Aug 2, 2024 · 2 mentions

  • ▶ 20:12 Alessio Fanelli First of all, shout out to our friend Tridao, uh, who released Flash Attention Three, Flash Attention Two. 2 times in the scene

Building an open AI company - with Ce and Vipul of Together AI Feb 8, 2024 · 2 mentions

The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert Jan 11, 2024 · 1 mention

  • ▶ 13:46 unnamed speaker Because when we had three DAO on the pockets, it's a flash of tension came to be because at AZ, they have so much overlap between systems engineering, like, uh, deep learning engineers.

The End of Finetuning — with Jeremy Howard of Fast.ai Oct 20, 2023 · 2 mentions

  • ▶ 1:15:32 unnamed speaker As we had, um, three DAO on the podcast who created Flash Attention. 2 times in the scene
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.