Tri Dao, every mention
13 scenes · ← back to Tri Dao
tap a year for its mentions
every year anyone Alessio Fanelli 6Ali Taha 2Quentin Anthony 1Diego Bachman 1Chris Lattner 1
Verbatim, from the transcripts: the passages where Tri Dao comes up
Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten
How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
- ▶ 14:16 Quentin Anthony You know, tree dial is like one example of making very public kernels, but how many more do exist?
Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21
- ▶ 4:55 unnamed speaker Like, I think like researching hybrid LLMs is something that basically only Tridao and Albert Gu were kind of like looking into.
⚡️ Beyond Transformers with Power Retention
- ▶ 27:20 Diego Bachman There's things like flash attention from TreeDAO, which produced these massive, massive improvements in performance.
The Shape of Compute (Chris Lattner of Modular)
- ▶ 30:59 Chris Lattner We're beating the tree DAO reference implementation that everybody uses, for example, right?
2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
- ▶ 0:44 Alessio Fanelli You know, we had like Tridao, we had Jeremy Howard, we had more folks like that.
Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph
- ▶ 6:19 Alessio Fanelli We had Trida on the podcast that you mentioned, uh, he was really inspired working with like systems people to think about flash attention.
Answer.ai & AI Magic with Jeremy Howard
- ▶ 43:11 unnamed speaker We had a three DAO on the podcast and, uh, we talked about how a lot of it is like systems work to make some of these things work.
The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap)
- ▶ 20:12 Alessio Fanelli First of all, shout out to our friend Tridao, uh, who released Flash Attention Three, Flash Attention Two. 2 times in the scene
Building an open AI company - with Ce and Vipul of Together AI
- ▶ 5:45 Alessio Fanelli You know, we already had three DAO from together, and we talked about Hazy.
- ▶ 35:07 Alessio Fanelli Because we at ThreeDAO, obviously, flash attention to is open source.
The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
- ▶ 13:46 unnamed speaker Because when we had three DAO on the pockets, it's a flash of tension came to be because at AZ, they have so much overlap between systems engineering, like, uh, deep learning engineers.
The End of Finetuning — with Jeremy Howard of Fast.ai
- ▶ 1:15:32 unnamed speaker As we had, um, three DAO on the podcast who created Flash Attention. 2 times in the scene