Fireworks AI, every mention

30 scenes · ← back to Fireworks AI

tap a year for its mentions
002545082023202420252026episodesmentions
0482023202420252026episodes it came up in
0044882023202420252026episodesmentions per episode

every year anyone Shawn Wang 44Alessio Fanelli 5Lin Qiao 3Mikhail Parakhin 2Beyang Liu 2Sarah Sachs 1

Verbatim, from the transcripts: the passages where Fireworks AI comes up

loading…

Next 100x in AI: Inference, Networking, & Self-Optimizing Models — Philip Kiely & Ali Taha, Baseten Aug 3, 2026 · 1 mention

  • ▶ 12:53 Shawn Wang Ok, so, like, you know, a lot of people, all you guys, right, whenever a new model launch, like, people rush to say, like, oh, Hugging Face supports this, Fireworks supports this, Base 10 supports this, and I'm like, yeah, of course you…

The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO Jul 8, 2026 · 1 mention

  • ▶ 14:39 unnamed speaker There's a bunch of inference providers, which, you know, provide this fireworks, does this as a service together or whatnot.

AI-Native Engineering: 100% adoption, 5x search throughput, unlimited tokens — Mikhail Parakhin Apr 22, 2026 · 2 mentions

  • ▶ 45:59 Mikhail Parakhin I work with, uh, uh, Fireworks and Sentinel, uh, you know, to help with optimizations and browser base, as you mentioned. 2 times in the scene

Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work Apr 15, 2026 · 1 mention

  • ▶ 3:21 Sarah Sachs Like before function calling came out, we were trying to fine tune with the frontier labs and with fireworks, like a function calling model on notion functions.

Artificial Analysis: The Independent LLM Analysis House — with George Cameron and Micah Hill-Smith Jan 9, 2026 · 2 mentions

The Future of Email: Superhuman CTO on Your Inbox As the Real AI Agent (Not ChatGPT) — Loïc Houssier Dec 11, 2025 · 2 mentions

  • ▶ 32:55 Shawn Wang The inference provider for open models compared to, let's say the fireworks and the together AIs. 2 times in the scene

Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave) Oct 16, 2025 · 1 mention

⚡️ Beyond Transformers with Power Retention Sep 23, 2025 · 1 mention

  • ▶ 24:54 unnamed speaker So if you are together or fireworks or any of these like inference providers, should you just be doing this for every single model?

The AI Agenda: GPT5 leaks and the business of AI News — Steph Palazzolo, The Information Aug 6, 2025 · 2 mentions

  • ▶ 8:03 Shawn Wang I saw the other day fireworks is raising at four billion. 2 times in the scene

The Shape of Compute (Chris Lattner of Modular) Jun 13, 2025 · 1 mention

  • ▶ 36:34 Shawn Wang And I think a lot of the other inference team, because effectively every team is, is a startup, like the fireworks together, you know, all those guys, your business model is very different from them.

DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing) Jan 19, 2025 · 1 mention

  • ▶ 8:43 unnamed speaker I'll let Yeneng answer the, the, the sort of patterns around, you know, uh, using FPA in training, but I want to draw one, one distinction that, that I think gets to the, the latter part of your question, Sean, which is that unlike…

2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents Jan 1, 2025 · 3 mentions

  • ▶ 41:03 Shawn Wang We have now interviewed both together and fireworks and replicates.
  • ▶ 59:25 Shawn Wang Like Canvas has incorporated the diff mode that both Anthropic and OpenAI and Fireworks has now shipped that I think is going to be the norm for next year, that everyone
  • ▶ 1:36:55 Shawn Wang Um, everyone knew like F-one, we had a preview at the Fireworks HQ, and then I think some, some other labs did it, but I think R-one and Q-W-Q, Quill, um, from the Quint team,

The new Claude 3.5 Sonnet, Computer Use, and Building SOTA Agents — with Erik Schluntz, Anthropic Nov 28, 2024 · 1 mention

Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI Nov 25, 2024 · 38 mentions

  • ▶ 0:12 Shawn Wang Hey, and today we're in a very special studio inside the Fireworks office with Lin Tian, CEO of Fireworks. 4 times in the scene
  • ▶ 2:08 Shawn Wang So let's, I guess, let's start at the prehistory, like the pre, the sort of, uh, the initial history of fireworks. 2 times in the scene
  • ▶ 6:59 Shawn Wang When you and I chatted about like the origins of fireworks, it was originally envisioned more as a PyTorch platform, and then later became much more focused on generative AI. 6 times in the scene
  • ▶ 20:30 Alessio Fanelli Because when I hear your explanation, it's almost like you're centralizing a lot of the decisions through the Fireworks platform on like the quality and whatnot. 3 times in the scene
  • ▶ 38:43 Shawn Wang First of all, I realized that, I don't know if I've ever given you this feedback, but I think you guys are one of, like, one of the reasons I agreed to advise you, because, like, you know, I think when, when you first met me, I was kind of… 10 times in the scene
  • ▶ 46:31 Shawn Wang What is your take on what happens, and maybe you want to set the record straight on how Fireworks does quantization, because I think a lot of people may have outdated perceptions, or they didn't read the clarification post on your approach… 10 times in the scene
  • ▶ 54:23 Lin Qiao Yeah, so we really want to get a lot of feedback from the application developers who are starting to build on GNI or, you know, on the already adopted or starting about thinking about new use cases and so on to try out on Fireworks first… 3 times in the scene

Production AI Engineering starts with Evals Oct 11, 2024 · 1 mention

  • ▶ 1:37:21 unnamed speaker You can use Together, Fireworks, all these guys.

The Winds of AI Winter (Q2 Four Wars of the AI Stack Recap) Aug 2, 2024 · 2 mentions

  • ▶ 11:23 Alessio Fanelli I also don't know what the hosting up hosting options are as far as like scaling, you know, I don't know if the fireworks and togethers of the world, how much capacity they actually have to serve this model, because at the end of the day,…
  • ▶ 25:10 unnamed speaker No one thinks that he can do inference at 13 X cheaper than the fireworks together, right?

[LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models Jul 29, 2024 · 2 mentions

  • ▶ 12:51 unnamed speaker Um, you know, fireworks is somehow really undercutting inference price.
  • ▶ 1:05:43 unnamed speaker I'm writing an email now between cloud providers for 3.1 70 B just to see if like, you know, together's, uh, the eight versus, I don't know, fireworks or something like this has a difference or versus Glock.

Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI Mar 6, 2024 · 3 mentions

  • ▶ 19:51 unnamed speaker I know you have Fireworks AI, 3 times in the scene

The Four Wars of the AI Stack - Dec 2023 Recap Jan 26, 2024 · 1 mention

  • ▶ 32:24 unnamed speaker Like, uh, Fireworks, um, recently announced, uh, Fire Attention, where they wrote a custom cruder kernel for mixed drawl, uh, on H 100.

The "Normsky" architecture for AI coding agents — with Beyang Liu + Steve Yegge of SourceGraph Dec 17, 2023 · 2 mentions

Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.