Caching

topic on 1 show · 2 statements across 2 episodes

Latent Space

2 statements about Caching, every show

Swix: Proprietary LLM caching creates vendor lock-in
“I feel like this is definitely a form of lock-in because you ideally want to be able to run prompts across multiple providers and all that. And yeah, caching is a hard problem. Like, I think ultimately, like, you control your destiny if you can run your own op…”
Shawn Wang Sep 11, 2025 ▶ 38:30 Context Engineering for Agents - Lance Martin, LangChain
Embeddings will not scale for cross-session memory in real-time AI systems
“I don't think that like embeddings are going to be able to scale to, I think they work well for some of this as like kind of the MVP version of the experience, but I think you're going to need a different experience and it's Probably something like really smar…”
Logan Kilpatrick Feb 28, 2025 ▶ 23:41 Gemini 2.0 Flash and Flash Thinking: the new SOTA models for the agentic era

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.