Hu: Voice AI Startups Meta-Prompt with Frontier Models Before Distillation
Diana Hu · State-Of-The-Art Prompting For AI Agents · Y Combinator · May 30, 2025 · at 11:12
Y Combinator partner Diana Hu explains how AI agent startups optimize latency by refining prompts on frontier models before deploying on smaller models.
“I think that's a common pattern sometimes for companies when they need to get responses from elements, elements in their product a lot quicker. They do the meta-prompting with a bigger, beefier model, any of the, I don't know, hundreds of billions of parameters plus models like I guess, cloud four, 3.7, or your GPT-O three, and they do this meta-prompting, and then they have a very good working one that Then they use into the distilled model, so they use it on, ah, for example, an FRO, and it ends up working pretty well, specifically sometimes for, ah, voice AI agents, companies, because, ah, latency is very important to, ah, get this whole Turing test to pass, because if you have too much pause before, before the agent responds, I think humans can detect something is off, so they use a faster model”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →