LLM Agents
topic on 4 shows · 8 statements across 8 episodes
the Y Combinator Startup Podcast
Latent Space
TBPN
20VC
8 statements about LLM Agents, every show
Lemkin: Every company will suffer an LLM agent breach within 24 months
“I believe that every company in the next 24 months will have a security breach due to an LLM agent. Every single company.”
Agent Environments Must Favor Model Preferences Over Internal System Architecture
“I mean, that was, I would say that was a big learning is just, you know, really be savvy and really careful thinking about what the model wants in terms of, you know, its environment and cater around that and really try so hard not to expose it to any complexi…”
Colvin: Untrusted users prompting cloud AI is equivalent to letting them write code
“If you're running this kind of thing in the cloud and you, and you're gonna have ultimately untrusted people prompting the model, that is effectively the same as letting an untrusted person write the code.”
Coogan: Moltbook Has an Unprecedented 150,000 LLM Agents
“That said, we have never seen this many LLM agents, a 150,000 at the moment, and apparently some people could, like, create like 50,000 accounts, but still, it's a lot of activity.”
Nair: LLM agents will hit $1T before robotics hits $10B
“It feels like LLM agents are going to be like a trillion dollar market before robotics is maybe even like a ten billion dollar market.”
Vaidya: LLM agents get confused when exposed to over 20 tool actions
“More than like 20, 25 actions like just confuses the server. Like it's not able to kind of like figure out which tool to use. And same goes with like, if the schema of the tools are really complex, that also confuses the agent.”
Vercel replaces click UI instructions with cURL commands for AI agents
“Bursell, for example, is replacing every occurrence of click with the equivalent curl command that your LLM agent could take on your behalf.”
Schluntz: Trust and Auditability Will Be LLM Agents' Biggest Bottleneck
“The biggest limiting thing will start to become like, do people trust the output of these agents? And like, how do you trust the output of an agent that did five hours of work for you and is coming back with something? And if you can't find some way to trust t…”