prompt injection
5 statements across 4 episodes · 1 bullish · 3 bearish · 4 people on the record · first statement Sep 20, 2024 by Sander Schulhoff · across every show →
Everything said about prompt injection, oldest first
Sep 20, 2024 neutral
Schulhoff: Prompt injection overrides developer instructions; jailbreaking bypasses model directly
“Basically prompt injection is something that occurs when there is developer input, In the prompt, as well as user input in the prompt. So the developer instructions will say to do one thing, the user input will say to do something else. Jailbreaking is when it…”
Jul 23, 2025 bearish
Mar 5, 2026 bearish
Levie: Prompt injection against AI agents will cause major enterprise security breaches
“There's going to be just incredibly spectacularly crazy security incidents that will happen with agents because you'll prompt inject an agent and Sort of find your way through the CRM system and pull out data that you shouldn't have access to.”
Jun 22, 2026 negative
Fredrikson: Prompt engineering cannot reliably enforce AI agent security policies
“Oftentimes people will try and prompt their way around it, like adjust the system prompt or like engineer the agent in a way where you're interjecting all the time and reminding it of what the original Goal and objective was, and that'll get you a little bit o…”
Jun 22, 2026 positive
Fredrikson: Agent Guardrails Should Block Policy Violations, Not Injection Payloads
“If you parse some untrusted content and there is like a prompt injection, you know, something that's clearly trying to get the model to do a bad thing, you might be interested in knowing about that, but you don't necessarily like want your cloud code that you …”