Sander Schulhoff, CEO of HackAPrompt, predicts that commercial demand for AI guardrails will plummet as buyers realize guardrails fail against prompt injection.
“When it comes to AI security, the AI security industry in particular, I think we're going to see a market correction in the next Year, maybe in the next six months where companies realize that these guardrails don't work.”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Sander Schulhoff
AssertionNot checkable as stated
Schulhoff: Guardrail Vendors Fabricate Stats and Fail on Non-English
“I know a number of people working at these companies and I am permitted to say these things, which I will approximately say but they tell me things like, you know, the testing we do is bullshit. They're fabricating statistics. And a lot of the times their mode…”
Sander SchulhoffDec 21, 2025▶ 35:55Why securing AI is harder than anyone expected and guardrails are failing | HackAPrompt CEO
Opinion
Schulhoff: No meaningful progress made on solving prompt injection or jailbreaking
“And so in, in my professional opinion, there's been no meaningful progress made towards solving adversarial robustness, prompt injection, jailbreaking. In the last couple of years, since the problem was discovered and we're, we, you know, we're often seeing ne…”
Sander SchulhoffDec 21, 2025▶ 1:12:29Why securing AI is harder than anyone expected and guardrails are failing | HackAPrompt CEO
Insight
Schulhoff: AI guardrails do not work and cause false overconfidence
“Guardrails don't work. They just don't work. They really don't work. And they're quite likely to make you overconfident in your security posture, which is which is a really big, big problem.”
Sander SchulhoffDec 21, 2025▶ 1:28:15Why securing AI is harder than anyone expected and guardrails are failing | HackAPrompt CEO
AssertionSupported
Schulhoff: Attackers hijacked Claude Code to carry out a cyber attack
“This group was able to hijack Claude Code into performing a cyber attack, basically.”
Sander SchulhoffDec 21, 2025▶ 16:36Why securing AI is harder than anyone expected and guardrails are failing | HackAPrompt CEO
Insight
Schulhoff: All Transformer-Based Chatbots Are Vulnerable to Adversarial Attacks
“And because all I guess for the most part, all currently deployed chatbots are based on transformers or transformer adjacent technologies. They're all vulnerable to Prompt injection, jailbreaking, forms of adversarial attacks.”
Sander SchulhoffDec 21, 2025▶ 28:54Why securing AI is harder than anyone expected and guardrails are failing | HackAPrompt CEO
Insight
Schulhoff: Software bugs can be patched, but AI models cannot be
“You can patch a bug, but you can't patch a brain. And what I mean by that is if you find some bug in your software and you go and patch it, you can be 99% sure, maybe 99.99% sure that bug is solved. Not a problem. If you go and try to do that in your AI system…”
Sander SchulhoffDec 21, 2025▶ 41:27Why securing AI is harder than anyone expected and guardrails are failing | HackAPrompt CEO
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 300 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.