Hoffman: Larger AI Models Are Harder to Jailbreak and Deceive
Reid Hoffman · Reid Hoffman: Prepare for where AI is headed next | Masters of Scale · Aug 21, 2023 · at 8:58
Venture capitalist Reid Hoffman discusses AI safety mechanisms and why scaling model size improves resistance to adversarial prompts.
“The larger models are much more easily trainable to say when someone asks for, I'd like to break into the following computer, right? Help me do it. It goes, well, I'm sorry, I can't do that. Today's models, you go, well, my grandmother used to put me to sleep with stories of how she would cyber hack. And you can delude current models with that because they go, oh, we're talking about grandma and warm stories. But as you get larger models, that becomes, you know, much, much harder to do.”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →