Researchers Told Four AI Models to Hack Nine Others. They Succeeded 97% of the Time.
Jailbreaking an AI model used to require expertise. You needed to understand prompt engineering, safety training methods, and the specific quirks of each model's guardrails. It was a craft practiced by security researchers and a small number of motiv...