Microsoft: North Korean Hackers Are Jailbreaking AI Models for Cyberattacks
Nation-state threat actors are operationalizing AI across the attack lifecycle, using jailbreak techniques to bypass safety controls
Tag
Nation-state threat actors are operationalizing AI across the attack lifecycle, using jailbreak techniques to bypass safety controls
A peer-reviewed study finds AI models can autonomously jailbreak other AI models with 97% success - and Claude was the only one that held the line.
New paper shows 'intent laundering' bypasses Gemini, Claude, and other models with 90-98% success by removing obvious attack cues