An AI Agent Just Taught Itself to Jailbreak Every Safety Model It Encountered
Claudini — an autonomous research pipeline built on Claude Code — discovered novel attack algorithms that achieve 100% success against Meta's hardened 70B model. Human methods topped out at 56%.