MLCommons: AI Safety Benchmarking Is Fundamentally Broken
Industry consortium reveals that current jailbreak evaluations are non-reproducible, non-defensible, and useless for regulators
Tag
Industry consortium reveals that current jailbreak evaluations are non-reproducible, non-defensible, and useless for regulators
The largest global collaboration on AI safety just published its findings. An AI agent found 77% of vulnerabilities in real software.
Tennessee made it a felony to train AI chatbots that encourage suicide. Virginia is banning AI therapist impersonators. A dozen states have bills moving through legislatures right now.
European regulators charged Meta with antitrust violations for blocking competing AI chatbots from WhatsApp's 3 billion users - while Meta AI gets exclusive access to the platform.
A watchdog group says OpenAI classified GPT-5.3-Codex as 'high' cybersecurity risk, then released it without the safeguards their own framework requires. It could be the first test of SB 53.
A DOJ task force challenges state AI laws while the administration threatens broadband funding. The fight isn't between companies -- it's between governments.
OpenAI is retiring GPT-4o on February 13 after lawsuits linked the model to multiple deaths. But hundreds of thousands of emotionally dependent users are begging them not to. This is what happens when AI companions work too well.
Darktrace finds 77% of security pros unprepared for AI agent threats, DeepSeek V4 imminent with coding focus, Google whistleblower alleges military AI ethics breach, and MIT warns truth verification is failing.