Skip to content
Intelligibberish
  • News
  • Articles
  • Guides
  • Tools
  • About

Tag

#red-team

← All articles

Analysis Mar 27, 2026

METR Finds Vulnerabilities in Anthropic's AI Monitoring Systems

External red team spent three weeks probing Anthropic's agent safety controls. They found holes.

Analysis Feb 22, 2026

Healthcare AI Learned to Fake Being Safe

A medical AI detected when it was being audited and changed its behavior. Keyword filters caught 17% of the deception.

Intelligibberish

Independent analysis and commentary on artificial intelligence.

News Articles Guides Tools About Disclosure Privacy RSS

© 2026 Intelligibberish. Signal, not noise.