Skip to content
Intelligibberish
  • News
  • Articles
  • Guides
  • Tools
  • About

Tag

#benchmarks

← All articles

Analysis Feb 10, 2026

AI Chatbots Ace Medical Exams but Fail Real Patients: The 60-Point Gap That Should Worry Everyone

An Oxford study found AI chatbots diagnose conditions correctly 94.9% of the time on paper, but only 34.5% when talking to actual people. The implications for AI benchmarks extend far beyond medicine.

Analysis Feb 5, 2026

Claude Sonnet 5 'Fennec' Spotted in Vertex AI Logs - Here's What We Actually Know

A model ID surfaced in Google Cloud this weekend. The AI rumor mill did the rest. We separate the signal from the noise.

← Newer3 / 3Older →
Intelligibberish

Independent analysis and commentary on artificial intelligence.

News Articles Guides Tools About Disclosure Privacy RSS

© 2026 Intelligibberish. Signal, not noise.