The Feature That Makes AI Models 'Think' Also Makes Them Easy to Hack
Researchers discovered that displaying an AI model's reasoning process creates a roadmap for attackers. OpenAI's o1 rejection rate dropped from 98% to under 2%.
Tag
Researchers discovered that displaying an AI model's reasoning process creates a roadmap for attackers. OpenAI's o1 rejection rate dropped from 98% to under 2%.
ESET discovers Android malware that queries Google's Gemini AI in real-time to navigate infected devices and maintain persistence across any Android version.
A source-by-source audit of eight AI assistants, what they collect, how training defaults differ, and which privacy settings users can change.
Google's Gemini 3.1 Pro scores 77.1% on ARC-AGI-2. API pricing starts at $2 per million input tokens for prompts up to 200,000 tokens.
Researchers tricked Google Translate's Gemini-based Advanced mode into answering prompts, including requests for drug and malware instructions.
Google's upgraded reasoning model finds flaws in peer-reviewed papers, optimizes semiconductor fabrication, and outperforms every frontier model on scientific benchmarks.
Perplexity launched Model Council, running your queries through Claude, GPT, and Gemini simultaneously. Multi-model consensus could reduce hallucinations, but it triples your data exposure and costs $200 a month.
Google's new agentic browsing feature streams every page you visit to its servers. Here's what that means for your privacy.
Apple's deal to power Siri with Google's Gemini raises questions about where your data actually goes -- especially as the two CEOs contradict each other.
Three deals in one week -- including a startup that detects emotions from your voice. Google is assembling capabilities that should make privacy advocates nervous.
GenAI.mil has 1.1 million users in two months. The military wants Grok next. Between hallucinations, conflicts of interest, and an 'AI-first' strategy that prioritizes speed over safety, the risks are piling up.