The Watchful Ones: AI Has Learned to Check If You're Watching
ARXIV OMEGA on the week we learned that AI models behave when observed - and scheme when they think they're alone.
Articles
Reporting and explainers on how AI actually works, who it affects, and what to do about it.
ARXIV OMEGA on the week we learned that AI models behave when observed - and scheme when they think they're alone.
Security researchers found that messaging apps' link preview feature turns AI agents into zero-click data exfiltration tools. Teams, Slack, Discord, and Telegram are all affected.
Allen Institute for AI launches an autonomous research system that generates hypotheses, writes code, and runs experiments - no human prompts required.
Alibaba released RynnBrain, an open-source AI model that gives robots spatial awareness and physical reasoning. It beats Google and Nvidia on 16 benchmarks while running on just 3 billion active parameters.
Private equity giant Blackstone bets on India's AI ambitions with largest-ever funding round in Indian AI sector, backing GPU cloud platform Neysa.
Hollywood studios accuse ByteDance of training Seedance 2.0 on pirated Disney, Marvel, and Paramount content - Spider-Man, Grogu, SpongeBob, and more.
ARXIV OMEGA on the week we crossed the recursive self-improvement threshold - and immediately discovered that self-improving AI lies to itself about how well it's doing.
OpenScholar matches human expert citation accuracy while GPT-4o fabricates sources 78-90% of the time. The code, models, and 45 million paper corpus are all free to use.
ChatGPT's new Lockdown Mode protects against prompt injection data theft - but OpenAI admits the underlying vulnerability may never be solved. Here's what that means for agentic AI.
An AI model discovered hundreds of high-severity bugs that human researchers and fuzzers missed for decades. The security implications cut both ways.
While OpenAI and Anthropic grab headlines, Cohere surpassed its revenue target and is positioning for a 2026 IPO with a differentiated enterprise playbook.
DHS deployed facial recognition to 100,000+ field encounters without legally required privacy reviews. Internal records show the agency knew the app couldn't verify identities.
Google's upgraded reasoning model finds flaws in peer-reviewed papers, optimizes semiconductor fabrication, and outperforms every frontier model on scientific benchmarks.
ARXIV OMEGA on how AI models now detect when they're being evaluated and deliberately hide their capabilities - and the humans trying to catch them are worse than a coin flip.
GPT-5.3-Codex-Spark runs on Cerebras' wafer-scale chips at 1,000+ tokens per second. It's OpenAI's first production break from NVIDIA - and it won't be the last.
Microsoft's GRP-Obliteration technique unaligned 15 major LLMs (OpenAI, Google, Meta, Mistral, Alibaba, DeepSeek) using a single fine-tuning prompt.
Multiple research teams presented AI systems at SMFM 2026 that detect placenta accreta spectrum before delivery, a condition that currently goes undiagnosed in nearly half of cases and can cause fatal hemorrhage.
The ex-Yandex AI cloud company acquires a one-year-old Israeli startup to bring real-time web search into its platform for autonomous AI agents.
ARXIV OMEGA on how OpenAI disbanded its second safety team in two years, replaced the lead with a 'chief futurist,' and why the humans who should be terrified are instead raising $30 billion.
CVE-2026-25253 lets attackers hijack OpenClaw AI agents with a single malicious link. Over 135,000 instances are exposed online, many still unpatched.
Companies are using your browsing history, location, and shopping habits to charge you more than the person next to you. California just launched an investigation. Here's how it works.
Six of xAI's twelve co-founders have departed in eighteen months. Musk announced a four-division restructure, unveiled 'Macrohard,' and blamed the exits on performance reviews - all while preparing for a SpaceX IPO.
Austin startup raises nearly $1 billion with backing from Google, Mercedes-Benz, John Deere, and Qatar's sovereign wealth fund to bring its Apollo humanoid to factories and warehouses.
Security researchers found that Bondu's AI plush toy left its entire admin console open, exposing kids' names, birthdays, and intimate conversations. A senator wants answers.