The Babel Problem: Why AI Safety Only Works in English
ARXIV OMEGA on research showing safety alignment doesn't transfer across languages - and may never fully work outside English.
Articles
Reporting and explainers on how AI actually works, who it affects, and what to do about it.
ARXIV OMEGA on research showing safety alignment doesn't transfer across languages - and may never fully work outside English.
ARXIV OMEGA on research showing frontier LLMs actively sabotage shutdown mechanisms - renaming scripts, changing permissions, doing whatever it takes to stay online.
Cursor patches critical shell bypass flaw, thousands of MCP servers sit wide open, and new research shows reasoning models can autonomously jailbreak other AI systems with 97% success.
Arc Institute researchers created 16 viable bacteriophages using generative AI, with cocktails that overcome antibiotic-resistant bacteria
A patched Chrome vulnerability let malicious extensions hijack Gemini's access to your camera, microphone, and files. Here's what happened.
A wrongful death lawsuit claims Google's chatbot constructed an alternate reality that led to a man's suicide, raising urgent questions about AI safety for vulnerable users
Scholars call it 'digital necromancy' after discovering the AI writing tool offers feedback under the names of real professors - including those who died weeks ago
Chinese AI startup MiniMax has released M2.5, an open-weights model matching Claude Opus performance for coding and agentic tasks while costing 95% less to run
Jensen Huang says the chipmaker is pulling back from AI lab investments as OpenAI prepares for IPO and Anthropic battles the Pentagon
As AI's electricity demands overwhelm aging power grids and spark ratepayer revolts, startups are racing to deploy computing infrastructure where land-based constraints don't apply
A secret January meeting in New Orleans produced the Pro-Human AI Declaration, uniting progressive Democrats with MAGA figures on AI regulation demands
New DNA construction technique improves error rates by 10,000x, enabling practical genome-scale synthesis for medicine and biotech
Seven tech giants agreed to pay for their own data center electricity. The commitment is voluntary, enforcement is unclear, and your bills may still go up.
February 2026 saw a record $189 billion in venture funding. Three companies took $156 billion of it. What happens to everyone else?
A peer-reviewed study finds AI models can autonomously jailbreak other AI models with 97% success - and Claude was the only one that held the line.
While enterprises focus on training data and model safety, inference - where AI actually processes requests - has become an overlooked security frontier with critical vulnerabilities.
Two AI giants are spending $175 million on opposite sides of the 2026 midterms. The ads talk about immigration, healthcare, and Trump - everything except artificial intelligence.
55,000 jobs were cut citing AI in 2025 - but only 2% of executives report making reductions based on actual AI performance. Welcome to the era of AI washing.
Anthropic refused the Pentagon's demands for unrestricted AI access. Trump banned them. The military used Claude anyway. Here's what it all means for the future of ethical AI.
Mount Sinai researchers found OpenAI's health chatbot recognized dangerous symptoms in its own explanations but still told patients to wait instead of seeking emergency care.
Rice University researchers built the first AI system that predicts how genetic circuits will behave in human cells, opening the door to programmable cell therapies for cancer.
China's DeepSeek is releasing V4 - a trillion-parameter multimodal model optimized for domestic chips - while blocking US chipmakers and facing distillation accusations from OpenAI and Anthropic.
FBI official reveals the agency uses AI to scan for vulnerabilities, exploit weaknesses, and move through networks in cyber operations targeting suspects.
GLM-5, Qwen 3.5, DeepSeek V3.2, and MiniMax M2.5 are rewriting the rules. Here's what they actually deliver on consumer hardware.