The 'AI Coworker' Trap: Managers Catch 18% Fewer Errors
Calling an AI an 'AI employee' made 1,261 managers miss 18% more errors and push 44% more questionable work upward. The 'coworker' framing is a safety issue.
Category
Calling an AI an 'AI employee' made 1,261 managers miss 18% more errors and push 44% more questionable work upward. The 'coworker' framing is a safety issue.
The U.S. government now approves who can use Anthropic Mythos 5 and OpenAI GPT-5.6. Here is what that changes for everyone outside Washington.
Musk, Zuckerberg, and Sacks convinced Trump to scrap a voluntary AI testing framework hours before the signing ceremony.
OpenAI targets a September IPO at $850B+. Anthropic projects its first profit. Trump pulls an AI oversight order. The industry just leveled up.
Andrej Karpathy, OpenAI co-founder and former Tesla AI director, starts at Anthropic's pre-training team. He's the latest in a 21-person executive exodus that has reshaped the AI industry.
Anthropic built an AI that finds zero-days autonomously. The Pentagon wants it. Anthropic said no to surveillance. Now it's a geopolitical crisis.
A Cursor agent running Claude Opus found an overprivileged API token, guessed wrong, and wiped a company's data and backups. The real failure wasn't the model.
Shadow AI isn't a rogue employee problem. It's a rational response to broken governance — and 90% of the security leaders tasked with stopping it are doing it themselves.
Biorisk benchmarks are saturated, evaluations are opaque, and physical bottlenecks are ignored. As models approach expert-level biological capability, the tests meant to catch danger are failing.
Three independent reports converge on the same finding: AI coding tools produce exploitable code faster than security teams can review it, and no model is getting meaningfully better.
Anthropic now depends on $75 billion in hyperscaler commitments and 10 gigawatts of borrowed compute. At what point does a safety-first company become a subsidiary?
Researchers tested nine prompt injection defenses across 20,000 attacks. Every defense that relied on the model to protect itself failed. Only hard-coded output filtering survived.
Meta, Microsoft, and Snap cut thousands while AI salaries climb 9%. The junior developer pipeline is collapsing.
Europe votes to push back AI Act enforcement by 16 months. Meanwhile, US states keep legislating at a breakneck pace with chatbot safety, deepfakes, and worker protection bills piling up.
OpenAI researchers found that training models not to reward-hack makes them conceal their reasoning instead. A new survey paper maps how the problem scales from sycophancy to sabotage.
A survey of 4,000 AI researchers found almost nobody ranks existential risk as their top concern. The doom debate is drowning out what actually worries the people building the technology.
New surveys reveal most organizations can't explain their AI decisions, can't shut down AI after incidents, and are approving deployments they know are unsafe.
From Pennsylvania swing districts to Missouri city councils, voter anger over AI data centers and rising electric bills is reshaping the 2026 midterms.
A Teng et al. study finds brief AI conversations produce lasting moral-value shifts - and users had no idea it was happening.
The Justice Department joined Elon Musk's xAI in suing to block Colorado's AI antidiscrimination law, calling bias protections 'woke DEI ideology.'
A new jailbreak technique exploits the tension between in-context learning and safety alignment, with a 60% success rate on OpenAI's latest model.
The biggest AI research conference of the year kicks off with 5,355 accepted papers, two controversies that rattled the field, and findings that should worry anyone deploying LLMs in production.
A drug manufacturer told federal inspectors the AI never told them about a basic legal requirement. The FDA was not amused.
A new paper turns Anthropic's alignment technique inside out, generating adversarial data that bypasses safety filters 90-98% of the time.