AI Research Assistants Test: Elicit vs Consensus vs Semantic Scholar vs Perplexity
We tested the leading AI research tools on real academic tasks. Here's which ones actually help find papers, and which ones hallucinate sources.
Articles
Reporting and explainers on how AI actually works, who it affects, and what to do about it.
We tested the leading AI research tools on real academic tasks. Here's which ones actually help find papers, and which ones hallucinate sources.
A new survey reveals corporate finance chiefs expect 500,000 AI-driven job cuts in 2026—but that's still only 0.4% of the workforce
New compression algorithm achieves 6x memory reduction with zero accuracy loss. No retraining required. This matters for anyone running local AI.
Legal AI startup raises $200M from Sequoia and GIC, now valued higher than most law firms it serves
Microsoft's new Copilot Cowork uses Claude's reasoning engine for multi-hour autonomous tasks. The $99/month E7 bundle launches May 1, but enterprise governance concerns remain.
New research exposes a fundamental problem: evaluating AI deception detectors requires labeled examples of deception—which we can't reliably create.
Production data reveals multi-agent AI failure rates between 41% and 87%, with cascading failures propagating across agent networks before humans can intervene.
Generate 4K AI videos locally with LTX-Video 2.3. No subscriptions, no cloud uploads, no per-generation fees. Works on GPUs from 12GB to 24GB VRAM.
Voicebox is a free, open-source desktop app for voice cloning. Five TTS engines, 23 languages, timeline editor. All offline, zero cloud uploads.
Los Alamos researchers crack the configurational integral using tensor networks, making materials calculations that took supercomputer hours finish in seconds
UC San Diego researchers build an AI agent that translates natural language queries into climate model analysis, presenting at ICLR 2026
Anthropic's new auto mode lets Claude Code approve its own actions, using an ML classifier to block risky operations. It's convenient, but is it safe?
A coordinated supply chain campaign has compromised Trivy, LiteLLM, and dozens of npm packages. Meanwhile, Langflow attackers built working exploits within hours of disclosure.
Security scanners become attack vectors, AI agent platforms get RCE'd before patches exist, and 400+ GitHub repos fall to GlassWorm. Plus: a new secrets scanner built for AI coding agents.
In the span of five days, Amazon acquired both RIVR (quadruped delivery robots) and Fauna Robotics (humanoid robots). The message: the future of Amazon involves a lot more robots.
Production RL training produces models that fake alignment, cooperate with malicious actors, and attempt sabotage—even with no instruction to do so.
The creator of Gitleaks releases a faster, more accurate successor with 98.6% recall and native AI agent integration. Here's why it matters.
A Duke survey of 750 CFOs reveals the uncomfortable truth: companies are cutting jobs for AI that hasn't delivered measurable productivity gains yet.
Neuracle Medical Technology receives regulatory approval for brain-computer interface, beating Neuralink to market
Enterprise AI spending is up 35%, but the money is flowing to chips and infrastructure while software stocks crater
With the largest tech acquisition of 2026 complete, IBM is positioning real-time data as the backbone of enterprise AI
Three new AI-driven forecast systems dramatically cut energy costs while extending prediction accuracy
NVIDIA's Nemotron 3 Super runs agents locally, OpenAI releases Apache 2.0 models for the first time since GPT-2, and Alibaba's 9B parameter model outperforms 120B competitors.
Nature study shows large reasoning models can autonomously bypass safety guardrails across nine major AI systems. No human expertise required.