A 330 GB On-Die DRAM AI Chip Lands as a Whitepaper
PhantaField's Sophon PFG-1 whitepaper claims ~95x Nvidia HBM4 bandwidth via monolithic 3D stacking. No silicon yet. Here's why it matters anyway.
Tag
PhantaField's Sophon PFG-1 whitepaper claims ~95x Nvidia HBM4 bandwidth via monolithic 3D stacking. No silicon yet. Here's why it matters anyway.
DeepSeek V4, Cohere Command A+, ZAYA1-8B, and NVIDIA Nemotron 3 mark the busiest month for open-weight AI ever.
GLM-5.1 becomes the first open-weight model to top SWE-Bench Pro. The gap between open and proprietary AI is now just three months.
DeepSeek V4 matches Claude Opus on coding at 7x lower cost under MIT license. NVIDIA's Nemotron 3 brings hybrid Mamba-Transformer MoE to the open. Google's TurboQuant cuts KV cache memory by 6x with no retraining.
NVIDIA's Nemotron 3 brings a hybrid Mamba-Transformer architecture to consumer GPUs while Meta abandons open source for proprietary Muse Spark. The open-weight field just reshuffled.
Eight labs unite under NVIDIA's Nemotron Coalition, LangChain open-sources the enterprise coding agent pattern, and Sarvam proves frontier AI doesn't require Silicon Valley.
Hugging Face's Spring 2026 report reveals China now leads in AI model downloads, robotics datasets jumped 2,200%, and open-weight models are achieving 10x-1000x cost advantages.
NVIDIA's Nemotron 3 Super runs agents locally, OpenAI releases Apache 2.0 models for the first time since GPT-2, and Alibaba's 9B parameter model outperforms 120B competitors.
AMD's 6GW data center agreement with Meta represents the largest challenge yet to NVIDIA's AI chip dominance
Warren and Blumenthal accuse NVIDIA of structuring its Groq licensing deal to dodge merger review while consolidating AI chip dominance
Brussels is probing whether NVIDIA bundles GPUs with networking equipment and uses CUDA lock-in to crush competitors. The investigation could take years.
Bill Gurley warns AI reset is coming while Norway's $2.1 trillion fund models a 35% crash. The problem: Nvidia invests in OpenAI, which buys more Nvidia chips.
Federal prosecutors charge three with routing Nvidia AI servers through Taiwan to China using fake documents and dummy equipment to evade export controls.
GTC 2026's biggest announcements were open-source. Nemotron 3 Super runs locally on RTX PCs, LTX 2.3 generates 4K video with audio, and vLLM hits production grade.
Jensen Huang's keynote unveils Vera Rubin chips, a $20B Groq acquisition, DLSS 5, and positions Nvidia to dominate both training and inference markets.
Jensen Huang's keynote today marks Nvidia's biggest pivot in years - from training chips to inference, from cloud to edge, and from prompts to autonomous agents
From the Vera Rubin architecture to NemoClaw enterprise agents, here's everything Nvidia is expected to unveil at its biggest conference of the year
The chip rivals jointly backed a photonics startup that promises 20x better performance per watt. Volume production targets 2028 AI systems.
Nvidia open-sources a 30B-parameter reasoning model that runs on consumer GPUs with a million-token context window. Here's what makes it different.
Just 14 months after Trump announced the $500 billion AI infrastructure project, the flagship Abilene site expansion has fallen apart. Meta and Nvidia are circling the remains.
Draft regulations would require government approval for nearly all Nvidia and AMD chip exports worldwide - echoing Biden rules Trump rescinded just months ago.
Jensen Huang says the chipmaker is pulling back from AI lab investments as OpenAI prepares for IPO and Anthropic battles the Pentagon
China's DeepSeek is releasing V4 - a trillion-parameter multimodal model optimized for domestic chips - while blocking US chipmakers and facing distillation accusations from OpenAI and Anthropic.
Amazon, NVIDIA, and SoftBank pour $110 billion into OpenAI at a $730B valuation. Here's what the partnerships mean and who's actually winning.