Ollama's New Pricing: What the Credit-Pool Change Actually Means
Ollama swapped GPU-hour billing for per-token credits across Pro, Max and Team plans. What the tiers cost, what's free, and how the no-logging promise holds up.
Tag
Ollama swapped GPU-hour billing for per-token credits across Pro, Max and Team plans. What the tiers cost, what's free, and how the no-logging promise holds up.
Open weights is not an open licence. Verified Hugging Face licence tags for 46 model repos, and the size and version traps that block shipping.
Qwen's 27B vision-language model is on Hugging Face ungated under Apache 2.0, and AMD says it needs roughly 24GB of VRAM to run comfortably.
At Ai4 2026, Hinton, Li, and Ng shared a stage and split over open-weight AI. The only consensus: regulation belongs in the conversation.
Alibaba announced 2.4T-parameter Qwen3.8-Max open weights and a 27B sibling that Unsloth says will run in 17GB of RAM or VRAM.
A new industry letter asks Washington to protect open-weight AI while separating legitimate distillation from alleged theft of closed models.
Backed by $400M, the nonprofit Current AI wants to build a free, public-interest AI stack modeled on the early Web. Here's who's funding it and what's shipping.
A leaked board email proposed a local GPT-3-class model in 2022. OpenAI's later gpt-oss release shows how that strategy changed.
UK AISI's July 17 evaluation finds GLM-5.2 and DeepSeek V4-Pro match closed frontier models from 4-7 months ago at a fraction of the cost.
A tiny Australian startup is fine-tuning a model to break ChatGPT-style groupthink. The 'Time is a river' problem is the symptom.
DeepSeek V4, Cohere Command A+, ZAYA1-8B, and NVIDIA Nemotron 3 mark the busiest month for open-weight AI ever.
A tier-by-tier comparison of the top open-weight LLMs you can run locally, from 8GB laptops to 24GB gaming GPUs to Apple Silicon Macs.
Qwen 3.5 offers a 397B MoE flagship and smaller local models under Apache 2.0, but Alibaba's benchmarks need independent testing.