Local AI Without a GPU: What Actually Runs (August 2026)
Which open models run on CPU and RAM alone, how to size them, why mixture-of-experts helps, and what a machine with no discrete GPU cannot do.
Tag
Which open models run on CPU and RAM alone, how to size them, why mixture-of-experts helps, and what a machine with no discrete GPU cannot do.
Alibaba's new 35B model matches Claude Sonnet 4.5 on benchmarks while running locally on an RTX 4090. Here's what you need to know.
Qwen 3.5 offers a 397B MoE flagship and smaller local models under Apache 2.0, but Alibaba's benchmarks need independent testing.