Best Local Vision Models by VRAM (October 2026)
Open-weight vision-language models sized by VRAM: Qwen3-VL 2B to 235B-A22B, Mistral Small 3.2, Gemma 3, Pixtral, and where each tier breaks.
Tag
Open-weight vision-language models sized by VRAM: Qwen3-VL 2B to 235B-A22B, Mistral Small 3.2, Gemma 3, Pixtral, and where each tier breaks.