Qwen3-VL 8B

Apache 2.0

Alibaba · 8.8B · Dense

The community-favourite local VLM — superb OCR, receipts & captioning お使いの GPU や Mac で Qwen3-VL 8B が動くか確認——最小 4.9 GB、推奨 8.2 GB。

2025-10256K context

量子化オプション

量子化ビットVRAM品質状態
Q2_K23.3 GBlow
Q3_K_M34.4 GBmoderate
Q4_K_M45 GBgood
Q5_K_M56.1 GBgood
Q6_K67.3 GBexcellent
Q8_089.5 GBexcellent
F161618.5 GBlossless

Can I run Qwen3-VL 8B locally?

Can I run Qwen3-VL 8B locally?
Qwen3-VL 8B needs about 4.9 GB of memory at a minimum and 8.2 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 8B need?
At Q4_K_M, Qwen3-VL 8B uses about 5 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.