Qwen3-VL 8B

Apache 2.0

Alibaba · 8.8B · 密集(Dense)

The community-favourite local VLM — superb OCR, receipts & captioning 看看你的 GPU 或 Mac 跑不跑得動 Qwen3-VL 8B——最低 4.9 GB,建議 8.2 GB。

2025-10256K context

量化選項

量化位元VRAM品質狀態
Q2_K23.3 GBlow
Q3_K_M34.4 GBmoderate
Q4_K_M45 GBgood
Q5_K_M56.1 GBgood
Q6_K67.3 GBexcellent
Q8_089.5 GBexcellent
F161618.5 GBlossless

Can I run Qwen3-VL 8B locally?

Can I run Qwen3-VL 8B locally?
Qwen3-VL 8B needs about 4.9 GB of memory at a minimum and 8.2 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 8B need?
At Q4_K_M, Qwen3-VL 8B uses about 5 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.