Qwen3-VL 8B

Apache 2.0

Alibaba · 8.8B · Dense

The community-favourite local VLM — superb OCR, receipts & captioning Check if your GPU or Mac can run Qwen3-VL 8B locally — 4.9 GB min, 8.2 GB recommended.

2025-10256K context

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K23.3 GBlow
Q3_K_M34.4 GBmoderate
Q4_K_M45 GBgood
Q5_K_M56.1 GBgood
Q6_K67.3 GBexcellent
Q8_089.5 GBexcellent
F161618.5 GBlossless

Can I run Qwen3-VL 8B locally?

Can I run Qwen3-VL 8B locally?
Qwen3-VL 8B needs about 4.9 GB of memory at a minimum and 8.2 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 8B need?
At Q4_K_M, Qwen3-VL 8B uses about 5 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.