Qwen3-VL 4B

Apache 2.0

Alibaba · 4.4B · Dense

Compact dedicated vision-language model — OCR & image chat on edge お使いの GPU や Mac で Qwen3-VL 4B が動くか確認——最小 2.5 GB、推奨 4.1 GB。

2025-10256K context

量子化オプション

量子化ビットVRAM品質状態
Q2_K21.9 GBlow
Q3_K_M32.5 GBmoderate
Q4_K_M42.8 GBgood
Q5_K_M53.3 GBgood
Q6_K63.9 GBexcellent
Q8_085 GBexcellent
F16169.5 GBlossless

Can I run Qwen3-VL 4B locally?

Can I run Qwen3-VL 4B locally?
Qwen3-VL 4B needs about 2.5 GB of memory at a minimum and 4.1 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 4B need?
At Q4_K_M, Qwen3-VL 4B uses about 2.8 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.