Qwen3-VL 30B-A3B

Apache 2.0

Alibaba · 31B (3B active) · Mixture of Experts

Efficient vision MoE — 3B active, strong temporal & document understanding Check if your GPU or Mac can run Qwen3-VL 30B-A3B locally — 17.3 GB min, 28.9 GB recommended.

2025-09256K context

Mixture of Experts

Total experts: 128
Active experts: 8
Active params: 3.0B

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K210.4 GBlow
Q3_K_M314.4 GBmoderate
Q4_K_M416.4 GBgood
Q5_K_M520.3 GBgood
Q6_K624.3 GBexcellent
Q8_0832.3 GBexcellent
F161664 GBlossless

Can I run Qwen3-VL 30B-A3B locally?

Can I run Qwen3-VL 30B-A3B locally?
Qwen3-VL 30B-A3B needs about 17.3 GB of memory at a minimum and 28.9 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen3-VL 30B-A3B need?
At Q4_K_M, Qwen3-VL 30B-A3B uses about 16.4 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.