Qwen 3.6 35B-A3B

Apache 2.0

Alibaba · 36B (3B active) · Mixture of Experts

Big-model quality at 3B-active speed — the mid-hardware sweet spot Check if your GPU or Mac can run Qwen 3.6 35B-A3B locally — 20.1 GB min, 33.5 GB recommended.

2026-04256K context

Mixture of Experts

Total experts: 256
Active experts: 8
Active params: 3.0B

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K212 GBlow
Q3_K_M316.6 GBmoderate
Q4_K_M418.9 GBgood
Q5_K_M523.6 GBgood
Q6_K628.2 GBexcellent
Q8_0837.4 GBexcellent
F161674.3 GBlossless

Can I run Qwen 3.6 35B-A3B locally?

Can I run Qwen 3.6 35B-A3B locally?
Qwen 3.6 35B-A3B needs about 20.1 GB of memory at a minimum and 33.5 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen 3.6 35B-A3B need?
At Q4_K_M, Qwen 3.6 35B-A3B uses about 18.9 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.