Alibaba · 480B (35B active) · 专家混合(MoE)
Largest open coding MoE — 35B active 看看你的 GPU 或 Mac 跑不跑得動 Qwen 3 Coder 480B——最低 268.2 GB,建議 447 GB。
2025-07256K context
专家混合(MoE)
专家总数: 128
启用专家: 8
启用参数: 35.0B
| 量化 | 位元 | VRAM | 品質 | 状態 |
|---|---|---|---|---|
| Q2_K | 2 | 154.2 GB | low | — |
| Q3_K_M | 3 | 215.6 GB | moderate | — |
| Q4_K_M | 4 | 246.4 GB | good | — |
| Q5_K_M | 5 | 307.8 GB | good | — |
| Q6_K | 6 | 369.3 GB | excellent | — |
| Q8_0 | 8 | 492.2 GB | excellent | — |
| F16 | 16 | 984 GB | lossless | — |
关于这個模型
Qwen3-Coder is the most agentic code model to date in the Qwen series.
Get started
480B
Cloud
ollama run qwen3-coder:480b-cloud
Local
ollama run qwen3-coder:480b
Running locally requires a minimum of 250GB of memory or unified memory.
30B
ollama run qwen3-coder:30b
Overview
qwen3-coder:30b offers 30B total parameters with only 3.3B activated, delivering strong performance while maintaining efficiency.
- Exceptional agentic capabilities for real-world software engineering tasks through advanced long-horizon reinforcement learning on SWE-Bench and similar benchmarks.
- Long context support with 256K tokens natively and up to 1M tokens using extrapolation methods, optimized for repository-scale understanding.
- Scaled pretraining on 7.5T tokens (70% code ratio) while preserving strong general and mathematical abilities.
- Execution-driven reinforcement learning that significantly boosts code execution success rates across diverse real-world coding tasks.
Reference
Can I run Qwen 3 Coder 480B locally?
- Can I run Qwen 3 Coder 480B locally?
- Qwen 3 Coder 480B needs about 268.2 GB of memory at a minimum and 447 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
- How much VRAM does Qwen 3 Coder 480B need?
- At Q4_K_M, Qwen 3 Coder 480B uses about 246.4 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.