Gemma 4 12B IT

Apache 2.0

Google · 12B · Dense

Gemma 4 mid-size instruct — multimodal any-to-any Check if your GPU or Mac can run Gemma 4 12B IT locally — 6.7 GB min, 11.2 GB recommended.

2026-04256K context

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K24.3 GBlow
Q3_K_M35.9 GBmoderate
Q4_K_M46.6 GBgood
Q5_K_M58.2 GBgood
Q6_K69.7 GBexcellent
Q8_0812.8 GBexcellent
F161625.1 GBlossless

Can I run Gemma 4 12B IT locally?

Can I run Gemma 4 12B IT locally?
Gemma 4 12B IT needs about 6.7 GB of memory at a minimum and 11.2 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Gemma 4 12B IT need?
At Q4_K_M, Gemma 4 12B IT uses about 6.6 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.