Nemotron Nano 9B v2

NVIDIA Open

NVIDIA · 9B · Dense

Hybrid Mamba2 architecture for reasoning Check if your GPU or Mac can run Nemotron Nano 9B v2 locally — 5 GB min, 8.4 GB recommended.

2025-06128K context

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K23.4 GBlow
Q3_K_M34.5 GBmoderate
Q4_K_M45.1 GBgood
Q5_K_M56.3 GBgood
Q6_K67.4 GBexcellent
Q8_089.7 GBexcellent
F161618.9 GBlossless

Can I run Nemotron Nano 9B v2 locally?

Can I run Nemotron Nano 9B v2 locally?
Nemotron Nano 9B v2 needs about 5 GB of memory at a minimum and 8.4 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Nemotron Nano 9B v2 need?
At Q4_K_M, Nemotron Nano 9B v2 uses about 5.1 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.