Nemotron Nano 9B v2

NVIDIA Open

NVIDIA Β· 9B Β· Dense

Hybrid Mamba2 architecture for reasoning Check if your GPU or Mac can run Nemotron Nano 9B v2 locally β€” 5 GB min, 8.4 GB recommended.

2025-06128K context

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K23.4 GBlowβ€”
Q3_K_M34.5 GBmoderateβ€”
Q4_K_M45.1 GBgoodβ€”
Q5_K_M56.3 GBgoodβ€”
Q6_K67.4 GBexcellentβ€”
Q8_089.7 GBexcellentβ€”
F161618.9 GBlosslessβ€”

Can I run Nemotron Nano 9B v2 locally?

Can I run Nemotron Nano 9B v2 locally?
Nemotron Nano 9B v2 needs about 5 GB of memory at a minimum and 8.4 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Nemotron Nano 9B v2 need?
At Q4_K_M, Nemotron Nano 9B v2 uses about 5.1 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.