Nemotron Nano 9B V2
NVIDIA · 2025-08-18 · 8.9B parameters
Nemotron Nano 9B V2 is NVIDIA's 9B-class hybrid Mamba-Transformer language model, released in August 2025 and produced by pruning and distilling a 12B base model pretrained on approximately 20 trillion tokens. Its architecture replaces most self-attention layers with Mamba-2 layers to improve inference throughput on long reasoning traces, with reasoning behavior controllable via system prompt. Because the base model's training-token figure is approximate and does not describe this checkpoint's cumulative exposure, it is not stored as an exact count.