Creator
NVIDIA AI
Released
2025-08-18
Intelligence
7.4
Artificial Analysis Index
Coding
—
Artificial Analysis Index
In $/1M
$0.05
input tokens
Out $/1M
$0.20
output tokens
Blended $/1M
$0.09
3:1 blended
Speed
177
tokens / sec
Profile
License, openness, and modality — the governance layer.
License
NVIDIA Open Model License Agreement
commercial OK
Weights
Open
downloadable
Modalities
—
Parameters
9B
total
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.
Capability profile
Category strength across 20 domains, via LLM Stats.
Benchmark breakdown
Independent evaluation scores, normalized to 0–100.