NVIDIA model product
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.
Updated Aug 10, 2026. Default version: Nemotron Nano 9B v2
Structured fields from the published default version.
This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.
Published benchmark records for the scored version Nemotron Nano 9B v2.
| BFCL_v3_MultiTurn | 0.7 | 2 | 2 | 0.0% | C | |
| MATH-500 | 1.0 | 5 | 32 | 87.1% | C | |
| IFEval | 0.9 | 14 | 65 | 79.7% | C | |
| LiveCodeBench | 0.7 | 19 | 73 | 75.0% | C | |
| AIME 2025 | 0.7 | 85 | 114 | 25.7% | C | |
| GPQA | 0.6 | 153 | 233 | 34.5% | C |
Preference and agent-evaluation signals from published Arena datasets.
The default version has no published Arena rows, or its source alias has not been resolved.
Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.
The default version has no provider offering with current input or output token prices.
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.
| Nemotron Nano 9B v2 | 61.2 | 8.9B | 131.1K | 131.1K | Yes | NVIDIA Open Model License Agreement |
A concise description based on the published model registry.
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.
Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.
Data snapshot: 2026-08-07. Editorial model content is not available in the backend.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest published LLMBoard score.
Common questions about Nemotron Nano 9B.
Nemotron Nano 9B's default version was released on Aug 18, 2025.
No official standard PAYG price is currently available for Nemotron Nano 9B.
Nemotron Nano 9B is published under NVIDIA in the model registry.
The default version has a 131.1K token context window.
Yes. The default version is marked as open weight under NVIDIA Open Model License Agreement .
No published provider offering is currently linked to the default version.
Nearby ranked alternatives include o3 mini, Qwen3.5 122B A10B, LongCat Flash Thinking.