NVIDIA Nemotron Nano 9B V2 is NVIDIA’s unified reasoning and non-reasoning model with /think controls, tool calling, and 128K context.
Capabilities, design details, and architectural traits
Nemotron Nano 9B V2 is NVIDIA’s model for both reasoning and non-reasoning tasks. It is built to generate a reasoning trace first, then a final response, with runtime control over that behavior.
| Distinctive trait | What is documented |
|---|---|
| Unified task mode | Designed as a unified model for both reasoning and non-reasoning tasks. |
| Reasoning trace control | Uses /think and /no_think style control for reasoning on or off. |
| Tool calling support | Documented with native tool-calling use. |
| Long context | Supports up to 128K context. |
The model can answer with intermediate reasoning traces or without them, depending on the prompt control. Its documented behavior is to produce the trace first and the final response afterward. The model card also describes a runtime thinking budget for controlling how much it reasons.
The model uses a hybrid Mamba-2 and MLP design with four attention layers. The official docs place it in the Nemotron Nano v2 family and describe it as optimized for long-sequence use.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | NVIDIA | Anthropic | Anthropic |
| Release Date | August 18, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Mar 2025 | May 2026 | - |
| Context & Limits | |||
| Context Window | 128K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.05 Best Input Pricing | $5 | $10 |
| Output Pricing | $0.20 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 7.2 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.