Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Nemotron Nano 12B 2 VL
NVIDIA
Released October 28, 2025

Nemotron Nano 12B 2 VL

NVIDIA Nemotron Nano 12B v2 VL is a vision-language model for multimodal document intelligence, multi-image reasoning, video understanding, visual Q&A.

Visit NVIDIAAnnouncement
Inputs
Image
Text
Video
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Nemotron Nano 12B v2 VL - Vision-language document model

Nemotron Nano 12B v2 VL is built for multimodal document intelligence. It is set up for image, video, and text inputs, with reasoning controlled by the system prompt for text and images.

Distinct traitDocumented behavior
Multi-image reasoningSupports reasoning over multiple images, including up to five input images.
Video understandingHandles video inputs for tasks such as video understanding and video Q&A.
Prompt-controlled reasoningUses /think to enable reasoning and /no_think to disable it for text and images.
Document intelligence focusDescribed for document tasks such as invoices, receipts, manuals, visual Q&A, and summarization.

Input pattern

The model accepts image, video, and text inputs, and its documented use centers on multimodal document workflows. It also supports a 128K input plus output token context window.

Output behavior

The model returns text output and is presented as a commercial-use model. Its documented operation emphasizes document analysis and visual question answering rather than general chat.

Benchmark Performance

Independent evaluations · Artificial Analysis

8.8%
Intelligence

Accuracy & Capability Details

GPQA - Graduate Science57.2%
Humanity's Last Exam5.5%
SciCode - Scientific Coding26.2%
Instruction Following31.9%
Long Context Reasoning42.0%
τ²-Bench - Agentic Tasks21.3%
TerminalBench - System Control4.5%
Specs
Context window
128Ktokens
Input pricing
$0.20per 1M tokens
Output pricing
$0.60per 1M tokens

Prices in USD.

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderNVIDIAAnthropicAnthropic
Release DateOctober 28, 2025July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window128K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.20
Best Input Pricing
$5$10
Output Pricing
$0.60
Best Output Pricing
$25$50
Modalities
Inputs
imagetextvideo
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index4.2
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Nemotron Nano 12B 2 VL
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

44%
Nemotron Nano 12B 2 VL
GPQA Benchmark
Score: 44%
Nemotron Nano 12B 2 VL
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

4%
Nemotron Nano 12B 2 VL
Humanity's Last Exam
Score: 4%
Nemotron Nano 12B 2 VL
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

20%
Nemotron Nano 12B 2 VL
Long Context Reasoning
Score: 20%
Nemotron Nano 12B 2 VL
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models