Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Nemotron 3 Nano 30B A3B
NVIDIA
Released December 15, 2025

Nemotron 3 Nano 30B A3B

NVIDIA's Nemotron 3 Nano 30B A3B: a hybrid Mamba-Transformer MoE model activating 3.2B parameters with configurable reasoning traces.

Visit NVIDIAAnnouncement
Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Nemotron 3 Nano 30B A3B - Hybrid Mamba-Transformer MoE

Nemotron 3 Nano 30B A3B utilizes an architectural design that merges sequence processing with selective associative recall. It leverages sparse activation to maintain broad knowledge while executing with the computational footprint of a much smaller network.

Architectural ComponentImplementation Details
Hybrid Layer InterleavingIntegrates 23 Mamba-2 and MoE layers for linear-time sequence processing with 6 Attention layers for precise recall.
Expert Routing MechanismUtilizes 128 standard experts plus 1 shared expert per MoE layer, routing each token to exactly 6 active experts.
Parameter SparsityMaintains a total base of 31.6B parameters while activating only 3.2B parameters per forward pass.
Configurable Reasoning TracesIncludes an enable_thinking toggle to bypass <think> block generation for low-latency inference workloads.

Pretraining Optimization

The underlying foundation was pretrained using a specific Warmup-Stable-Decay learning rate schedule before undergoing multi-environment reinforcement learning to stabilize its reasoning outputs.

Benchmark Performance

Independent evaluations · Artificial Analysis

14.5%
Intelligence
14.4%
Coding Index
2.0%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science75.7%
Humanity's Last Exam11.4%
SciCode - Scientific Coding29.6%
Instruction Following71.1%
Long Context Reasoning37.3%
τ²-Bench - Agentic Tasks40.9%
TerminalBench - System Control13.6%
Specs
Context window
256Ktokens
Input pricing
$0.05per 1M tokens
Output pricing
$0.20per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Legal & E-Discovery70%
Financial Analysis70%
General Knowledge70%
Creative Writing70%
Translation & Lang70%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderNVIDIAAnthropicAnthropic
Release DateDecember 15, 2025July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window256K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.05
Best Input Pricing
$5$10
Output Pricing
$0.20
Best Output Pricing
$25$50
Modalities
Inputs
text
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index7.2
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Nemotron 3 Nano 30B A3B
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

40%
Nemotron 3 Nano 30B A3B
GPQA Benchmark
Score: 40%
Nemotron 3 Nano 30B A3B
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

5%
Nemotron 3 Nano 30B A3B
Humanity's Last Exam
Score: 5%
Nemotron 3 Nano 30B A3B
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

9%
Nemotron 3 Nano 30B A3B
Long Context Reasoning
Score: 9%
Nemotron 3 Nano 30B A3B
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models