Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Qwen3 30B A3B
Alibaba
Released July 29, 2025Cutoff March 2025

Qwen3 30B A3B

Qwen3 30B A3B is Alibaba's Mixture-of-Experts model with 30.5B total parameters and 3.3B activated per token, featuring hybrid thinking/non-thinking modes, 128 experts with 8 active, and 119-language support.

Visit AlibabaAnnouncement
Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Qwen3 30B A3B - Mixture-of-Experts Model with Hybrid Reasoning Modes

Qwen3 30B A3B is a Mixture-of-Experts large language model from the Qwen3 family designed to balance high performance with efficient inference. It has 30.5 billion total parameters with 3.3 billion active per forward pass.

What Makes It Different: MoE Architecture with Hybrid Thinking Modes

Qwen3 30B A3B's MoE design and dual reasoning modes distinguish it from dense models:

Distinctive TraitWhy It MattersHow It Works
3.3B activated parametersEfficient inference without performance lossUses 8 experts out of 128 total during inference, activating only 3.3B of 30.5B total per token
Hybrid reasoning modesComplex tasks need step-by-step thinking, simple queries need speedSupports both thinking mode for deep reasoning and non-thinking mode via /no_think for rapid responses
128-expert poolBroader knowledge coverage with selective activationSelects 8 optimal experts from 128 pool for each forward pass based on input context
131K token context with YaRNLong documents and massive datasets require extended contextNative 32K context extended to 131,072 tokens using YaRN rope scaling method
Strong-to-weak distillationCompetitive performance from smaller modelTrained via distillation from larger Qwen3 models, maintaining reasoning and coding quality with fewer activated parameters

Why It Exists

Qwen3 30B A3B was built to deliver competitive performance on reasoning, coding, and multilingual benchmarks while using significantly fewer activated parameters than previous models. It outperforms QwQ-32B with 10 times fewer activated parameters.

Core Application Focus

The model is optimized for conversational AI, code assistance, agentic systems, search, multimedia, and enterprise RAG. It supports over 100 languages for multilingual instruction and translation tasks.

Benchmark Performance

Independent evaluations · Artificial Analysis

9.2%
Intelligence
12.1%
Coding Index
1.8%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science61.6%
Humanity's Last Exam6.2%
SciCode - Scientific Coding28.5%
Instruction Following41.5%
Long Context Reasoning0.0%
τ²-Bench - Agentic Tasks26.0%
TerminalBench - System Control2.3%
Specs
Context window
131Ktokens
Input pricing
$0.20per 1M tokens
Output pricing
$0.80per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Creative Writing90%
Creativity90%
Advanced Math80%
Scientific Biology70%
General Knowledge70%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderAlibabaAnthropicAnthropic
Release DateJuly 29, 2025July 24, 2026June 9, 2026
Knowledge CutoffMar 2025May 2026-
Context & Limits
Context Window131K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.20
Best Input Pricing
$5$10
Output Pricing
$0.80
Best Output Pricing
$25$50
Modalities
Inputs
text
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index8.9
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Qwen3 30B A3B
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

66%
Qwen3 30B A3B
GPQA Benchmark
Score: 66%
Qwen3 30B A3B
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

7%
Qwen3 30B A3B
Humanity's Last Exam
Score: 7%
Qwen3 30B A3B
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

25%
Qwen3 30B A3B
Long Context Reasoning
Score: 25%
Qwen3 30B A3B
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models