Qwen3 14B by Qwen is a dense causal language model with thinking and non-thinking modes, 32K context, YaRN scaling, and tool use support.
Capabilities, design details, and architectural traits
Qwen3 14B is a dense causal language model in the Qwen3 series. Its documented identity centers on switching between a thinking mode and a non-thinking mode, with support for instruction-following, tool use, creative writing, and multilingual tasks.
The model is positioned as a dense 14B-class Qwen3 model rather than a mixture-of-experts variant.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Alibaba | Anthropic | Anthropic |
| Release Date | April 28, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Mar 2025 | May 2026 | - |
| Context & Limits | |||
| Context Window | 132K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.35 Best Input Pricing | $5 | $10 |
| Output Pricing | $1.40 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 6.8 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.