Qwen3.5 397B A17B is Qwen's flagship model that launched the Qwen3.5 series, unifying separate vision, text, and reasoning model lines into one.
Capabilities, design details, and architectural traits
Qwen3.5 397B A17B is documented as a multimodal foundation model featuring a Hybrid Mixture-of-Experts architecture with early fusion vision-language training. It is the first model Qwen released in the Qwen3.5 series, built to fold capabilities that used to live in separate model lines into one.
| Trait | Detail |
|---|---|
| Defining purpose | A multimodal foundation model with early fusion vision-language training, designed for chat, retrieval-augmented generation, vision-language understanding, video understanding, and agentic workflows. |
| First Qwen3.5 release | Documented as the first model released in the Qwen3.5 series, serving as the flagship entry point for the generation rather than a smaller companion size. |
| Unifies vision and text lines | Built as one native multimodal model, rather than maintaining separate vision focused and text only model lines as in the prior Qwen3 generation. |
| One model, two reasoning modes | Supports both a thinking mode and a non-thinking mode inside a single model, rather than shipping separate instruct and thinking variants as before. |
Qwen documents the model's efficiency as coming from Gated Delta Networks combined with a sparse Mixture-of-Experts design, aimed at high throughput inference with reduced latency and overhead.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Alibaba | Anthropic | Anthropic |
| Release Date | February 16, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 262K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.60 Best Input Pricing | $5 | $10 |
| Output Pricing | $3.60 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimagevideo | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 32.7 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.