Qwen3.5 122B A10B is Qwen's multimodal Mixture-of-Experts model for text, image, and video input, built on Gated Delta Networks and early fusion training.
Capabilities, design details, and architectural traits
Qwen3.5 122B A10B is documented as a multimodal vision-language Mixture-of-Experts model designed for native multimodal agent applications. Unlike text-only models, it is built from the ground up to take text, image, and video as input within the same conversation.
| Trait | Detail |
|---|---|
| Defining purpose | A multimodal vision-language Mixture-of-Experts model designed for native multimodal agent applications, supporting text, image, and video inputs. |
| Video input support | Documented with video input examples that let it answer questions about specific details inside a video clip, rather than only static images. |
| Gated Delta plus MoE | Built on a hybrid architecture combining Gated Delta Networks with a sparse Mixture-of-Experts design, aimed at high-throughput inference. |
| Early fusion training | Trained with early fusion on multimodal tokens, a unified approach to vision-language learning rather than attaching a separate vision encoder to a text only model. |
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Alibaba | Anthropic | Anthropic |
| Release Date | February 24, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 262K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.40 Best Input Pricing | $5 | $10 |
| Output Pricing | $3.20 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimagevideo | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 28.2 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 43.3 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 16.2 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.