GLM-5.1 by Z.ai is a flagship agentic engineering model: a 744B/40B-active MoE with SOTA SWE-Bench Pro results and sustained long-horizon optimization.
Capabilities, design details, and architectural traits
GLM-5.1 is Z.ai's next-generation flagship model for agentic engineering: an update to GLM-5 built specifically to stay effective on long-horizon coding and tool-use tasks where earlier models, including GLM-5 itself, tend to plateau.
Built to fix a specific failure mode in prior models, including GLM-5: exhausting familiar techniques early and plateauing on agentic tasks, regardless of how much time or compute they're given.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Z AI | Anthropic | Anthropic |
| Release Date | April 7, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 203K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $1.39 Best Input Pricing | $5 | $10 |
| Output Pricing | $4.40 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 36.3 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.