Alibaba
Released May 19, 2026

Qwen3.7 Max

Qwen3.7 Max by Alibaba is a text-only reasoning agent model with 1M-token context, cross-harness training, preserve_thinking API, and closed-weight API-only.

Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Qwen3.7 Max - Alibaba's Text-Only Agent Frontier Model

Qwen3.7 Max is officially framed by Alibaba under the title "The Agent Frontier". It is a text-only reasoning model - it accepts text input and returns text output only, with no vision capability. Its training methodology explicitly separates the task, the agentic harness, and the verifier as distinct components to prevent the model from learning tricks tied to any single agent framework.

Documented Traits That Distinguish Qwen3.7 Max

TraitDetail
Cross-harness agent trainingTrained on many combinations of task, harness, and verifier - explicitly designed to prevent single-harness overfitting and deliver stable performance across diverse agent frameworks
preserve_thinking API parameterRetains <think> reasoning blocks across conversation turns, preventing loss of reasoning chain between tool calls in long-horizon sessions
1M-token context windowText-only; 64K output token limit; context window doubled from Qwen3.6 Max Preview's 256K
Native OpenAI and Anthropic API compatibilityDocumented at launch; drops into existing OpenAI-compatible or Anthropic-compatible integrations without infrastructure changes
Closed weightsAPI-only access via Alibaba Cloud Model Studio - a documented break from the prior Qwen open-weight release pattern

How Cross-Harness Training Defines the 3.7 Generation

Alibaba's documented approach separates three typically coupled components in agent training: the task, the agentic harness (the tool-calling scaffold), and the verifier (the success evaluator). Training across many combinations of these components prevents the model from learning patterns specific to one setup. This is the official explanation for Qwen3.7 Max's documented ability to generalize across frameworks including OpenClaw, Hermes Agent, and Qwen Code without per-framework tuning.

Benchmark Performance

Independent evaluations · Artificial Analysis

29.5%
Intelligence
66.0%
Coding Index
22.5%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science92.3%
Humanity's Last Exam40.5%
SciCode - Scientific Coding49.5%
Instruction Following80.5%
Long Context Reasoning79.0%
τ²-Bench - Agentic Tasks94.7%
TerminalBench - System Control50.8%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderAlibabaAnthropicAnthropic
Release DateMay 19, 2026September 22, 2026September 28, 2026
Knowledge Cutoff--Jun 2026
Context & Limits
Context Window1M1M1M
Pricing (per 1M tokens)
Input Pricing$2.50$4
$2
Best Input Pricing
Output Pricing
$7.50
Best Output Pricing
$20$10
Modalities
Inputs
text
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index29.5
57.6
Best Intelligence Index
56.0
Coding Index66.0--
Agentic Index22.5--
Qwen3.7 Max
Claude Opus 5.5
Claude Sonnet 5.5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

41%
Qwen3.7 Max
Humanity's Last Exam
Score: 41%
Qwen3.7 Max
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5

Long Context Reasoning

Logical reasoning over long context windows.

79%
Qwen3.7 Max
Long Context Reasoning
Score: 79%
Qwen3.7 Max
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5

SciCode Benchmark

Scientific coding and mathematical modeling.

50%
Qwen3.7 Max
SciCode Benchmark
Score: 50%
Qwen3.7 Max
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from Alibaba

Other models by Alibaba

Top AI Models

Leading alternatives by intelligence score

View all