Anthropic
Released September 28, 2026Cutoff June 2026

Claude Sonnet 5.5

Claude Sonnet 5.5 by Anthropic: mid-tier workhorse model with fewer tool calls, near-flagship performance, and effort-based reasoning for coding and agents.

Inputs
Text
Image
File
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Claude Sonnet 5.5 - Anthropic's efficient mid-tier workhorse

Claude Sonnet 5.5 is Anthropic's faster, more efficient update to its mid-tier workhorse model. Its defining idea is completing a task with fewer tokens and fewer tool calls. In several of Anthropic's own tests it comes strikingly close to the flagship, Opus 5.5.

What sets it apart

TraitDetail
Leaner task completionFinishes jobs with fewer tokens and fewer tool calls
Near-flagship performanceComes strikingly close to Opus 5.5 in several of Anthropic's tests, narrowing the gap between the two tiers
Positioned for defined workBest suited to well-defined everyday work: software debugging, coding, documents, presentations, spreadsheets, and interface design, while Opus 5.5 handles ambiguous or open-ended work
Effort-based reasoningRuns at Low, Medium, or High effort; Claude Code and consumer apps default to Medium, the Claude Platform defaults to High
Leaner agent stepsEarly customers report roughly one-third fewer tool calls and about half as many shell executions to finish coding jobs

Efficiency as the product

Anthropic's pitch with Sonnet 5.5 is that raw step counts tell only part of the story. Every unnecessary tool call, failed action, and retry adds latency to an autonomous workflow. Sonnet 5.5 is built to complete the same jobs in fewer steps, which matters most for coding agents and long-running automated work.

Effort settings shape the tradeoff

Lower effort settings trade some reasoning depth for lower latency and token consumption. Higher settings let the model spend longer checking and refining its work. Anthropic reports that Sonnet 5.5 at Low or Medium effort can surpass Sonnet 5's best result, making effort selection a practical lever for production deployments.

Benchmark Performance

Independent evaluations · Artificial Analysis

56.0%
Intelligence

Accuracy & Capability Details

Humanity's Last Exam55.0%
SciCode - Scientific Coding61.0%
Long Context Reasoning82.7%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderAnthropicAnthropicAnthropic
Release DateSeptember 28, 2026September 22, 2026September 1, 2026
Knowledge CutoffJun 2026--
Context & Limits
Context Window1M1M1M
Pricing (per 1M tokens)
Input Pricing
$2
Best Input Pricing
$4$10
Output Pricing
$10
Best Output Pricing
$20$50
Modalities
Inputs
textimagefile
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index56.0
57.6
Best Intelligence Index
53.4
Coding Index--81.6
Agentic Index--57.9
Claude Sonnet 5.5
Claude Opus 5.5
Claude Fable 5.1

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
59%
Claude Fable 5.1
Humanity's Last Exam
Score: 59%
Claude Fable 5.1

Long Context Reasoning

Logical reasoning over long context windows.

83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
85%
Claude Fable 5.1
Long Context Reasoning
Score: 85%
Claude Fable 5.1

SciCode Benchmark

Scientific coding and mathematical modeling.

61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
63%
Claude Fable 5.1
SciCode Benchmark
Score: 63%
Claude Fable 5.1

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from Anthropic

Other models by Anthropic

Top AI Models

Leading alternatives by intelligence score

View all