Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Claude Sonnet 4.6
Anthropic
Released February 17, 2026

Claude Sonnet 4.6

Claude Sonnet 4.6 by Anthropic: default model on claude.ai, near-Opus intelligence at Sonnet pricing, major computer use upgrade, adaptive thinking, and 1M token context window in beta.

Visit AnthropicAnnouncement
Inputs
Text
Image
File
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Claude Sonnet 4.6 - Sonnet-Tier Model With Near-Opus Performance

Claude Sonnet 4.6 is the default model across Claude.ai and Claude Cowork for Free and Pro users. Its defining identity is closing the gap to Opus-class intelligence while remaining at Sonnet pricing - to the point where early users in Claude Code preferred it over Opus 4.5, the previous generation flagship, 59% of the time.

TraitWhat it means for Sonnet 4.6 specifically
Default model on claude.ai and CoworkReplaces prior Sonnet as the default for all Free and Pro users across claude.ai and Claude Cowork at launch
Computer use - major generational stepOfficially described as a "major improvement" over Sonnet 4.5; prompt injection resistance matches Opus 4.6, upgraded from experimental toward production use
Both adaptive and extended thinking supportedSonnet 4.6 supports both thinking: {type: "adaptive"} and the older budget_tokens extended thinking mode (deprecated but still functional), unlike Opus 4.7+ which dropped budget_tokens entirely
1M token context window in betaAvailable in beta on the API; Anthropic documents that the model reasons effectively across this full context, not just ingests it
Strong performance with thinking offOfficially documented to deliver strong performance even with extended thinking disabled - explicitly recommended to explore across the effort spectrum
Context compaction in betaAutomatically summarizes older context as conversations approach limits, extending effective length for long agentic sessions

The Prompt Injection Improvement

Anthropicexplicitly highlights prompt injection resistance as a safety advance specific to Sonnet 4.6 in its computer use context. The system card confirms Sonnet 4.6 is a major improvement over Sonnet 4.5 on this vector, and now performs similarly to Opus 4.6 - a documented safety parity that was not present in prior Sonnet models.

Benchmark Performance

Independent evaluations · Artificial Analysis

48.4%
Intelligence
63.0%
Coding Index
42.1%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science87.5%
Humanity's Last Exam33.6%
SciCode - Scientific Coding46.8%
Instruction Following56.6%
Long Context Reasoning74.0%
τ²-Bench - Agentic Tasks75.7%
TerminalBench - System Control53.0%
Specs
Context window
1Mtokens
Input pricing
$3per 1M tokens
Output pricing
$15per 1M tokens
Cached input
$0.30per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Scientific Biology90%
Physics Problems90%
Translation & Lang90%
Chemistry Concepts90%
Communication90%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderAnthropicAnthropicAnthropic
Release DateFebruary 17, 2026July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window1M1M1M
Pricing (per 1M tokens)
Input Pricing
$3
Best Input Pricing
$5$10
Output Pricing
$15
Best Output Pricing
$25$50
Modalities
Inputs
textimagefile
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index48.4
63.1
Best Intelligence Index
62.1
Coding Index63.0
78.0
Best Coding Index
76.5
Agentic Index42.1
59.2
Best Agentic Index
56.6
Claude Sonnet 4.6
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

88%
Claude Sonnet 4.6
GPQA Benchmark
Score: 88%
Claude Sonnet 4.6
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

34%
Claude Sonnet 4.6
Humanity's Last Exam
Score: 34%
Claude Sonnet 4.6
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

74%
Claude Sonnet 4.6
Long Context Reasoning
Score: 74%
Claude Sonnet 4.6
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models