Claude 3 Haiku by Anthropic: a compact Claude 3 variant optimized for near-instant responses, 200K-token context window, 4K max output, knowledge cutoff Aug 2023.
Capabilities, design details, and architectural traits
Claude 3 Haiku is the compact member of Anthropic's Claude 3 family designed for near-instant responses and high efficiency. It trades top-end generative depth for lower latency and smaller runtime footprint while keeping the large context window used across the Claude 3 family.
| Trait | Documented detail that distinguishes Claude 3 Haiku |
|---|---|
| Identity within family | The smallest and fastest Claude 3 variant in Anthropic's Claude 3 lineup (Haiku vs Sonnet vs Opus). |
| Optimization target | Explicitly optimized for near-instant response speed and runtime efficiency. |
| Context capacity | Uses the Claude 3 family context window of up to 200,000 tokens/characters as documented for Haiku. |
| Output limit | Documented maximum single response output capped at 4,000 characters/tokens. |
| Primary positioning | Positioned for use cases needing low latency and lower compute cost compared with larger Claude 3 models. |
| Official availability | Offered through Anthropic channels and available on cloud marketplaces (for example, Amazon Bedrock) as the Haiku edition. |
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Anthropic | Anthropic | Anthropic |
| Release Date | March 4, 2024 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Aug 2023 | May 2026 | - |
| Context & Limits | |||
| Context Window | 200K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.25 Best Input Pricing | $5 | $10 |
| Output Pricing | $1.25 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimage | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 3.5 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.