Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Command A
Cohere
Released March 13, 2025Cutoff August 2024

Command A

Command A by Cohere is a performant enterprise model for tool use, RAG, agents, multilingual use cases, and 256K context.

Visit CohereAnnouncement
Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Command A - Enterprise agentic model

Command A is Cohere’s most performant model for enterprise use cases, positioned around tool use, retrieval-augmented generation, agents, and multilingual workflows. The official docs describe it as a text-only model with a 256K context window and support for structured outputs and reasoning.

Core Design

Cohere states that Command A has 111 billion parameters and requires two GPUs to run, specifically A100s or H100s. The model is described as significantly more efficient at inference time, with 150% higher throughput than Command R+ 08-2024. The docs also say it is interactive by default and optimized for conversation.

Operating Pattern

The model is documented as especially strong in real-world enterprise tasks that require breaking questions into subgoals, active information seeking, and financial numerical manipulation. Cohere also notes that it can respond in the user’s language or follow instructions to output a different language. The model supports citations when working from grounded document snippets.

Benchmark Performance

Independent evaluations · Artificial Analysis

7.5%
Intelligence

Accuracy & Capability Details

GPQA - Graduate Science52.7%
Humanity's Last Exam4.0%
SciCode - Scientific Coding28.1%
Instruction Following36.5%
Long Context Reasoning20.0%
τ²-Bench - Agentic Tasks15.2%
TerminalBench - System Control0.8%
Specs
Context window
256Ktokens
Input pricing
$2.50per 1M tokens
Output pricing
$10per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Translation & Lang80%
Advanced Math70%
Legal & E-Discovery70%
Financial Analysis70%
General Knowledge70%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderCohereAnthropicAnthropic
Release DateMarch 13, 2025July 24, 2026June 9, 2026
Knowledge CutoffAug 2024May 2026-
Context & Limits
Context Window256K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$2.50
Best Input Pricing
$5$10
Output Pricing
$10
Best Output Pricing
$25$50
Modalities
Inputs
text
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index7.5
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Command A
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

53%
Command A
GPQA Benchmark
Score: 53%
Command A
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

4%
Command A
Humanity's Last Exam
Score: 4%
Command A
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

20%
Command A
Long Context Reasoning
Score: 20%
Command A
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models