Sapiens AI
Released August 26, 2026

Agnes 2.5 Pro Beta

Agnes 2.5 Pro Beta by Sapiens AI is a paid reasoning model with text and image input, tool calling, Thinking mode, and agentic terminal task support.

Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Agnes 2.5 Pro Beta - Paid reasoning model with text and image input

Agnes 2.5 Pro Beta is a paid reasoning model from Sapiens AI, positioned above flash-class models in reasoning depth. It targets production workloads such as complex code tasks, science and math reasoning, long-context understanding, knowledge-intensive Q&A, and agentic terminal or workflow tasks. It uses the same request parameters, response formats, supported text protocols, and context limits as the other Agnes 2.5 Pro models.

Key characteristics

TraitDetail
Model typePaid reasoning model with text and image input, text output
Multi-protocol APIServes three endpoints: Chat Completions, Responses, and Messages
Thinking modeOptional thinking field enables Thinking mode in Anthropic-compatible requests
Agentic terminal tasksDesigned for agentic coding workflows and terminal or workflow tasks
Tool callingOpenAI-compatible function calling and external tool orchestration
Image understandingAccepts image URL inputs for visual analysis and multimodal reasoning
Long-context analysisHandles long documents, structured context, and multi-turn reasoning

Integration pattern

The model is called through the Agnes AI API using the live model ID agnes-2.5-pro-beta. It supports streaming responses and standard sampling parameters such as temperature, top_p, and max_tokens, plus the chat_template_kwargs extension field for OpenAI-compatible requests.

Benchmark Performance

Independent evaluations · Artificial Analysis

35.2%
Intelligence
62.3%
Coding Index
43.8%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science90.5%
Humanity's Last Exam37.5%
SciCode - Scientific Coding48.8%
Long Context Reasoning83.0%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderSapiens AIAnthropicAnthropic
Release DateAugust 26, 2026September 22, 2026September 28, 2026
Knowledge Cutoff--Jun 2026
Context & Limits
Context Window262K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.10
Best Input Pricing
$4$2
Output Pricing
$0.30
Best Output Pricing
$20$10
Modalities
Inputs
text
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index35.2
57.6
Best Intelligence Index
56.0
Coding Index62.3--
Agentic Index43.8--
Agnes 2.5 Pro Beta
Claude Opus 5.5
Claude Sonnet 5.5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

38%
Agnes 2.5 Pro Beta
Humanity's Last Exam
Score: 38%
Agnes 2.5 Pro Beta
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5

Long Context Reasoning

Logical reasoning over long context windows.

83%
Agnes 2.5 Pro Beta
Long Context Reasoning
Score: 83%
Agnes 2.5 Pro Beta
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5

SciCode Benchmark

Scientific coding and mathematical modeling.

49%
Agnes 2.5 Pro Beta
SciCode Benchmark
Score: 49%
Agnes 2.5 Pro Beta
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from Sapiens AI

Other models by Sapiens AI

Top AI Models

Leading alternatives by intelligence score

View all