IBM
Released August 25, 2026

Granite 4.2 8B

IBM Granite 4.2 8B is a mid-size dense reasoning model with chain-of-thought, flexible thinking modes, and reasoning-augmented tool calling under Apache 2.0.

Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Granite 4.2 8B - Mid-size dense reasoning model with flexible thinking modes

Granite 4.2 8B is the balanced, general-purpose member of IBM's Granite 4.2 family of dense reasoning language models. It performs step-by-step chain-of-thought reasoning before answering, and lets the user switch reasoning depth on a per-query basis within a single model.

TraitDetail
Built-in reasoningNative chain-of-thought inside `` tags, improving math, coding, and multi-step logic performance
Flexible thinking modesThree modes in one model: thinking (default), non-thinking, and low-effort, selected via chat-template parameters to balance depth vs. latency
Reasoning-augmented tool callingReasons about which tools to invoke and why before making the call, producing more accurate function calls for agentic workflows
Context windowNatively supports 128K, with long-context extension to 512K
Enterprise positioningBalanced reasoning model for general-purpose enterprise applications, post-trained from Granite-4.1-8B-Base
Open licensingApache 2.0 with cryptographic signatures, ISO certification, and full transparency disclosures
MultilingualTested across 12 languages including English, German, Spanish, French, Japanese, Arabic, Korean, and Chinese

Thinking modes as a first-class control

Rather than shipping separate reasoning and non-reasoning variants, Granite 4.2 8B exposes reasoning depth as a runtime switch. Full thinking produces complete chain-of-thought, non-thinking gives a direct answer with no reasoning overhead, and low-effort applies brief reasoning for simpler queries.

Trained for agentic work

The supervised fine-tuning stage combined publicly available permissive datasets, internally generated synthetic data targeting reasoning and tool calling, agentic traces collected across diverse tasks, and curated human-authored data.

Benchmark Performance

Independent evaluations · Artificial Analysis

11.1%
Intelligence
22.4%
Coding Index
1.3%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science63.1%
Humanity's Last Exam9.7%
SciCode - Scientific Coding31.5%
Long Context Reasoning45.0%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderIBMAnthropicAnthropic
Release DateAugust 25, 2026September 22, 2026September 28, 2026
Knowledge Cutoff--Jun 2026
Context & Limits
Context Window131K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.06
Best Input Pricing
$4$2
Output Pricing
$0.25
Best Output Pricing
$20$10
Modalities
Inputs
text
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index11.1
57.6
Best Intelligence Index
56.0
Coding Index22.4--
Agentic Index1.3--
Granite 4.2 8B
Claude Opus 5.5
Claude Sonnet 5.5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

10%
Granite 4.2 8B
Humanity's Last Exam
Score: 10%
Granite 4.2 8B
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5

Long Context Reasoning

Logical reasoning over long context windows.

45%
Granite 4.2 8B
Long Context Reasoning
Score: 45%
Granite 4.2 8B
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5

SciCode Benchmark

Scientific coding and mathematical modeling.

32%
Granite 4.2 8B
SciCode Benchmark
Score: 32%
Granite 4.2 8B
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from IBM

Other models by IBM

Top AI Models

Leading alternatives by intelligence score

View all