IBM
Released August 25, 2026

Granite 4.2 30B

IBM Granite 4.2 30B is a dense reasoning model with switchable thinking modes, reasoning-augmented tool calling, 512K context, and an Apache 2.0 license.

Inputs
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Granite 4.2 30B - IBM's flagship dense reasoning model with switchable thinking

Granite 4.2 30B is the largest variant of IBM's dense reasoning model family, built for complex reasoning, coding, and agentic tool-calling workloads. Its defining feature is native chain-of-thought reasoning inside `` tags, combined with the ability to turn that reasoning on, off, or dial it down per query.

What sets Granite 4.2 30B apart

TraitDetail
Flexible thinking modesSwitch between thinking (default), non-thinking, and low-effort modes within a single model to balance depth against latency per query
Reasoning-augmented tool callingReasons about which tools to invoke and why before making the call, for more accurate function calls in agentic workflows
Extended long context128K context window natively, with a long-context extension to 512K available only on the 30B model
Open with enterprise assurancesApache 2.0 license with cryptographic signatures, ISO certification, and full transparency disclosures
Agentic RL trainingA specialized agentic RL phase covering software engineering, terminal-based coding, and search-driven workflows, on top of foundational RL and RLHF alignment
Synthetic code trainingTrained on 1 trillion tokens of synthetic code from IBM's CodeAlchemy pipeline, plus a mid-training step to unlock more reasoning power
Built-in speculative decodingA speculative decoding layer speeds up text output while serving more users

Built for enterprise agentic work

IBM positions Granite 4.2 as purpose-built for agentic workflows where tasks are ambiguous and multi-step. The 30B model serves as the flagship for deeper reasoning and complex coding, while smaller siblings (3B, 8B) handle high-throughput agentic tasks. Software engineering agents built on it can navigate codebases, handle multi-step development tasks, and operate in terminal environments. It is tested across 12 languages, including English, German, Japanese, Arabic, Korean, and Chinese.

Benchmark Performance

Independent evaluations · Artificial Analysis

12.7%
Intelligence
29.9%
Coding Index
3.5%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science64.4%
Humanity's Last Exam11.2%
SciCode - Scientific Coding37.8%
Long Context Reasoning49.0%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderIBMAnthropicAnthropic
Release DateAugust 25, 2026September 22, 2026September 28, 2026
Knowledge Cutoff--Jun 2026
Context & Limits
Context Window131K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.16
Best Input Pricing
$4$2
Output Pricing
$0.65
Best Output Pricing
$20$10
Modalities
Inputs
text
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index12.7
57.6
Best Intelligence Index
56.0
Coding Index29.9--
Agentic Index3.5--
Granite 4.2 30B
Claude Opus 5.5
Claude Sonnet 5.5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

11%
Granite 4.2 30B
Humanity's Last Exam
Score: 11%
Granite 4.2 30B
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5

Long Context Reasoning

Logical reasoning over long context windows.

49%
Granite 4.2 30B
Long Context Reasoning
Score: 49%
Granite 4.2 30B
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5

SciCode Benchmark

Scientific coding and mathematical modeling.

38%
Granite 4.2 30B
SciCode Benchmark
Score: 38%
Granite 4.2 30B
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from IBM

Other models by IBM

Top AI Models

Leading alternatives by intelligence score

View all