Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Gemma 4 26B A4B
Google
Released April 2, 2026

Gemma 4 26B A4B

Google Gemma 4 26B A4B: Mixture-of-Experts open model with 26B total but only 4B active parameters per token, running near 4B model speed with 256K context.

Visit GoogleAnnouncement
Inputs
Image
Text
Video
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Gemma 4 26B A4B - MoE Open Model at Dense 4B Inference Speed

Gemma 4 26B A4B is Google's Mixture-of-Experts (MoE) model in the Gemma 4 family. The A4B in its name means 4B active parameters - only 4 billion of the 26 billion total parameters activate per token during inference. Google officially documents it as running almost as fast as a 4B-parameter model, while carrying the capacity of a 26B model for routing and expert diversity.

TraitDetail
MoE inference pattern26B total parameters loaded into memory; only a 4B active subset used per token - explicitly designed for high-throughput with near-4B latency
Documented speedRuns almost as fast as a dense 4B model despite 26B total weight, contrasting with the 31B dense sibling which maximizes raw quality
MTP (Multi-Token Prediction)Supported but MoE-specific: gains depend on batch size; at batch size 1, expert weight reuse is limited and speedups are not guaranteed across all hardware
Context window256K tokens - same as the 12B and 31B variants, larger than E2B/E4B (128K)
Architecture reuseServes as the base architecture for DiffusionGemma, Google's experimental discrete text-diffusion model

High-Throughput Positioning within Gemma 4

Within the Gemma 4 family, the 26B A4B is the variant explicitly designated for high-throughput use. Where the 31B dense model is positioned to maximize raw quality and fine-tuning depth, the 26B A4B trades some of that quality ceiling for significantly faster tokens-per-second in production inference workloads.

Benchmark Performance

Independent evaluations · Artificial Analysis

26.1%
Intelligence
39.3%
Coding Index
11.0%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science79.2%
Humanity's Last Exam19.3%
SciCode - Scientific Coding40.0%
Instruction Following72.4%
Long Context Reasoning61.7%
τ²-Bench - Agentic Tasks43.6%
TerminalBench - System Control13.6%
Specs
Context window
262Ktokens
Input pricing
$0.13per 1M tokens
Output pricing
$0.40per 1M tokens
Cached input
$0.13per 1M tokens

Prices in USD.

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderGoogleAnthropicAnthropic
Release DateApril 2, 2026July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window262K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.13
Best Input Pricing
$5$10
Output Pricing
$0.40
Best Output Pricing
$25$50
Modalities
Inputs
imagetextvideo
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index20.4
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Gemma 4 26B A4B
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

71%
Gemma 4 26B A4B
GPQA Benchmark
Score: 71%
Gemma 4 26B A4B
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

12%
Gemma 4 26B A4B
Humanity's Last Exam
Score: 12%
Gemma 4 26B A4B
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

41%
Gemma 4 26B A4B
Long Context Reasoning
Score: 41%
Gemma 4 26B A4B
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models