Google
Released September 30, 2026

Gemini 4 Argon

Gemini 4 Argon is Google's frontier model designed for long-horizon coding, legal and finance work, and cyber defense, with a 1M token output limit.

Inputs
Text
Image
File
Video
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Gemini 4 Argon - Google's next era of frontier intelligence

Gemini 4 Argon is Google's frontier model built to sustain deep reasoning across complex, long-horizon workflows. It targets real-world software engineering, enterprise knowledge work such as legal and finance, and cybersecurity defense. It also introduces a new naming scheme for Google's model lineup.

Defining traitWhat it means
1M token output limitExpanded from the previous 64K, so the model can think deeply and generate hundreds of thousands of tokens in a single trajectory to solve tough problems in one go.
Cybersecurity defense focusTrained to autonomously find, validate, and patch critical software vulnerabilities, and made available without cyber guardrails to trusted defenders and internal Google teams.
Fairwind Program rolloutFirst released to a limited set of trusted cyber defenders through Google's Fairwind Program, with more than 650 participating partners globally.
Misalignment monitoringDeployed mitigations monitor Argon's chain-of-thought and actions and stop execution when necessary, with precautions against feeding monitoring findings back into training.

Changing how Google works and builds

Argon is already powering internal workflows for thousands of Googlers. Teams of Argon agents analyzed fleet-wide profiling telemetry to autonomously apply memory optimizations across Google's data centers, freeing over 300 TiB of memory. Other Argon agents are migrating C/C++ codebases to Rust, scaling from core libraries like re2 and libgav1 up to 800K+ lines of the Fuchsia Zircon kernel, with all rewrites undergoing rigorous auditing and review before production.

Phased release under frontier safeguards

Google is engaged in the U.S. government's voluntary process for pre-release model access while gradually expanding availability. Before broader release, the company is hardening sandboxed environments, monitoring internal activations to spot misuse, and strengthening resilience against indirect prompt injection attacks. Wider access will start with Google AI Ultra subscribers and paid API customers.

Benchmark Performance

Independent evaluations · Artificial Analysis

52.6%
Intelligence

Accuracy & Capability Details

Humanity's Last Exam57.1%
SciCode - Scientific Coding61.8%
Long Context Reasoning79.7%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderGoogleAnthropicAnthropic
Release DateSeptember 30, 2026September 22, 2026September 28, 2026
Knowledge Cutoff--Jun 2026
Context & Limits
Context Window-1M1M
Pricing (per 1M tokens)
Input Pricing
$2
Best Input Pricing
$4
$2
Best Input Pricing
Output Pricing
$10
Best Output Pricing
$20
$10
Best Output Pricing
Modalities
Inputs
textimagefilevideo
textimagefile
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index52.6
57.6
Best Intelligence Index
56.0
Coding Index---
Agentic Index---
Gemini 4 Argon
Claude Opus 5.5
Claude Sonnet 5.5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

57%
Gemini 4 Argon
Humanity's Last Exam
Score: 57%
Gemini 4 Argon
61%
Claude Opus 5.5
Humanity's Last Exam
Score: 61%
Claude Opus 5.5
55%
Claude Sonnet 5.5
Humanity's Last Exam
Score: 55%
Claude Sonnet 5.5

Long Context Reasoning

Logical reasoning over long context windows.

80%
Gemini 4 Argon
Long Context Reasoning
Score: 80%
Gemini 4 Argon
85%
Claude Opus 5.5
Long Context Reasoning
Score: 85%
Claude Opus 5.5
83%
Claude Sonnet 5.5
Long Context Reasoning
Score: 83%
Claude Sonnet 5.5

SciCode Benchmark

Scientific coding and mathematical modeling.

62%
Gemini 4 Argon
SciCode Benchmark
Score: 62%
Gemini 4 Argon
67%
Claude Opus 5.5
SciCode Benchmark
Score: 67%
Claude Opus 5.5
61%
Claude Sonnet 5.5
SciCode Benchmark
Score: 61%
Claude Sonnet 5.5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Explore more from Google

Other models by Google

Top AI Models

Leading alternatives by intelligence score

View all