Grok 4.6
Grok 4.6 by SpaceXAI: flagship code model with 500K context, configurable reasoning effort, minimal hallucinations, and no realtime access without search
Model Overview
Capabilities, design details, and architectural traits
Grok 4.6 - Flagship Code and Agentic Model with Configurable Reasoning
Grok 4.6 is xAI's flagship model for code and general-purpose tasks, defined by three stated pillars: agentic tool calling, minimal hallucinations, and configurable reasoning.
| Trait | Detail |
|---|---|
| Configurable reasoning | Reasoning depth adjustable via reasoning_effort parameter |
| Context window | 500,000 tokens |
| Minimal hallucinations | Explicit design goal stated in model documentation |
Post-Training Over Scale
Grok 4.6 reuses the same V9 foundation as its predecessor Grok 4.5, with capability gains delivered through improved supervised fine-tuning and reinforcement learning rather than increased parameter scale. The stated goal is matching or exceeding competing frontier models while preserving the inference speed and token efficiency of Grok 4.5. Supplemental training incorporates SpaceX engineering data, excluding ITAR-restricted material.
Benchmark Performance
Independent evaluations · Artificial Analysis
Accuracy & Capability Details
Compare Models Side-by-Side
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | SpaceXAI | Anthropic | Anthropic |
| Release Date | August 12, 2026 | September 22, 2026 | September 28, 2026 |
| Knowledge Cutoff | Feb 2026 | - | Jun 2026 |
| Context & Limits | |||
| Context Window | 500K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $2 Best Input Pricing | $4 | $2 Best Input Pricing |
| Output Pricing | $6 Best Output Pricing | $20 | $10 |
| Modalities | |||
| Inputs | textimagefile | textimagefile | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 44.3 | 57.6 Best Intelligence Index | 56.0 |
| Coding Index | 76.8 | - | - |
| Agentic Index | 53.0 | - | - |
Humanity's Last Exam
Extremely difficult logical reasoning and knowledge.
Long Context Reasoning
Logical reasoning over long context windows.
SciCode Benchmark
Scientific coding and mathematical modeling.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.
Explore more from SpaceXAI
Other models by SpaceXAI
Top AI Models
Leading alternatives by intelligence score