Llama 4 Maverick by Meta is a 400B MoE model with 128 experts, 17B active parameters, early-fusion multimodality, 1M token context, and single H100 host deployment.
Capabilities, design details, and architectural traits
Llama 4 Maverick is Meta's open-weight general-purpose model in the Llama 4 series. It uses a mixture-of-experts (MoE) architecture with 128 routed experts plus one shared expert, totaling 400B parameters while activating only 17B parameters per token. Each token is routed to the shared expert and exactly one of the 128 routed experts per MoE layer.
Maverick uses alternating dense and MoE layers — not a pure MoE stack — combined with early fusion for native multimodality, integrating text and image inputs at the token level. It is released in both BF16 and FP8 quantized weights, with the FP8 version fitting on a single H100 DGX host.
Maverick supports a 1 million token context window, positioned for applications requiring extensive document or multi-image reasoning within a single host deployment.
Pretrained on approximately 22 trillion tokens of multimodal data, including publicly available content, licensed data, and Meta platform data (Instagram, Facebook, Meta AI interactions).
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Meta | Anthropic | Anthropic |
| Release Date | April 5, 2025 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Aug 2024 | May 2026 | - |
| Context & Limits | |||
| Context Window | 1.0M Best Context Window | 1M | 1M |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.26 Best Input Pricing | $5 | $10 |
| Output Pricing | $0.91 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimage | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 14.5 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 16.3 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 1.2 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.