Hermes 3 70B by NousResearch: Llama-3.1 70B fine-tune with native tool-call and scratchpad tags, 131K token context window, steerable instruct tuning.
Capabilities, design details, and architectural traits
Hermes 3 70B is NousResearch's supervised fine-tune of Meta's Llama-3.1 70B focused on steerable instruction behavior and native tool interaction. The model is configured to accept large contexts and explicit in-prompt control tokens for structured tool calls and internal scratchpad reasoning.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Nous Research | Anthropic | Anthropic |
| Release Date | August 15, 2024 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | Dec 2023 | May 2026 | - |
| Context & Limits | |||
| Context Window | 131K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.70 Best Input Pricing | $5 | $10 |
| Output Pricing | $0.70 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | text | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 4.8 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | - | 78.0 Best Coding Index | 76.5 |
| Agentic Index | - | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Scientific coding and mathematical modeling.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.