Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. GPT-5.4 Mini
OpenAI
Released March 17, 2026Cutoff August 2025

GPT-5.4 Mini

GPT-5.4 Mini by OpenAI is a fast, efficient small model optimized for coding, subagent delegation, computer use, and multimodal reasoning at 400k context.

Visit OpenAIAnnouncement
Inputs
File
Image
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

GPT-5.4 Mini - Distilled Speed for Subagents and Coding Workflows

GPT-5.4 Mini is OpenAI's distilled small model built to bring GPT-5.4's core strengths - coding, tool use, multimodal reasoning, and computer use - to a faster, lower-cost form. It runs more than 2x faster than GPT-5 Mini and approaches GPT-5.4-level performance on several evaluations, making it one of the strongest performance-per-latency tradeoffs in the small-model tier.

Designed for the Subagent Layer

GPT-5.4 Mini is explicitly positioned as the execution layer in multi-model pipelines. In Codex, for example, a larger model like GPT-5.4 handles planning and coordination, while GPT-5.4 Mini runs as a subagent completing narrower parallel tasks - codebase search, file review, document processing. This division is built into how Codex allocates quota: GPT-5.4 Mini uses only 30% of the GPT-5.4 quota per task.

What Separates It

TraitDetail
Computer use at small-model speedInterprets dense UI screenshots quickly; approaches GPT-5.4 on OSWorld-Verified while outperforming GPT-5 Mini substantially
Subagent-native designOfficially described as the target model for parallel subagent execution inside Codex multi-model workflows
400k context windowHandles large codebase navigation and long document tasks at the mini tier
Multimodal reasoningSupports text and image inputs with real-time image interpretation as a documented primary use case
Tool use reliabilitySupports function calling, web search, file search, computer use, and skills natively via the /v1/responses endpoint

Available in Three Surfaces

GPT-5.4 Mini is deployed across the API, Codex (app, CLI, IDE extension, web), and ChatGPT (as the Thinking option for Free and Go users, and as a rate-limit fallback for other tiers). This cross-surface availability is unique among mini-tier models in the GPT-5.4 family - GPT-5.4 Nano, by contrast, is API-only.

Benchmark Performance

Independent evaluations · Artificial Analysis

40.9%
Intelligence
56.1%
Coding Index
31.5%
Agentic Index

Accuracy & Capability Details

GPQA - Graduate Science87.5%
Humanity's Last Exam28.1%
SciCode - Scientific Coding49.9%
Instruction Following73.3%
Long Context Reasoning73.0%
τ²-Bench - Agentic Tasks83.3%
TerminalBench - System Control52.3%
Specs
Context window
400Ktokens
Input pricing
$0.75per 1M tokens
Output pricing
$4.50per 1M tokens
Cached input
$0.07per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Scientific Biology90%
Physics Problems90%
Chemistry Concepts90%
Communication90%
Structured Outputs90%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderOpenAIAnthropicAnthropic
Release DateMarch 17, 2026July 24, 2026June 9, 2026
Knowledge CutoffAug 2025May 2026-
Context & Limits
Context Window400K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$0.75
Best Input Pricing
$5$10
Output Pricing
$4.50
Best Output Pricing
$25$50
Modalities
Inputs
fileimagetext
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index30.5
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
GPT-5.4 Mini
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

82%
GPT-5.4 Mini
GPQA Benchmark
Score: 82%
GPT-5.4 Mini
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

19%
GPT-5.4 Mini
Humanity's Last Exam
Score: 19%
GPT-5.4 Mini
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

62%
GPT-5.4 Mini
Long Context Reasoning
Score: 62%
GPT-5.4 Mini
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models