Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. GPT-5.2
OpenAI
Released December 11, 2025

GPT-5.2

GPT-5.2 by OpenAI is the professional knowledge work frontier model with three variants: Instant, Thinking, and Pro, plus a Codex-optimized coding release.

Visit OpenAIAnnouncement
Inputs
File
Image
Text
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

GPT-5.2 - Frontier Model for Professional Knowledge Work

GPT-5.2 is OpenAI's professional knowledge work flagship, explicitly designed around the economic output of knowledge workers across 44 occupations. It is the first model in the GPT-5 series where OpenAI measured success on GDPval - an eval that scores model outputs against the work of top industry professionals - making occupational-level performance the central release identity rather than abstract benchmark scores.

Three-Variant Structure and the reasoning.effort='none' Default

GPT-5.2 ships as three distinct models with separate API identifiers:

  • gpt-5.2-chat-latest (GPT-5.2 Instant) - fast, designed for info-seeking, how-tos, and everyday writing
  • gpt-5.2 (GPT-5.2 Thinking) - reasoning model for coding, long-document analysis, and multi-step tasks
  • gpt-5.2-pro (GPT-5.2 Pro) - highest compute, maximum accuracy on complex tasks

All three variants default to reasoning.effort='none', a setting introduced with GPT-5.2. This is a documented departure from earlier GPT-5 models, which defaulted to medium. It lowers latency at default and gives developers explicit control to increase reasoning depth as needed.

What Separates GPT-5.2

TraitDetail
GDPval-centered release identityFirst OpenAI model released with professional occupational performance as the primary evaluation axis, covering 44 defined knowledge work roles
reasoning.effort='none' as defaultFirst GPT-5 series model to ship with reasoning off by default, making low-latency the baseline and high-effort reasoning opt-in
/compact endpoint compatibilityThinking variant supports the Responses API /compact endpoint, extending effective context beyond the model's physical context window for tool-heavy, long-running workflows
Dedicated Codex variantReleased a separate GPT-5.2-Codex model optimized specifically for agentic software engineering and defensive cybersecurity workloads
xhigh reasoning effortBoth Thinking and Pro variants are the first to support the fifth reasoning effort level xhigh, in addition to none, low, medium, and high

Benchmark Performance

Independent evaluations · Artificial Analysis

43.3%
Intelligence

Accuracy & Capability Details

GPQA - Graduate Science90.3%
Humanity's Last Exam37.7%
SciCode - Scientific Coding52.1%
Instruction Following75.4%
Long Context Reasoning79.3%
τ²-Bench - Agentic Tasks84.8%
TerminalBench - System Control47.0%
Specs
Context window
400Ktokens
Input pricing
$1.75per 1M tokens
Output pricing
$14per 1M tokens
Cached input
$0.96per 1M tokens

Prices in USD.

Key Capabilities & Ratings

Scientific Biology90%
General Knowledge90%
Physics Problems90%
Translation & Lang90%
Chemistry Concepts90%

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderOpenAIAnthropicAnthropic
Release DateDecember 11, 2025July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window400K
1M
Best Context Window
1M
Best Context Window
Pricing (per 1M tokens)
Input Pricing
$1.75
Best Input Pricing
$5$10
Output Pricing
$14
Best Output Pricing
$25$50
Modalities
Inputs
fileimagetext
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index38.9
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
GPT-5.2
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

86%
GPT-5.2
GPQA Benchmark
Score: 86%
GPT-5.2
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

27%
GPT-5.2
Humanity's Last Exam
Score: 27%
GPT-5.2
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

69%
GPT-5.2
Long Context Reasoning
Score: 69%
GPT-5.2
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models