Toolbit.aiToolbit.ai

Find, compare, and explore the best AI tools to match your specific tasks and use cases.

Explore

  • AI Search
  • Compare ToolsNew
  • Browse Categories
  • Trending Tools
  • Most Popular
  • New Additions

Resources

  • Updates HubNew
  • AI News
  • ModelsNew
  • Blog Articles
  • NewsletterNew

Company

  • Launch a Tool
  • Advertise with Us
  • Guest Post
  • Contact Us
© 2026 Toolbit.ai. All rights reserved.
Privacy PolicyTerms & ConditionsDisclaimer
There's An AI For That favicon
There's An AI For That•The front page of AI for everyone
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Toolbit.ai
Toolbit.ai
Toolbit.ai
Toolbit.ai
UpdatesNew
Blog
Sign in
  1. Home
  2. Updates
  3. Models
  4. Gemini 3 Flash Preview
Google
Released December 17, 2025

Gemini 3 Flash Preview

Google Gemini 3 Flash Preview is a Google model for multimodal understanding, agentic and vibe-coding workflows, and structured tool use.

Visit GoogleAnnouncement
Inputs
Text
Image
File
Audio
Video
Outputs
Text

Model Overview

Capabilities, design details, and architectural traits

Gemini 3 Flash Preview - multimodal agentic Flash model

Gemini 3 Flash Preview is built for multimodal understanding and fast interactive work. It is positioned as a Flash model with strong agentic behavior and vibe-coding use.

Distinctive traitWhat is documented
Multimodal focusDescribed as a model for multimodal understanding.
Agentic and vibe-coding useCalled a powerful agentic and vibe-coding model.
Interactive visual workPositioned for richer visuals and deeper interactivity.
Thinking and toolsSupports thinking, code execution, computer use, file search, function calling, grounding, structured outputs, and URL context.

Model behavior

The model is described as built on a foundation of state-of-the-art reasoning. Its documented use pattern combines multimodal input with tool use for interactive tasks.

Supported inputs

The model accepts text, image, video, audio, and PDF inputs. The documentation also lists caching and Batch API support, which fits its preview role as a developer-facing Flash model.

Benchmark Performance

Independent evaluations · Artificial Analysis

38.7%
Intelligence

Accuracy & Capability Details

GPQA - Graduate Science89.8%
Humanity's Last Exam36.6%
SciCode - Scientific Coding50.6%
Instruction Following78.0%
Long Context Reasoning73.0%
τ²-Bench - Agentic Tasks80.4%
TerminalBench - System Control38.6%
Specs
Context window
1.0Mtokens
Input pricing
$0.50per 1M tokens
Output pricing
$3per 1M tokens
Cached input
$0.05per 1M tokens

Prices in USD.

Compare Models Side-by-Side

Evaluate specifications, pricing, and independent benchmark indices

Model Details
General Info
ProviderGoogleAnthropicAnthropic
Release DateDecember 17, 2025July 24, 2026June 9, 2026
Knowledge Cutoff-May 2026-
Context & Limits
Context Window
1.0M
Best Context Window
1M1M
Pricing (per 1M tokens)
Input Pricing
$0.50
Best Input Pricing
$5$10
Output Pricing
$3
Best Output Pricing
$25$50
Modalities
Inputs
textimagefileaudiovideo
textimage
textimagefile
Outputs
text
text
text
Benchmarks (0-100)
Intelligence Index27.9
63.1
Best Intelligence Index
62.1
Coding Index-
78.0
Best Coding Index
76.5
Agentic Index-
59.2
Best Agentic Index
56.6
Gemini 3 Flash Preview
Claude Opus 5
Claude Fable 5

GPQA Benchmark

Graduate-level reasoning and expert Q&A evaluation.

81%
Gemini 3 Flash Preview
GPQA Benchmark
Score: 81%
Gemini 3 Flash Preview
93%
Claude Opus 5
GPQA Benchmark
Score: 93%
Claude Opus 5
93%
Claude Fable 5
GPQA Benchmark
Score: 93%
Claude Fable 5

Humanity's Last Exam

Extremely difficult logical reasoning and knowledge.

15%
Gemini 3 Flash Preview
Humanity's Last Exam
Score: 15%
Gemini 3 Flash Preview
55%
Claude Opus 5
Humanity's Last Exam
Score: 55%
Claude Opus 5
56%
Claude Fable 5
Humanity's Last Exam
Score: 56%
Claude Fable 5

Long Context Reasoning

Logical reasoning over long context windows.

53%
Gemini 3 Flash Preview
Long Context Reasoning
Score: 53%
Gemini 3 Flash Preview
76%
Claude Opus 5
Long Context Reasoning
Score: 76%
Claude Opus 5
77%
Claude Fable 5
Long Context Reasoning
Score: 77%
Claude Fable 5

Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.

Back to all models