Model
Model explorer
Gemini 3 Flash
CLOSEDGoogle · Gemini 3 family · released Dec 17, 2025
Pro-grade reasoning at Flash latency/cost via an adjustable thinking_level parameter; default model in the Gemini app.
ReasoningCodingVisionFunction callingTool useAgentic
2326.6
Elo · rank #48
Parameters
Undisclosed
Active params
Undisclosed
Context
1M tokens
Architecture
Sparse Mixture-of-Experts Transformer (natively multimodal, based on Gemini 3 Pro)
License
Proprietary (Google API Terms)
Languages
—
API price (in/out)
$0.5 / $3
Modalities
text · vision · audio · video
Benchmark results
Bar shows position within the tracked field; marker = field best
best: GPT-5.2 · 100.0%
best: GPT-6 Astra · 95.0%
best: Qwen3.8-Flash-Next · 90.6%
best: Gemini 2.5 Flash-Lite (Thinking) · 86.8%
best: GPT-6 Astra · 96.0%
best: Claude Opus 5 · 64.7%
best: Gemini 3.1 Pro · 2887
best: DeepSeek-V4-Pro (Think Max) · 93.5%
best: OpenAI o3 · 92.9%
best: Claude Opus 4.7 · 85.5%
best: Claude Opus 5 · 96.0%
best: Gemini 3.8 Flash · 89.4%
Run it locally
Closed weights — available via API only. No local deployment.
Input / M tok
$0.5
Output / M tok
$3
Gemini 3 family
Elo progression across releases
API price $0.5/$3 · each benchmark row carries its own source badge (see methodology)