Model
Model explorer

Gemini 3.5 Flash

CLOSED
Google · Gemini 3.5 family · released May 19, 2026

First Flash to beat the prior-gen Pro flagship; I/O 2026 headline, 4x faster output than rival frontier models.

ReasoningCodingVisionFunction callingTool useAgentic
2498.9
Elo · rank #31
Parameters
Undisclosed
Active params
Undisclosed
Context
1M tokens
Architecture
Sparse Mixture-of-Experts Transformer (configuration undisclosed)
License
Proprietary
Languages
API price (in/out)
$1.5 / $9
Modalities
text · vision · audio · video
Benchmark results
Bar shows position within the tracked field; marker = field best
AIMEMath95.0%#17
best: GPT-5.2 · 100.0%
ARC-AGI-2Reasoning72.1%#10
best: GPT-6 Astra · 95.0%
CharXivVision84.2%#8
best: Qwen3.8-Flash-Next · 90.6%
GDPval-AAAgents1656#10
best: Claude Fable 5 · 1932
GPQA DiamondReasoning92.2%#20
best: GPT-6 Astra · 96.0%
Humanity's Last ExamReasoning40.2%#29
best: Claude Opus 5 · 64.7%
IFBenchReasoning76.3%#15
best: MiniMax M3 · 83.0%
LiveCodeBenchCoding87.6%#16
best: DeepSeek-V4-Pro (Think Max) · 93.5%
MMMU-ProVision83.6%#2
best: Claude Opus 4.7 · 85.5%
MMMUVision83.0%#14
best: Claude Fable 5 · 89.3%
OSWorld-VerifiedAgents78.4%#8
best: Claude Fable 5 · 85.0%
SWE-bench ProCoding55.1%#35
best: Claude Fable 5.1 · 81.2%
SWE-bench VerifiedCoding78.8%#24
best: Claude Opus 5 · 96.0%
Terminal-Bench 2.0Coding76.2%#21
best: Gemini 3.8 Flash · 89.4%
Run it locally
Closed weights — available via API only. No local deployment.
Input / M tok
$1.5
Output / M tok
$9
API price $1.5/$9 · each benchmark row carries its own source badge (see methodology)