Model
Model explorer

DeepSeek-V4-Pro (Think Max)

OPEN
DeepSeek · DeepSeek-V4 family · released Apr 24, 2026

Fast/low-latency mode of the DeepSeek-V4-Pro preview checkpoint, the default reasoning tier a user gets without selecting anything. (maximum reasoning-effort tier / 'Pro-Max' mode, DeepSeek's most capable released configuration).

ReasoningCodingVisionFunction callingTool useAgentic
2487.3
Elo · rank #33
Parameters
1600B
Active params
49B (MoE)
Context
1M tokens
Architecture
Mixture-of-Experts (1.6T total / 49B active), hybrid Compressed Sparse Attention + Heavily Compressed Attention, Manifold-Constrained Hyper-Connections, Muon optimizer
License
MIT
Languages
API price (in/out)
$0.435 / $0.87
Modalities
text
Benchmark results
Bar shows position within the tracked field; marker = field best
ARC-AGI-2Reasoning46.0%#15
best: GPT-6 Astra · 95.0%
BrowseCompAgents83.4%#14
best: Kimi K3 · 91.2%
CodeforcesCoding3206#1
best: this model · 3206
GDPval-AAAgents1554#17
best: Claude Fable 5 · 1932
GPQA DiamondReasoning90.1%#30
best: GPT-6 Astra · 96.0%
Humanity's Last ExamReasoning37.7%#32
best: Claude Opus 5 · 64.7%
LiveCodeBenchCoding93.5%#1
best: this model · 93.5%
MMLU-ProKnowledge87.5%#9
best: Claude Fable 5 · 91.5%
SWE-bench ProCoding55.4%#34
best: Claude Fable 5.1 · 81.2%
SWE-bench VerifiedCoding80.6%#13
best: Claude Opus 5 · 96.0%
Terminal-Bench 2.0Coding67.9%#31
best: Gemini 3.8 Flash · 89.4%
Run it locally
VRAM @ Q4
VRAM @ FP16
Fits on (Q4)
Multi-node cluster required
Throughput data unavailable.
Quantizations
Fine-tune it
Permissive
QLoRA1088.0 GB8× H200 141GB
LoRA3408.0 GBbeyond 8× B200
Full fine-tune25680.0 GBbeyond 8× B200
QLoRA SFT on ~10k samples ≈ $25.34 (8× H200 141GB)
API price $0.435/$0.87 · each benchmark row carries its own source badge (see methodology)