Model
Model explorer

Claude Opus 5

CLOSED
Anthropic · Claude 5 family · released Jul 24, 2026

Model id claude-opus-5 (released 2026-07-24, Anthropic's fourth Claude 5 model). Positioned as a thoughtful, proactive agent that 'comes close to the frontier intelligence of Claude Fable 5 at half the price' — same $5/$25 per MTok as Opus 4.8 (half of Fable 5's $10/$50); it is the new default model on Claude Max and the strongest model on Claude Pro. familySlug is claude-5, but predecessor is null: it is the first Opus in the Claude 5 family and cross-family predecessors are disallowed (Opus 4 likewise had a null predecessor at the start of Claude 4). Fast mode runs ~2.5x default speed at twice the base price ($10/$50). Per Anthropic's Opus 5 System Card (Table 8.1.A) it is state-of-the-art on FrontierBench v0.1 and GDPval-AA v2, roughly triples the next-best model on ARC-AGI-3, and ~1.5x the next-best on Zapier AutomationBench; it trails GPT-5.6 Sol on DeepSWE and ARC-AGI-2, and Fable 5 on SWE-bench Pro and FrontierCode. HLE headline uses the with-tools configuration (64.7%; no-tools is 56.3%) to stay ordering-consistent with the sibling Sonnet 5 row, whose corpus HLE figure (57.4%) is also with-tools. GDPval-AA v2 is deliberately NOT filed for this model: Opus 5's launch figure (1861) is on the current, re-fitted Artificial Analysis leaderboard (where Fable 5=1747, Opus 4.8=1593), whereas this corpus's gdpval-aa column is an earlier snapshot (Fable 5=1932, Opus 4.8=1890); mixing the two scales would make Opus 5 spuriously lose its GDPval-AA battles to Opus 4.8/Fable 5, contradicting the published SOTA result, so the row is omitted rather than corrupt the Elo. Cyber classifiers engage ~85% less often than Fable 5; flagged cyber/bio requests fall back to Opus 4.8 by default.

ReasoningCodingVisionFunction callingTool useAgentic
3290.9
Elo · rank #1
Parameters
Undisclosed
Active params
Context
1M tokens
Architecture
Dense transformer, architecture undisclosed; adaptive thinking on by default, controlled by a runtime effort parameter (low/medium/high/xhigh/max)
License
Proprietary
Languages
API price (in/out)
$5 / $25
Modalities
text · vision
Benchmark results
Bar shows position within the tracked field; marker = field best
AA-BriefcaseAgents1720#1
best: this model · 1720
ARC-AGI-1Reasoning97.5%#1
best: this model · 97.5%
ARC-AGI-2Reasoning90.4%#2
best: GPT-5.6 Sol · 92.5%
ARC-AGI-3Reasoning30.2%#1
best: this model · 30.2%
AutomationBenchAgents26.0%#1
best: this model · 26.0%
BrowseCompAgents90.8%#2
best: Kimi K3 · 91.2%
DeepSWECoding68.8%#3
best: GPT-5.6 Sol · 72.7%
FrontierBenchAgents43.3%#1
best: this model · 43.3%
FrontierCodeCoding53.4%#2
best: Claude Fable 5 · 53.5%
Humanity's Last ExamReasoning64.7%#1
best: this model · 64.7%
OSWorld 2.0Agents70.6%#1
best: this model · 70.6%
best: this model · 89.5%
best: this model · 59.4%
SWE-bench ProCoding79.2%#2
best: Claude Fable 5 · 80.0%
SWE-bench VerifiedCoding96.0%#1
best: this model · 96.0%
Run it locally
Closed weights — available via API only. No local deployment.
Input / M tok
$5
Output / M tok
$25
API price $5/$25 · each benchmark row carries its own source badge (see methodology)