Model
Model explorer

Claude Fable 5.1

CLOSED
Anthropic · Claude 5 family · released Sep 1, 2026

Anthropic's September 1, 2026 Mythos-class flagship, shipping as `claude-fable-5-1` on the Claude API, AWS, Google Cloud and Microsoft Foundry. List price is unchanged from Fable 5 at $10/$50 per Mtok — the change is cache reads, cut 75% to $0.25/Mtok, which Anthropic measures as roughly 25% lower cost on typical workloads and up to ~45% on context-heavy agentic ones. Defaults to High effort in Claude Code and Medium in Claude Cowork and on claude.ai. Fable 5.1 and Mythos 5.1 are the SAME underlying model with different safeguard levels; this row is the generally-available one. The gains concentrate in long, tool-using work — Terminal-Bench-Science more than doubles (24.7 → 52.6) and AutomationBench nearly doubles (17.1 → 31.4) — while already-strong short coding moves ~3 points (CursorBench 70.5 → 73.4). PROVENANCE — the launch table is vendor-run with production safeguards ON, which Anthropic says likely depresses several rows (both Fable 5.1 and Fable 5 scored zero on the tasks where safeguards intervened on OSWorld 2.0, as did Fable 5 on AutomationBench). Anthropic publishes rows it loses: Fable 5 beats it on ARC-AGI-1 (98.5 vs 97.5), GPT-5.6 Sol leads ARC-AGI-2 (92.5 vs 90.0), and Opus 5 leads SWE-bench Multimodal and HealthBench Professional. It headlines NO SWE-bench Verified figure — the 95.0% widely re-attributed to it is Fable 5's June number. Independent checks agree on direction: Artificial Analysis ranks it first on Intelligence Index v4.1.1 (66), Vals AI first on the Vals Index (67.87%), and ARC Prize ran ARC-AGI-1/2 on its own semi-private set. OSWorld 2.0 here is the benchmark authors' August 2026 task release, which Anthropic says is not comparable with earlier OSWorld 2.0 publications.

ReasoningCodingVisionFunction callingTool useAgentic
3144.2
Elo · rank #3
Parameters
Undisclosed
Active params
Context
1M tokens
Architecture
Proprietary frontier model (architecture and size undisclosed); adaptive thinking with selectable effort (low → max), 1M-token context, text+image in / text out
License
Proprietary
Languages
API price (in/out)
$10 / $50
Modalities
text · vision
Benchmark results
Bar shows position within the tracked field; marker = field best
AA-BriefcaseAgents1694#2
best: Claude Opus 5 · 1720
ARC-AGI-1Reasoning97.5%#2
best: GPT-6 Astra · 98.5%
ARC-AGI-2Reasoning90.0%#4
best: GPT-6 Astra · 95.0%
AutomationBenchAgents31.4%#6
best: Qwen3.8-Max-0902 · 50.8%
best: GPT-6 Astra · 95.9%
CursorBenchCoding73.4%#1
best: this model · 73.4%
DeepSWECoding67.4%#6
best: Muse Spark 1.3 · 75.4%
best: Claude Fable 5 · 64.9%
FrontierCodeCoding50.9%#4
best: Claude Fable 5 · 53.5%
GDPval-AAAgents1853#3
best: Claude Fable 5 · 1932
GPQA DiamondReasoning93.7%#6
best: GPT-6 Astra · 96.0%
Humanity's Last ExamReasoning60.9%#2
best: Claude Opus 5 · 64.7%
OSWorld 2.0Agents77.9%#1
best: this model · 77.9%
best: Claude Opus 5 · 89.5%
best: Claude Opus 5 · 59.4%
SWE-bench ProCoding81.2%#1
best: this model · 81.2%
Terminal-Bench 4.0Agents55.8%#3
best: Claude Mythos 5.1 · 60.9%
best: GPT-6 Astra · 64.6%
Run it locally
Closed weights — available via API only. No local deployment.
Input / M tok
$10
Output / M tok
$50
API price $10/$50 · each benchmark row carries its own source badge (see methodology)