Anthropic logo

Claude Fable 5.1

Multimodal
Anthropic

Claude Fable 5.1 is Anthropic's generally available, production-safeguarded deployment of the same underlying weights as the trusted-access Claude Mythos 5.1. It targets demanding reasoning and long-horizon agentic work with a 1M-token context window, 128K max output, and adaptive thinking always enabled, and it leads self-reported launch results on Terminal-Bench 4.0 and Humanity's Last Exam. First-party pricing matches Fable 5 at $10/$50 per million tokens with cache reads cut to $0.25 per million.

Key Specifications

Parameters
-
Context
1.0M
Release Date
September 1, 2026
Average Score
59.7%

Timeline

Key dates in the model's history
Announcement
September 1, 2026
Last Update
September 2, 2026
Today
September 3, 2026

Technical Specifications

Parameters
-
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$10.00
Output (per 1M tokens)
$50.00
Max Input Tokens
1.0M
Max Output Tokens
128.0K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Other Tests

Specialized benchmarks
AutomationBench
Fable 5.1 with production safeguards enabled. Zapier AutomationBench private held-out set / business workflows.Self-reported
31.4%
CursorBench v3.2
Fable 5.1 with production safeguards enabled. CursorBench 3.2.0 (catalog id cursorbench-3.2).Self-reported
73.4%
GDPval-AA
Fable 5.1 with production safeguards enabled. GDPval-AA v2 Elo (max_score 3000); not the older GDPval-AA Elo figures from Fable 5.Self-reported
61.8%
Humanity's Last Exam
Fable 5.1 with production safeguards enabled. Multidisciplinary reasoning with tools.Self-reported
65.0%
OSWorld 2.0
Fable 5.1 with production safeguards enabled. OSWorld 2.0 partial score on the authors' August 2026 task release (not comparable to earlier OSWorld numbers). Safeguard interventions scored zero on OSWorld 2.0.Self-reported
77.9%
Terminal-Bench 4.0
Fable 5.1 with production safeguards enabled. Terminal-Bench 4.0 (not Terminal-Bench 2.1). Mythos 5.1 scored 60.9% on the same table.Self-reported
55.8%
Terminal-Bench-Science 0.1
Fable 5.1 with production safeguards enabled. Terminal-Bench-Science 0.1 accuracy; SE ±3.5–4.5 pts. Public leaderboard uses a different setup (3 trials/task, Claude Code harness).Self-reported
52.6%

License & Metadata

License
proprietary
Announcement Date
September 1, 2026
Last Updated
September 2, 2026

Compare Claude Fable 5.1

All comparisons

Articles about Claude Fable 5.1

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.