Claude Fable 5.1
- Input
- $10/M tok
- Output
- $50/M tok
- Blended
- $20/M tok · 3:1
- Speed
- 66tok/s
- Context
- 1Mtokens
Rankings
Coding
Mixed#3 / 103rank by quality87.6±3.7quality / 1005/7 benchmarks measuredWriting & chat
Verified#1 / 146rank by quality97.6±3.5quality / 1003/5 benchmarks measuredPrice frontierSpeed frontierVision
Verified#3 / 82rank by quality87.4±7.3quality / 1002/6 benchmarks measuredHard reasoning
Verified#2 / 182rank by quality95.7±3.7quality / 1003/4 benchmarks measuredAgents
Verified#2 / 173rank by quality89.6±5.8quality / 1003/7 benchmarks measured
Benchmark scores
~ italic, dashed = no published score yet, estimated from related benchmarks.
Coding
SWE-bench VerifiedEstimated
~#2 of 14 (estimated)~93.5%
No published score — estimated from related benchmarks
DeepSWE v1.1Verified
#16 of 3064.3%
- #1DeepSeek V4.1 Flash74.2%
- #2Grok 4.772.6%
- #15Grok 4.664.9%
- #16Claude Fable 5.164.3%
- #16Hy4 preview64.3%
artificialanalysis.ai · 2026-09-24
Terminal-Bench 4.0Verified
#4 of 8352.0%
- #1Claude Sonnet 5.563.6%
- #2Claude Opus 5.559.6%
- #3GPT-6 Astra59.1%
- #4Claude Fable 5.152.0%
- #5Claude Opus 549.0%
artificialanalysis.ai · 2026-09-24
LiveCodeBenchEstimated
~#2 of 20 (estimated)~92.2%
No published score — estimated from related benchmarks
Writing & chat
#2 of 252162 Elo
eqbench.com · 2026-09-24
Vision
CharXiv ReasoningEstimated
~#2 of 9 (estimated)~86.0%
No published score — estimated from related benchmarks
OmniDocBench v1.5Estimated
~#2 of 8 (estimated)~91.2%
No published score — estimated from related benchmarks
ChartographyVerified
#4 of 2246.2%
- #1GPT-6 Astra71.0%
- #2Claude Opus 5.566.3%
- #3GPT-6 Sol53.6%
- #4Claude Fable 5.146.2%
- #5Gemini 3.8 Flash40.9%
surgehq.ai · 2026-09-24
Hard reasoning
AIME (latest)Estimated
~#2 of 25 (estimated)~96.9%
No published score — estimated from related benchmarks
Agents
τ²-bench (Telecom)Estimated
~#2 of 116 (estimated)~98.3%
No published score — estimated from related benchmarks
OSWorld-VerifiedEstimated
~#2 of 12 (estimated)~84.6%
No published score — estimated from related benchmarks
OSWorld 2.0Estimated
~#3 of 5 (estimated)~58.0%
No published score — estimated from related benchmarks
BrowseCompEstimated
~#2 of 18 (estimated)~92.2%
No published score — estimated from related benchmarks
Not comparable across labs (1)
Kept for reference, not counted in rankings: these use a lab-specific task set, answer key or scoring, so scores can't be compared fairly across labs. Only published scores are shown.