Claude Opus 5.5
- Input
- $4.00/M tok
- Output
- $20/M tok
- Blended
- $8.00/M tok · 3:1
- Speed
- 92tok/s
- Context
- 1Mtokens
Speed measured on the xhigh effort variant.
Rankings
Coding
Mixed#1 / 103rank by quality96.9±3.4quality / 1005/7 benchmarks measuredPrice frontierSpeed frontierWriting & chat
Verified#3 / 146rank by quality93.0±4.3quality / 1003/5 benchmarks measuredPrice frontierSpeed frontierVision
Verified#2 / 82rank by quality94.5±4.8quality / 1003/6 benchmarks measuredPrice frontierSpeed frontierHard reasoning
Verified#1 / 182rank by quality99.9±2.6quality / 1002/4 benchmarks measuredPrice frontierSpeed frontierAgents
Verified#1 / 173rank by quality99.6±4.1quality / 1002/7 benchmarks measuredPrice frontierSpeed frontier
Benchmark scores
~ italic, dashed = no published score yet, estimated from related benchmarks.
Coding
SWE-bench VerifiedEstimated
~#2 of 14 (estimated)~94.7%
No published score — estimated from related benchmarks
LiveCodeBenchEstimated
~#2 of 20 (estimated)~92.3%
No published score — estimated from related benchmarks
Writing & chat
#7 of 252050 Elo
eqbench.com · 2026-09-24
Vision
CharXiv ReasoningEstimated
~#2 of 9 (estimated)~86.1%
No published score — estimated from related benchmarks
OmniDocBench v1.5Estimated
~#2 of 8 (estimated)~91.4%
No published score — estimated from related benchmarks
GDP.pdf (AA)Verified
#3 of 4226.2%
- #1GPT-6 Astra31.0%
- #2Muse Spark 1.326.6%
- #3Claude Fable 5.126.2%
- #3Claude Opus 5.526.2%
- #5GPT-6 Sol24.8%
artificialanalysis.ai · 2026-09-24
Hard reasoning
GPQA DiamondEstimated
~#2 of 173 (estimated)~96.1%
No published score — estimated from related benchmarks
AIME (latest)Estimated
~#2 of 25 (estimated)~96.7%
No published score — estimated from related benchmarks
Agents
τ²-bench (Telecom)Estimated
~#2 of 116 (estimated)~98.3%
No published score — estimated from related benchmarks
OSWorld-VerifiedEstimated
~#2 of 12 (estimated)~84.6%
No published score — estimated from related benchmarks
OSWorld 2.0Estimated
~#2 of 5 (estimated)~58.7%
No published score — estimated from related benchmarks
BrowseCompEstimated
~#2 of 18 (estimated)~92.3%
No published score — estimated from related benchmarks
τ-Bench Banking (AA)Estimated
~#3 of 89 (estimated)~50.0%
No published score — estimated from related benchmarks
Not comparable across labs (1)
Kept for reference, not counted in rankings: these use a lab-specific task set, answer key or scoring, so scores can't be compared fairly across labs. Only published scores are shown.