Skip to content
Model Pareto

Compare models

Side by side: quality per category, price, speed, and every benchmark cell with where it came from. Pick up to four.
Claude Opus 5.5Granite 4.2 3BNano Banana 2 Lite (Gemini 3.1 Flash Lite Image)

Overview

Lab
Anthropic
IBM Granite
Google DeepMind
Released
Sep 22, 2026
Aug 25, 2026
Jun 30, 2026
Weights
Proprietary
Open
Proprietary
Input price
$/M tok
$4.00
$0.03Best
—
Output price
$/M tok
$20
$0.12Best
—
Blended price
$/M tok · 3:1 in:out
$8.00
$0.052Best
—
Output speed
tok/s
92
220Best
—
Context
tokens
1MBest
131K
—
Price per image
—
—
$0.034
Time per image
—
—
3.1s

Quality by category

0–100 within each category (100 = best tracked model). Rank is among all models ranked there. Click a row to focus it.

96.9±3.4Best
#1/103Mixedbest value
17.2±8.0
#95/103Verifiedbest value
Not ranked
93.0±4.3
#3/146Verifiedbest value
Not ranked
Not ranked
94.5±4.8
#2/82Verifiedbest value
Not ranked
Not ranked
99.9±2.6Best
#1/182Verifiedbest value
23.4±3.7
#128/182Verified
Not ranked
99.6±4.1Best
#1/173Verifiedbest value
17.0±5.8
#140/173Verified
Not ranked
Not ranked
Not ranked
47.9±3.2
#19/38Verified

Coding benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

~last/15below measured range (<41.1)Estimated
estimated from related benchmarks
—

Writing & chat benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Vision benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Hard reasoning benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Agents benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

~#2/12~84.6%Estimated
estimated from related benchmarks
—
~#2/5~58.7%Estimated
estimated from related benchmarks
~last/5below measured range (<47.5%)Estimated
estimated from related benchmarks
—
Not comparable across labs (1)

Image generation benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.