Skip to content
Model Pareto

Compare models

Side by side: quality per category, price, speed, and every benchmark cell with where it came from. Pick up to four.
Bonsai 27BGemini Omni 1.1 FlashGemini Omni FlashWan 3.0Up to 4 models — remove one to add another.

Overview

Lab
PrismML
Google DeepMind
Google DeepMind
Alibaba Qwen
Released
Jul 4, 2026
Aug 27, 2026
May 19, 2026
Aug 19, 2026
Weights
Open
Proprietary
Proprietary
Proprietary
Context
tokens
262K
—
—
—
Price per second
1080p
—
$0.10Best
$0.10Best
$0.20

Quality by category

0–100 within each category (100 = best tracked model). Rank is among all models ranked there. Click a row to focus it.

54.7±10.3
#35/114Lab-reported
Not ranked
Not ranked
Not ranked
56.7±9.7
#58/155Lab-reported
Not ranked
Not ranked
Not ranked
45.1±8.7
#57/90Lab-reported
Not ranked
Not ranked
Not ranked
60.7±9.9
#52/193Lab-reported
Not ranked
Not ranked
Not ranked
Not ranked
94.3±6.3
#2/36Verified
96.9±3.0Best
#1/36Verifiedbest value
93.8±4.1
#3/36Verified

Coding benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Writing & chat benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Vision benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Hard reasoning benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Video generation benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

VBench
day 0
—
—
—
—