Skip to content
Model Pareto

Compare models

Side by side: quality per category, price, speed, and every benchmark cell with where it came from. Pick up to four.
Claude Haiku 5.5Wan 3.0Gemini Omni FlashMiniMax H3Up to 4 models — remove one to add another.

Overview

Lab
Anthropic
Alibaba Qwen
Google DeepMind
MiniMax
Released
Oct 7, 2026
Aug 19, 2026
May 19, 2026
Jul 30, 2026
Weights
Proprietary
Proprietary
Proprietary
Open
Input price
$/M tok
$0.10
—
—
—
Output price
$/M tok
$0.50
—
—
—
Blended price
$/M tok · 3:1 in:out
$0.20
—
—
—
Context
tokens
1M
—
—
—
Price per second
1080p
—
$0.20
$0.10Best
$0.13
Render time
per clip
—
—
—
121s

Quality by category

0–100 within each category (100 = best tracked model). Rank is among all models ranked there. Click a row to focus it.

66.2±8.0
#19/119Verified
Not ranked
Not ranked
Not ranked
78.7±9.7
#7/91Lab-reportedbest value
Not ranked
Not ranked
Not ranked
81.3±5.1
#17/200Verifiedbest value
Not ranked
Not ranked
Not ranked
Not ranked
93.1±4.1Best
#1/36Verifiedbest value
92.9±3.7
#2/36Verifiedbest value
92.9±3.0
#3/36Verified

Coding benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Vision benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Hard reasoning benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

Video generation benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

VBench
day 0
—
—
—
—