Skip to content
Model Pareto

Compare models

Side by side: quality per category, price, speed, and every benchmark cell with where it came from. Pick up to four.
GPT Image 2Grok Imagine Image 2.0MAI-Image-2.6
Focus

Image generation: quality vs price

Compared models are ringed; the other 35 ranked here are greyed.

Lower price is better. Pareto frontier: Z-Image Turbo, Muse Image, MAI-Image-2.6-Flash, MAI-Image-2.6, GPT Image 2.5 Sunburst. 1 models are not plotted: Qwen-Image-2.1.

1 model with no price data — shown in the strip at the left edge

Qwen-Image-2.1

  • Best-value frontier (nothing is both cheaper and better)
  • Evidencestrong → weak
  • Estimated from other categories
  • GPT Image 2
    Quality
    82.6
    Rank
    #3/38
    Price
    $0.21/img
    Speed
    118s

    vs MAI-Image-2.6: +7.3 quality · 5.4× price

  • Grok Imagine Image 2.0
    Quality
    70.7
    Rank
    #5/38
    Price
    $0.060/img
    Speed
    —

    vs GPT Image 2: −12.0 quality · 0.28× price

  • MAI-Image-2.6
    Quality
    75.3
    Rank
    #4/38
    Price
    $0.039/img
    Speed
    —

    vs GPT Image 2: −7.3 quality · 0.18× price

Overview

Lab
OpenAI
SpaceXAI (xAI)
Microsoft AI
Released
Apr 21, 2026
Aug 7, 2026
Aug 10, 2026
Weights
Proprietary
Proprietary
Proprietary
Price per image
$0.21
$0.060
$0.039Best
Time per image
118s
—
—

Quality by category

0–100 within each category (100 = best tracked model). Rank is among all models ranked there. Click a row to focus it.

82.6±3.2Best
#3/38Verified
70.6±3.2
#5/38Verified
75.3±3.2
#4/38Verifiedbest value

Image generation benchmarks

Rank among models with a published score, and the leaderboard around each model. ~ italic = no published score, estimated from related benchmarks.

GenEval
day 0
~#2/4~88.3%Estimated
estimated from related benchmarks
~#2/4~87.5%Estimated
estimated from related benchmarks
~#2/4~87.7%Estimated
estimated from related benchmarks