Skip to content
Model Pareto

GPT-6 Sol

OpenAIProprietaryopenai.com ↗
Input
$2.00/M tok
Output
$10/M tok
Blended
$4.00/M tok · 3:1
Speed
110tok/s
Context
872Ktokens

Rankings

Benchmark scores

~ italic, dashed = no published score yet, estimated from related benchmarks.

Coding

Writing & chat

Vision

Hard reasoning

Agents

~#2 of 116 (estimated)~98.2%
  1. #1Step 3.7 Flash98.5%
  2. ~#2GPT-6 SolEstimated98.2%
  3. #2Gemini 3.1 Pro95.6%
No published score — estimated from related benchmarks
~#3 of 12 (estimated)~82.8%
  1. #1Kimi K384.8%
  2. #2Qwen3.8 27B84.3%
  3. ~#3GPT-6 SolEstimated82.8%
  4. #3Nex-N2.5-Pro82.2%
No published score — estimated from related benchmarks
OSWorld 2.0Estimated
~#3 of 5 (estimated)~54.2%
  1. #1GLM-5.3 Flash59.1%
  2. #2Kimi K358.3%
  3. ~#3GPT-6 SolEstimated54.2%
  4. #3GPT-5.6 Terra50.2%
No published score — estimated from related benchmarks
BrowseCompEstimated
~#7 of 18 (estimated)~88.6%
  1. #1Atria Dawn Preview92.5%
  2. #2GPT-6 Astra91.5%
  3. #6Step 5 Preview88.7%
  4. ~#7GPT-6 SolEstimated88.6%
  5. #7GPT-5.6 Terra87.5%
No published score — estimated from related benchmarks
~#16 of 89 (estimated)~39.5%
  1. #1Muse Spark 1.350.5%
  2. #2GLM-5.350.3%
  3. #15DeepSeek V4 Pro (0813)39.6%
  4. ~#16GPT-6 SolEstimated39.5%
  5. #16GPT-5.539.0%
No published score — estimated from related benchmarks
Not comparable across labs (1)

Kept for reference, not counted in rankings: these use a lab-specific task set, answer key or scoring, so scores can't be compared fairly across labs. Only published scores are shown.