Skip to content
Model Pareto

GPT-6.1 Sol

OpenAIProprietaryopenai.com ↗
Input
$2.00/M tok
Output
$10/M tok
Blended
$4.00/M tok · 3:1
Speed
~69tok/sest.
Context
922Ktokens

Context: 1.05M total, 922k max input (1.05M minus 128k output), same basis as gpt-6-sol's 872k. OpenAI says a faster "Ultrafast" serving tier of Sol is coming soon (up to 8x faster tokens in Codex at roughly 6x the price; https://thenextweb.com/news/openai-gpt-6-1-sol-price-astra-devday); same weights, so it will be a serving tier rather than a separate model.

Rankings

Benchmark scores

~ italic, dashed = no published score yet, estimated from related benchmarks.

Coding

Writing & chatEstimated from other categories

IFBenchEstimated
~#3 of 126 (estimated)~82.5%
  1. #1MiniMax-M382.9%
  2. #2Qwen3.8 2.4T A95B82.8%
  3. ~#3GPT-6.1 SolEstimated82.5%
  4. #3Qwen3.8-Omni-Flash81.5%
No published score — estimated from related benchmarks
MMMLUEstimated
~#2 of 5 (estimated)~90.4%
  1. #1Gemini 3.1 Pro92.6%
  2. ~#2GPT-6.1 SolEstimated90.4%
  3. #2Gemma 4 31B88.4%
No published score — estimated from related benchmarks
~#2 of 54 (estimated)~1499 Elo
  1. #1Claude Opus 5.51501
  2. ~#2GPT-6.1 SolEstimated1499
  3. #2Claude Fable 5.1Few votes1496
No published score — estimated from related benchmarks
~#2 of 54 (estimated)~1507 Elo
  1. #1Claude Opus 5.51509
  2. ~#2GPT-6.1 SolEstimated1507
  3. #2Claude Fable 5.11501
No published score — estimated from related benchmarks
~#7 of 26 (estimated)~2063 Elo
  1. #1GPT-6 Astra2173
  2. #2Claude Fable 5.12162
  3. #6GLM-5.32075
  4. ~#7GPT-6.1 SolEstimated2063
  5. #7Claude Opus 5.52050
No published score — estimated from related benchmarks

Vision

MMMUEstimated
~#2 of 3 (estimated)~71.3%
  1. #1Claude Haiku 4.573.2%
  2. ~#2GPT-6.1 SolEstimated71.3%
  3. #2Llama 4 Maverick69.4%
No published score — estimated from related benchmarks
~#2 of 9 (estimated)~85.7%
  1. #1Gemini 3.8 Flash86.2%
  2. ~#2GPT-6.1 SolEstimated85.7%
  3. #2Kimi K384.8%
No published score — estimated from related benchmarks
~#4 of 8 (estimated)~90.6%
  1. #1MiniMax-M391.6%
  2. #2Kimi K391.1%
  3. #2Qwen3.8 27B91.1%
  4. ~#4GPT-6.1 SolEstimated90.6%
  5. #4Gemini 3.1 Pro85.3%
No published score — estimated from related benchmarks
GDP.pdf (AA)Estimated
~#9 of 43 (estimated)~21.9%
  1. #1GPT-6 Astra31.0%
  2. #2Muse Spark 1.326.6%
  3. #8Kimi K322.0%
  4. ~#9GPT-6.1 SolEstimated21.9%
  5. #9Claude Opus 521.6%
No published score — estimated from related benchmarks
ChartographyEstimated
~#5 of 23 (estimated)~41.3%
  1. #1GPT-6 Astra71.0%
  2. #2Claude Opus 5.566.3%
  3. #4Claude Fable 5.146.2%
  4. ~#5GPT-6.1 SolEstimated41.3%
  5. #5Gemini 3.8 Flash40.9%
No published score — estimated from related benchmarks

Hard reasoning

AgentsEstimated from other categories

~#2 of 116 (estimated)~97.8%
  1. #1Step 3.7 Flash98.5%
  2. ~#2GPT-6.1 SolEstimated97.8%
  3. #2Gemini 3.1 Pro95.6%
No published score — estimated from related benchmarks
~#3 of 12 (estimated)~83.4%
  1. #1Kimi K384.8%
  2. #2Qwen3.8 27B84.3%
  3. ~#3GPT-6.1 SolEstimated83.4%
  4. #3Nex-N2.5-Pro82.2%
No published score — estimated from related benchmarks
OSWorld 2.0Estimated
~#3 of 5 (estimated)~53.8%
  1. #1GLM-5.3 Flash59.1%
  2. #2Kimi K358.3%
  3. ~#3GPT-6.1 SolEstimated53.8%
  4. #3GPT-5.6 Terra50.2%
No published score — estimated from related benchmarks
BrowseCompEstimated
~#5 of 18 (estimated)~90.5%
  1. #1Atria Dawn Preview92.5%
  2. #2GPT-6 Astra91.5%
  3. #4Claude Opus 590.8%
  4. ~#5GPT-6.1 SolEstimated90.5%
  5. #5Nex-N2.5-Pro89.7%
No published score — estimated from related benchmarks
~#12 of 87 (estimated)~1612 Elo
  1. #1Claude Opus 5.51846
  2. #2Claude Fable 5.11735
  3. #11Qwen3.8 Flash-Next1612
  4. ~#12GPT-6.1 SolEstimated1612
  5. #12DeepSeek V4.1 Flash1600
No published score — estimated from related benchmarks
~#5 of 80 (estimated)~65.7%
  1. #1Claude Opus 5.569.5%
  2. #2DeepSeek V4.1 Flash68.9%
  3. #4Grok 4.667.0%
  4. ~#5GPT-6.1 SolEstimated65.7%
  5. #5Grok 4.765.6%
No published score — estimated from related benchmarks
~#9 of 90 (estimated)~45.6%
  1. #1Muse Spark 1.350.5%
  2. #2GLM-5.350.3%
  3. #8Kimi K346.0%
  4. ~#9GPT-6.1 SolEstimated45.6%
  5. #9Qwen3.8 Flash-Next45.4%
No published score — estimated from related benchmarks
Not comparable across labs (1)

Kept for reference, not counted in rankings: these use a lab-specific task set, answer key or scoring, so scores can't be compared fairly across labs. Only published scores are shown.