Skip to content
Model Pareto

Grok 4.6

SpaceXAI (xAI)Proprietaryx.ai ↗
Input
$2.00/M tok
Output
$6.00/M tok
Blended
$3.00/M tok · 3:1
Speed
60tok/s
Context
500Ktokens

Previous generation.

Rankings

Benchmark scores

~ italic, dashed = no published score yet, estimated from related benchmarks.

Coding

Writing & chat

Vision

MMMUEstimated
~#2 of 3 (estimated)~71.3%
  1. #1Claude Haiku 4.573.2%
  2. ~#2Grok 4.6Estimated71.3%
  3. #2Llama 4 Maverick69.4%
No published score — estimated from related benchmarks
MMMU-ProEstimated
~#10 of 77 (estimated)~81.5%
  1. #1Claude Opus 5.587.7%
  2. #2GPT-6 Astra86.9%
  3. #9Intern-S2-397B81.7%
  4. ~#10Grok 4.6Estimated81.5%
  5. #10GPT-5.6 Terra80.7%
No published score — estimated from related benchmarks
~#5 of 9 (estimated)~81.8%
  1. #1Gemini 3.8 Flash86.2%
  2. #2Kimi K384.8%
  3. #4Qwen3.8-Omni-Flash83.5%
  4. ~#5Grok 4.6Estimated81.8%
  5. #5Muse Glimmer78.8%
No published score — estimated from related benchmarks
~#4 of 8 (estimated)~88.2%
  1. #1MiniMax-M391.6%
  2. #2Kimi K391.1%
  3. #2Qwen3.8 27B91.1%
  4. ~#4Grok 4.6Estimated88.2%
  5. #4Gemini 3.1 Pro85.3%
No published score — estimated from related benchmarks

Hard reasoning

Agents