Grok 4.7
- Input
- $2.00/M tok
- Output
- $6.00/M tok
- Blended
- $3.00/M tok · 3:1
- Speed
- 40tok/s
- Context
- 500Ktokens
Rankings
Coding
Verified#11 / 103rank by quality72.7±4.5quality / 1004/7 benchmarks measuredWriting & chat
Verified#25 / 146rank by quality70.9±4.3quality / 1003/5 benchmarks measuredVision
Verified#23 / 82rank by quality64.1±7.3quality / 1002/6 benchmarks measuredHard reasoning
Verified#14 / 182rank by quality81.8±5.1quality / 1002/4 benchmarks measuredAgents
Verified#5 / 173rank by quality86.7±7.8quality / 1002/7 benchmarks measured
Benchmark scores
~ italic, dashed = no published score yet, estimated from related benchmarks.
Coding
SWE-bench VerifiedEstimated
~#2 of 14 (estimated)~85.7%
No published score — estimated from related benchmarks
SWE-bench Pro (Public)Estimated
~#6 of 19 (estimated)~68.1%
- #1Claude Opus 5.589.9%
- #2Claude Sonnet 5.581.3%
- #5Intern-S2-397B68.5%
- ~#6Grok 4.7Estimated68.1%
- #6Hy4 preview65.7%
No published score — estimated from related benchmarks
Terminal-Bench 4.0Verified
#15 of 8325.8%
- #1Claude Sonnet 5.563.6%
- #2Claude Opus 5.559.6%
- #14DeepSeek V4.1 Flash26.8%
- #15Grok 4.725.8%
- #16Qwen3.8 Flash-Next25.3%
artificialanalysis.ai · 2026-09-24
LiveCodeBenchEstimated
~#4 of 20 (estimated)~89.5%
- #1Qwen3.8-Omni-Flash92.6%
- #2Qwen3.8 Flash-Next91.9%
- #3Qwen3.8 27B90.3%
- ~#4Grok 4.7Estimated89.5%
- #4Nemotron 3 Ultra89.0%
No published score — estimated from related benchmarks
SciCodeVerified
#11 of 9557.4%
- #1Claude Opus 5.566.9%
- #2Claude Fable 5.163.1%
- #10GPT-6 Sol57.6%
- #11Grok 4.757.4%
- #12Gemini 3.8 Flash56.6%
artificialanalysis.ai · 2026-09-24
Writing & chat
IFBenchEstimated
~#21 of 126 (estimated)~73.8%
- #1MiniMax-M382.9%
- #2Qwen3.8 2.4T A95B82.8%
- #20Command A+73.9%
- ~#21Grok 4.7Estimated73.8%
- #21GPT-5.4 mini73.3%
No published score — estimated from related benchmarks
LMArena Text, Non-English (Elo)Verified
#26 of 531435 Elo
- #1Claude Opus 5.51501
- #2Claude Fable 5.1Few votes1496
- #25GPT-6 Luna1436
- #26Grok 4.71435
- #27Qwen3.5 397B A17B1434
lmarena.ai · 2026-09-25
#8 of 252007 Elo
eqbench.com · 2026-09-24
Vision
MMMU-ProEstimated
~#11 of 77 (estimated)~80.5%
- #1Claude Opus 5.587.7%
- #2GPT-6 Astra86.9%
- #10GPT-5.6 Terra80.7%
- ~#11Grok 4.7Estimated80.5%
- #11Kimi K380.5%
No published score — estimated from related benchmarks
CharXiv ReasoningEstimated
~#5 of 9 (estimated)~81.4%
- #1Gemini 3.8 Flash86.2%
- #2Kimi K384.8%
- #4Qwen3.8-Omni-Flash83.5%
- ~#5Grok 4.7Estimated81.4%
- #5Muse Glimmer78.8%
No published score — estimated from related benchmarks
OmniDocBench v1.5Estimated
~#4 of 8 (estimated)~87.1%
No published score — estimated from related benchmarks
ChartographyVerified
#18 of 2214.7%
- #1GPT-6 Astra71.0%
- #2Claude Opus 5.566.3%
- #17GLM-5.3 Flash16.3%
- #18Grok 4.714.7%
- #19DeepSeek V4.1 Flash11.6%
surgehq.ai · 2026-09-24
Hard reasoning
GPQA DiamondEstimated
~#16 of 173 (estimated)~92.3%
- #1GPT-6 Astra96.1%
- #2Gemini 3.8 Flash95.3%
- #15GPT-5.6 Terra92.5%
- ~#16Grok 4.7Estimated92.3%
- #16Hy4 preview92.3%
No published score — estimated from related benchmarks
Humanity's Last ExamVerified
#16 of 17843.1%
- #1Claude Opus 5.561.4%
- #2Claude Fable 5.159.1%
- #15Hy4 preview43.4%
- #16Grok 4.743.1%
- #16Qwen3.8 Max (0902)43.1%
artificialanalysis.ai · 2026-09-24
AIME (latest)Estimated
~#2 of 25 (estimated)~96.5%
No published score — estimated from related benchmarks
Agents
τ²-bench (Telecom)Estimated
~#2 of 116 (estimated)~96.4%
No published score — estimated from related benchmarks
OSWorld-VerifiedEstimated
~#3 of 12 (estimated)~82.3%
No published score — estimated from related benchmarks
OSWorld 2.0Estimated
~#3 of 5 (estimated)~58.3%
No published score — estimated from related benchmarks
BrowseCompEstimated
~#7 of 18 (estimated)~88.6%
- #1Atria Dawn Preview92.5%
- #2GPT-6 Astra91.5%
- #6Step 5 Preview88.7%
- ~#7Grok 4.7Estimated88.6%
- #7GPT-5.6 Terra87.5%
No published score — estimated from related benchmarks
τ-Bench Banking (AA)Estimated
~#11 of 89 (estimated)~43.4%
- #1Muse Spark 1.350.5%
- #2GLM-5.350.3%
- #10Gemini 3.8 Flash44.9%
- ~#11Grok 4.7Estimated43.4%
- #11Grok 4.643.3%
No published score — estimated from related benchmarks