Cogito v2.1
- Input
- $1.25/M tok
- Output
- $1.25/M tok
- Blended
- $1.25/M tok · 3:1
- Speed
- ~82tok/sest.
- Context
- 128Ktokens
Long-tail import from Artificial Analysis (2026-09-24); AA variant: Cogito v2.1 (Reasoning). Price = AA median host price.
Rankings
Self-hostingEstimated
How we estimate →What it would cost to run these open weights yourself on rented GPUs, compared with the API price above.
- Hardware
- 8× B200 180GB
- FP8 weights
- Throughput
- ~21,907 tok/s
- many requests batched
- Price
- $0.26–2.32 /M tok
- busy → light use
- On your own machine
- Mac Studio M5 Ultra 512GB (int4, ~45 tok/s single-stream)
- single consumer GPU or Mac
Assumptions (5)
- FP8 weights (671B params, 37B active per token) + 25% KV-cache headroom ≈ 839 GB
- 8× B200 180GB at $5.98–$14.24/GPU-hour on-demand (2026-09-24)
- ~21,907 output tok/s aggregate at batch 627 (bandwidth-bound); MoE compute scales with active params
- Blended 3:1 input:output; prefill ~229,865 tok/s
- Low = 75% utilization at the low GPU price; high = 20% at the high price
Benchmark scores
~ italic, dashed = no published score yet, estimated from related benchmarks.
Writing & chat
IFBenchVerified
#63 of 12546.3%
- #1MiniMax-M382.9%
- #2Qwen3.8 2.4T A95B82.8%
- #62Llama 3.3 Instruct 70B47.1%
- #63Cogito v2.146.3%
- #64LFM2 24B A2B45.9%
artificialanalysis.ai · 2026-09-24
MMMLUEstimated
~last of 5 (estimated)below measured range (<81.3%)
No published score — estimated from related benchmarks
LMArena Text, Non-English (Elo)Estimated
~#43 of 54 (estimated)~1348 Elo
- #1Claude Opus 5.51501
- #2Claude Fable 5.1Few votes1496
- #42Trinity Large Thinking1353
- ~#43Cogito v2.1Estimated1348
- #43gpt-oss-120b1338
No published score — estimated from related benchmarks
LMArena Text (Elo)Estimated
~#43 of 54 (estimated)~1369 Elo
- #1Claude Opus 5.51509
- #2Claude Fable 5.11501
- #42Trinity Large Thinking1369
- ~#43Cogito v2.1Estimated1369
- #43INTELLECT-31356
No published score — estimated from related benchmarks
EQ-Bench Creative Writing v3 (Elo)Estimated
~#21 of 26 (estimated)~1537 Elo
- #1GPT-6 Astra2173
- #2Claude Fable 5.12162
- #20DeepSeek V4.1 Flash1540
- ~#21Cogito v2.1Estimated1537
- #21Gemini 3.1 Pro1491
No published score — estimated from related benchmarks
Hard reasoning
GPQA DiamondVerified
#70 of 17276.8%
- #1GPT-6 Astra96.1%
- #2Gemini 3.8 Flash95.3%
- #69Mistral Small 476.9%
- #70Cogito v2.176.8%
- #71Command A+76.1%
artificialanalysis.ai · 2026-09-24
Humanity's Last ExamVerified
#80 of 17812.0%
- #1Claude Opus 5.561.4%
- #2Claude Fable 5.159.1%
- #79EXAONE 4.5 33B12.9%
- #80Cogito v2.112.0%
- #80Command A+12.0%
artificialanalysis.ai · 2026-09-24
AIME (latest)Estimated
~#14 of 25 (estimated)~73.7%
- #1Inkling97.1%
- #2Muse Glimmer94.7%
- #13Phi-4-reasoning-plus78.0%
- ~#14Cogito v2.1Estimated73.7%
- #14Mistral Large 338.0%
No published score — estimated from related benchmarks
AA Intelligence IndexEstimated
~#50 of 75 (estimated)~15.8
- #1Claude Opus 5.557.6
- #2Claude Sonnet 5.556.0
- #49Ring-2.6-1T16.6
- ~#50Cogito v2.1Estimated15.8
- #50K2 Horizon 3.7B15.6
No published score — estimated from related benchmarks