DeepSeek V4 Pro 0813 (max): benchmark scores, price and ranking

DeepSeek V4 Pro 0813 (max) from DeepSeek ranks #25 of 741 LLMs on the LLMs Tiger aggregate benchmark score (54.3), based on 12 benchmark results captured up to 2026-10-05.

Organisation
DeepSeek
Type
Open-weight / local model
Open weights
Yes
Released
2026-08-13 (source: epoch)
Cheapest API price
$0.660 in / $1.98 out per million tokens (via openrouter, observed 2026-10-05T21:47Z)
Aggregate rank
#25 of 741 (score 54.3)

Relative to other models, DeepSeek V4 Pro 0813 (max) is strongest on Chess Puzzles (#10 of 124) and weakest on FrontierMath T4 (#31 of 52).

DeepSeek V4 Pro 0813 (max) benchmark results

DeepSeek V4 Pro 0813 (max) benchmark scores, ranked against all models measured on each benchmark
BenchmarkScoreRankSourceCaptured
Chess Puzzles47.0%#10 of 124Epoch AI Benchmarking Hub2026-10-05
OTIS Mock AIME98.6%#16 of 188Epoch AI Benchmarking Hub2026-10-05
GPQA Diamond91.7%#24 of 274Epoch AI Benchmarking Hub2026-10-05
ALE-Bench1,403#10 of 78Epoch AI Benchmarking Hub2026-10-05
WeirdML66.2%#16 of 112Epoch AI Benchmarking Hub2026-10-05
CritPt18.0%#23 of 111Epoch AI Benchmarking Hub2026-10-05
Mystery Games43.0%#12 of 57Epoch AI Benchmarking Hub2026-10-05
SimpleQA Verified52.9%#14 of 66Epoch AI Benchmarking Hub2026-10-05
SciCode51.0%#29 of 100Epoch AI Benchmarking Hub2026-10-05
ARC-AGI-261.2%#28 of 89ARC Prize official leaderboard2026-10-05
FrontierMath T1-364.6%#29 of 69Epoch AI Benchmarking Hub2026-10-05
FrontierMath T426.8%#31 of 52Epoch AI Benchmarking Hub2026-10-05

DeepSeek V4 Pro 0813 (max) is also listed at 5 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.

Models ranked nearby

  • #22 gemini-4-argon-high (54.5)
  • #23 Claude 4.5 Opus (54.5)
  • #24 Claude 3.5-Sonnet-20241022 (54.4)
  • #26 GPT-5.3 Codex (xhigh) (54.2)
  • #27 Qwen 3.6 Max (Preview) (54.1)
  • #28 GPT-6 Astra (pro, max) (54.0)

Compare it on the live leaderboard → · How the aggregate score works