Kimi K3 (max): benchmark scores, price and ranking

Kimi K3 (max) from Moonshot ranks #17 of 741 LLMs on the LLMs Tiger aggregate benchmark score (55.7), based on 23 benchmark results captured up to 2026-10-05.

Organisation
Moonshot
Type
Open-weight / local model
Open weights
Yes
Released
2026-07-16 (source: epoch)
Cheapest API price
$1.29 in / $14.00 out per million tokens (via openrouter, observed 2026-10-05T21:47Z)
Aggregate rank
#17 of 741 (score 55.7)

Relative to other models, Kimi K3 (max) is strongest on SciCode (#5 of 100) and weakest on Furniture Assembly (#17 of 25).

Kimi K3 (max) benchmark results

Kimi K3 (max) benchmark scores, ranked against all models measured on each benchmark
BenchmarkScoreRankSourceCaptured
SciCode59.5%#5 of 100Epoch AI Benchmarking Hub2026-10-05
GPQA Diamond93.1%#15 of 274Epoch AI Benchmarking Hub2026-10-05
Surface Evolver93.0%#1 of 14Epoch AI Benchmarking Hub2026-10-05
WeirdML82.6%#9 of 112Epoch AI Benchmarking Hub2026-10-05
LMCA52.7%#11 of 111Epoch AI Benchmarking Hub2026-10-05
ALE-Bench1,524#8 of 78Epoch AI Benchmarking Hub2026-10-05
WebDev Arena1,674#7 of 63Epoch AI Benchmarking Hub2026-10-05
OTIS Mock AIME97.2%#21 of 188Epoch AI Benchmarking Hub2026-10-05
Chess Puzzles39.0%#17 of 124Epoch AI Benchmarking Hub2026-10-05
SimpleBench60.7%#9 of 64Epoch AI Benchmarking Hub2026-10-05
DTBench91.2%#21 of 149Epoch AI Benchmarking Hub2026-10-05
CritPt23.4%#17 of 111Epoch AI Benchmarking Hub2026-10-05
SimpleQA Verified50.6%#16 of 66Epoch AI Benchmarking Hub2026-10-05
FrontierMath T1-372.2%#19 of 69Epoch AI Benchmarking Hub2026-10-05
LMArena Code Elo1,658#8 of 29LMArena2026-10-05
LMArena Text Elo1,488#9 of 28LMArena2026-10-05
ARC-AGI-260.4%#30 of 89ARC Prize official leaderboard2026-10-05
FrontierMath T439.0%#21 of 52Epoch AI Benchmarking Hub2026-10-05
DeepSWE68.5%#8 of 19Epoch AI Benchmarking Hub2026-10-05
Mystery Games26.0%#25 of 57Epoch AI Benchmarking Hub2026-10-05
GDP.pdf19.0%#13 of 22Epoch AI Benchmarking Hub2026-10-05
FrontierSWE25.9%#10 of 16Epoch AI Benchmarking Hub2026-10-05
Furniture Assembly34.2%#17 of 25Epoch AI Benchmarking Hub2026-10-05

Kimi K3 (max) is also listed at 4 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.

Models ranked nearby

Compare it on the live leaderboard → · How the aggregate score works