Claude Opus 5 (max): benchmark scores, price and ranking

Claude Opus 5 (max) from Anthropic ranks #6 of 741 LLMs on the LLMs Tiger aggregate benchmark score (60.1), based on 27 benchmark results captured up to 2026-10-05.

Organisation
Anthropic
Type
Frontier model
Open weights
No
Released
2026-07-24 (source: curated)
Cheapest API price
$5.00 in / $25.00 out per million tokens (via aihubmix, observed 2026-10-05T21:47Z)
Aggregate rank
#6 of 741 (score 60.1)

Relative to other models, Claude Opus 5 (max) is strongest on LMCA (#2 of 111) and weakest on GDP.pdf (#7 of 22).

Claude Opus 5 (max) benchmark results

Claude Opus 5 (max) benchmark scores, ranked against all models measured on each benchmark
BenchmarkScoreRankSourceCaptured
LMCA63.3%#2 of 111Epoch AI Benchmarking Hub2026-10-05
DTBench97.6%#3 of 149Epoch AI Benchmarking Hub2026-10-05
GPQA Diamond93.9%#8 of 274Epoch AI Benchmarking Hub2026-10-05
WeirdML91.8%#5 of 112Epoch AI Benchmarking Hub2026-10-05
ARC-AGI-290.4%#5 of 89ARC Prize official leaderboard2026-10-05
BALROG63.4%#2 of 34Epoch AI Benchmarking Hub2026-10-05
FrontierCode53.4%#1 of 15Epoch AI Benchmarking Hub2026-10-05
OTIS Mock AIME98.9%#13 of 188Epoch AI Benchmarking Hub2026-10-05
Mystery Games59.0%#5 of 57Epoch AI Benchmarking Hub2026-10-05
ProofBench99.0%#4 of 43Epoch AI Benchmarking Hub2026-10-05
WebDev Arena1,687#6 of 63Epoch AI Benchmarking Hub2026-10-05
CritPt29.1%#11 of 111Epoch AI Benchmarking Hub2026-10-05
Chess Puzzles42.0%#13 of 124Epoch AI Benchmarking Hub2026-10-05
DeepSWE73.6%#2 of 19Epoch AI Benchmarking Hub2026-10-05
OSWorld 231.4%#1 of 8Epoch AI Benchmarking Hub2026-10-05
SciCode56.4%#15 of 100Epoch AI Benchmarking Hub2026-10-05
FrontierMath T1-385.6%#11 of 69Epoch AI Benchmarking Hub2026-10-05
SimpleQA Verified59.9%#11 of 66Epoch AI Benchmarking Hub2026-10-05
LMArena Code Elo1,695#5 of 29LMArena2026-10-05
LMArena Text Elo1,489#6 of 28LMArena2026-10-05
APEX-Agents65.8%#3 of 14Epoch AI Benchmarking Hub2026-10-05
FrontierMath T473.2%#12 of 52Epoch AI Benchmarking Hub2026-10-05
Furniture Assembly60.8%#6 of 25Epoch AI Benchmarking Hub2026-10-05
CursorBench46.6%#4 of 13Epoch AI Benchmarking Hub2026-10-05
FrontierSWE52.0%#5 of 16Epoch AI Benchmarking Hub2026-10-05
EBR-Bench45.7%#6 of 19Epoch AI Benchmarking Hub2026-10-05
GDP.pdf24.0%#7 of 22Epoch AI Benchmarking Hub2026-10-05

Claude Opus 5 (max) is also listed at 6 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.

Models ranked nearby

Compare it on the live leaderboard → · How the aggregate score works