Claude Fable 5.1 (max): benchmark scores, price and ranking

Claude Fable 5.1 (max) from Anthropic ranks #3 of 741 LLMs on the LLMs Tiger aggregate benchmark score (61.9), based on 21 benchmark results captured up to 2026-10-05.

Organisation
Anthropic
Type
Frontier model
Open weights
No
Released
2026-09-01 (source: epoch)
Cheapest API price
$10.00 in / $50.00 out per million tokens (via anthropic, observed 2026-10-05T21:47Z)
Aggregate rank
#3 of 741 (score 61.9)

Relative to other models, Claude Fable 5.1 (max) is strongest on OTIS Mock AIME (#1 of 188) and weakest on LMArena Vision Elo (#8 of 22).

Claude Fable 5.1 (max) benchmark results

Claude Fable 5.1 (max) benchmark scores, ranked against all models measured on each benchmark
BenchmarkScoreRankSourceCaptured
OTIS Mock AIME100.0%#1 of 188Epoch AI Benchmarking Hub2026-10-05
SciCode63.1%#2 of 100Epoch AI Benchmarking Hub2026-10-05
ProofBench100.0%#1 of 43Epoch AI Benchmarking Hub2026-10-05
WeirdML92.9%#3 of 112Epoch AI Benchmarking Hub2026-10-05
LMArena Document Elo1,513#1 of 18LMArena2026-10-05
FrontierMath T1-390.2%#4 of 69Epoch AI Benchmarking Hub2026-10-05
SimpleQA Verified70.8%#4 of 66Epoch AI Benchmarking Hub2026-10-05
WebDev Arena1,758#4 of 63Epoch AI Benchmarking Hub2026-10-05
ARC-AGI-290.0%#6 of 89ARC Prize official leaderboard2026-10-05
LMArena Text Elo1,501#2 of 28LMArena2026-10-05
Chess Puzzles47.0%#9 of 124Epoch AI Benchmarking Hub2026-10-05
CritPt29.7%#10 of 111Epoch AI Benchmarking Hub2026-10-05
Mystery Games58.0%#6 of 57Epoch AI Benchmarking Hub2026-10-05
FrontierMath T487.8%#6 of 52Epoch AI Benchmarking Hub2026-10-05
LMArena Code Elo1,749#4 of 29LMArena2026-10-05
EBR-Bench57.1%#3 of 19Epoch AI Benchmarking Hub2026-10-05
GDP.pdf27.6%#4 of 22Epoch AI Benchmarking Hub2026-10-05
Furniture Assembly70.0%#5 of 25Epoch AI Benchmarking Hub2026-10-05
CursorBench51.8%#3 of 13Epoch AI Benchmarking Hub2026-10-05
FrontierSWE56.3%#4 of 16Epoch AI Benchmarking Hub2026-10-05
LMArena Vision Elo1,288#8 of 22LMArena2026-10-05

Claude Fable 5.1 (max) is also listed at 6 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.

Models ranked nearby

Compare it on the live leaderboard → · How the aggregate score works