GPT-5.6 Sol (max): benchmark scores, price and ranking

GPT-5.6 Sol (max) from OpenAI ranks #7 of 741 LLMs on the LLMs Tiger aggregate benchmark score (60.1), based on 24 benchmark results captured up to 2026-10-05.

Organisation
OpenAI
Type
Frontier model
Open weights
No
Released
2026-07-09 (source: epoch)
Cheapest API price
$2.00 in / $10.00 out per million tokens (via openrouter, observed 2026-10-05T21:47Z)
Aggregate rank
#7 of 741 (score 60.1)

Relative to other models, GPT-5.6 Sol (max) is strongest on CritPt (#1 of 111) and weakest on CursorBench (#6 of 13).

GPT-5.6 Sol (max) benchmark results

GPT-5.6 Sol (max) benchmark scores, ranked against all models measured on each benchmark
BenchmarkScoreRankSourceCaptured
CritPt32.3%#1 of 111Epoch AI Benchmarking Hub2026-10-05
OTIS Mock AIME100.0%#4 of 188Epoch AI Benchmarking Hub2026-10-05
GPQA Diamond93.5%#9 of 274Epoch AI Benchmarking Hub2026-10-05
ARC-AGI-292.5%#3 of 89ARC Prize official leaderboard2026-10-05
ALE-Bench2,177#3 of 78Epoch AI Benchmarking Hub2026-10-05
GDP.pdf30.7%#1 of 22Epoch AI Benchmarking Hub2026-10-05
Chess Puzzles55.0%#6 of 124Epoch AI Benchmarking Hub2026-10-05
LMCA58.4%#6 of 111Epoch AI Benchmarking Hub2026-10-05
WeirdML87.0%#7 of 112Epoch AI Benchmarking Hub2026-10-05
DTBench95.5%#11 of 149Epoch AI Benchmarking Hub2026-10-05
FrontierMath T1-389.1%#6 of 69Epoch AI Benchmarking Hub2026-10-05
BALROG60.0%#3 of 34Epoch AI Benchmarking Hub2026-10-05
SimpleQA Verified69.7%#6 of 66Epoch AI Benchmarking Hub2026-10-05
SciCode57.1%#10 of 100Epoch AI Benchmarking Hub2026-10-05
Mystery Games58.0%#7 of 57Epoch AI Benchmarking Hub2026-10-05
FrontierMath T482.9%#7 of 52Epoch AI Benchmarking Hub2026-10-05
ProofBench83.0%#6 of 43Epoch AI Benchmarking Hub2026-10-05
DeepSWE72.7%#4 of 19Epoch AI Benchmarking Hub2026-10-05
OSWorld 227.3%#2 of 8Epoch AI Benchmarking Hub2026-10-05
Furniture Assembly56.7%#8 of 25Epoch AI Benchmarking Hub2026-10-05
EBR-Bench44.8%#7 of 19Epoch AI Benchmarking Hub2026-10-05
FrontierSWE32.2%#7 of 16Epoch AI Benchmarking Hub2026-10-05
CursorBench41.7%#6 of 13Epoch AI Benchmarking Hub2026-10-05
PostTrainBench36.2%#2 of 4Epoch AI Benchmarking Hub2026-10-05

GPT-5.6 Sol (max) is also listed at 7 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.

Models ranked nearby

Compare it on the live leaderboard → · How the aggregate score works