GPT-5.1 (high): benchmark scores, price and ranking
GPT-5.1 (high) from OpenAI ranks #52 of 741 LLMs on the LLMs Tiger aggregate benchmark score (52.9), based on 15 benchmark results captured up to 2026-10-05.
- Organisation
- OpenAI
- Type
- Frontier model
- Open weights
- No
- Released
- 2025-11-12 (source: curated)
- Cheapest API price
- $1.25 in / $10.00 out per million tokens (via azure, observed 2026-10-05T21:47Z)
- Aggregate rank
- #52 of 741 (score 52.9)
Relative to other models, GPT-5.1 (high) is strongest on CL-bench (#2 of 16) and weakest on MCP Atlas (#14 of 14).
GPT-5.1 (high) benchmark results
GPT-5.1 (high) is also listed at 5 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.
Models ranked nearby
- #49 Claude Mythos Preview (Early) (53.0)
- #50 Solar-Open2-250B (52.9)
- #51 Gemini 2.5 Pro Preview 05-06 (52.9)
- #53 Kimi K2.5 (Fireworks) (52.8)
- #54 CodeAct v2.1 (claude-3-5-sonnet-20241022) (52.8)
- #55 claude-3.5-sonnet-20240620 (52.7)
Compare it on the live leaderboard → · How the aggregate score works