Grok 4.6 (xhigh): benchmark scores, price and ranking
Grok 4.6 (xhigh) from xAI ranks #36 of 741 LLMs on the LLMs Tiger aggregate benchmark score (53.5), based on 19 benchmark results captured up to 2026-10-05.
- Organisation
- xAI
- Type
- Frontier model
- Open weights
- No
- Released
- 2026-08-12 (source: curated)
- Cheapest API price
- $1.25 in / $6.00 out per million tokens (via azure_ai, observed 2026-10-05T21:47Z)
- Aggregate rank
- #36 of 741 (score 53.5)
Relative to other models, Grok 4.6 (xhigh) is strongest on DTBench (#5 of 149) and weakest on FrontierSWE (#11 of 16).
Grok 4.6 (xhigh) benchmark results
Grok 4.6 (xhigh) is also listed at 5 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.
Models ranked nearby
- #33 Claude Opus 4.7 (max) (53.7)
- #34 GPT-6 Luna (max) (53.7)
- #35 DeepSeek R1 (0528) (53.6)
- #37 Gemini 3.8 Flash (high) (53.5)
- #38 gemini-2.5-pro-preview-06-05 (53.5)
- #39 Qwen3.8 Max (xhigh) (53.2)
Compare it on the live leaderboard → · How the aggregate score works