Grok 4.7 (xhigh): benchmark scores, price and ranking
Grok 4.7 (xhigh) from xAI ranks #45 of 751 LLMs on the LLMs Tiger aggregate benchmark score (53.1), based on 19 benchmark results captured up to 2026-10-07.
- Organisation
- xAI
- Type
- Frontier model
- Open weights
- No
- Released
- 2026-09-21 (source: curated)
- Cheapest API price
- $2.00 in / $6.00 out per million tokens (via azure_ai, observed 2026-10-08T00:16Z)
- Aggregate rank
- #45 of 751 (score 53.1)
Relative to other models, Grok 4.7 (xhigh) is strongest on GPQA Diamond (#17 of 274) and weakest on Furniture Assembly (#24 of 25).
Grok 4.7 (xhigh) benchmark results
Grok 4.7 (xhigh) is also listed at 5 other reasoning-effort settings; this page shows the highest-effort row, which is the one the leaderboard keeps by default.
Models ranked nearby
- #42 grok-4-0709 (53.2)
- #43 Grok 4 (53.2)
- #44 NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4 (53.1)
- #46 NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 (53.1)
- #47 o3 + gpt-4.1 (53.0)
- #48 Claude Code + Sonnet 5.5 (53.0)
Compare it on the live leaderboard → · How the aggregate score works