LiveBench Data Analysis leaderboard: which LLMs score highest
LiveBench Data Analysis (LiveBench release 2026_06_25) currently has scores for 17 models. The top score is 81.6% by GPT-5.5 (xhigh); the median model scores 71.8% and the lowest scores 62.7%.
| # | Model | Organisation | LiveBench Data Analysis score | Cheapest input $/M | Source | Captured |
|---|---|---|---|---|---|---|
| 1 | GPT-5.5 (xhigh) | OpenAI | 81.6% | $5.00 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 2 | GPT-5.4 (xhigh) | OpenAI | 79.3% | $2.50 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 3 | gemini-3.1-pro-preview-high | 78.5% | $2.00 | LiveBench 2026_06_25 official CSV | 2026-10-05 | |
| 4 | claude-opus-4-8-xhigh-effort | Anthropic | 78.3% | $5.00 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 5 | GPT-5.2 Codex | OpenAI | 78.2% | $1.75 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 6 | gpt-5.2-2025-12-11-high | OpenAI | 78.2% | $1.75 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 7 | MiniMax-M3 | MiniMax | 76.2% | $0.230 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 8 | claude-opus-4-5-20251101-thinking-64k-high-effort | Anthropic | 74.4% | $5.00 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 9 | Qwen3.7 Max | Alibaba | 71.8% | $1.25 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 10 | grok-build-0.1 | xAI | 70.8% | $1.00 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 11 | GPT-5.4 mini (xhigh) | OpenAI | 70.8% | $0.750 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 12 | Qwen3.6 27B | Alibaba | 70.4% | $0.150 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 13 | qwen3.6-plus | Alibaba | 69.9% | $0.325 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 14 | GPT-5.4 nano (xhigh) | OpenAI | 67.6% | $0.200 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 15 | kimi-k2.6-thinking | Moonshot AI | 65.1% | $0.650 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
| 16 | gemini-3.5-flash-high | 64.9% | $1.50 | LiveBench 2026_06_25 official CSV | 2026-10-05 | |
| 17 | Kimi K2.7 Code | Moonshot | 62.7% | $0.671 | LiveBench 2026_06_25 official CSV | 2026-10-05 |
Compare all benchmarks side by side on the live leaderboard →
Among the ten highest scorers, the cheapest listed API price belongs to MiniMax-M3 at $0.230 per million input tokens (score 76.2%).
What LiveBench Data Analysis measures
LiveBench Data Analysis measures LiveBench release 2026_06_25.
How to read these scores
Scores are percentages (higher is better): the share of questions or tasks the model completed correctly. Where a model is listed at several reasoning efforts, only its highest-effort row is shown. See the methodology for how rows are chosen.
Where the data comes from
LiveBench 2026_06_25 official CSV. Scores were last captured 2026-10-05; the Captured column gives each row's date.