LMArena Text Elo leaderboard: which LLMs score highest

LMArena Text Elo (head-to-head chat preference (crowd votes)) currently has scores for 28 models. The top score is 1,525 by gemini-4-argon-high; the median model scores 1,478 and the lowest scores 1,468.

LMArena Text Elo leaderboard: top 28 of 28 models (highest-effort row per model)
#ModelOrganisationLMArena Text Elo scoreCheapest input $/MSourceCaptured
1gemini-4-argon-highGoogle1,525—LMArena2026-10-05
2Claude Fable 5.1 (max)Anthropic1,501$10.00LMArena2026-10-05
3Gemini 3.8 Flash (high)Google DeepMind1,495$0.750LMArena2026-10-05
4Muse Spark 1.2 (xhigh)Meta AI1,494$1.25LMArena2026-10-05
5Muse Spark 1.3 (max)Meta AI1,494$1.25LMArena2026-10-05
6Claude Opus 5 (max)Anthropic1,489$5.00LMArena2026-10-05
7Muse SparkMeta AI1,489—LMArena2026-10-05
8Gemini 3.7 Flash (high)Google DeepMind1,488$0.750LMArena2026-10-05
9Kimi K3 (max)Moonshot1,488$1.29LMArena2026-10-05
10Gemini 3 ProGoogle DeepMind1,485$2.00LMArena2026-10-05
11Gemini 3.6 Flash (High)Google1,483$0.750LMArena2026-10-05
12GPT-6.1 Sol (max)OpenAI1,483$2.00LMArena2026-10-05
13mimo-v2.6-pro—1,480$0.435LMArena2026-10-05
14GLM-5.3 (max)Z.ai (Zhipu AI)1,478$0.070LMArena2026-10-05
15gemini-3.5-flash-highGoogle1,477$1.50LMArena2026-10-05
16GPT-6 Astra (max)OpenAI1,477$10.00LMArena2026-10-05
17glm-5.2-maxZ.ai1,476—LMArena2026-10-05
18gpt-5.2-chat-latest-20260210OpenAI1,476—LMArena2026-10-05
19grok-4.20-beta1xAI1,475—LMArena2026-10-05
20qwen3.7-max-previewAlibaba1,475$1.25LMArena2026-10-05
21claude-opus-4-5-20251101-high-32kAnthropic1,474$5.00LMArena2026-10-05
22deepseek-v4.1-flash-maxDeepSeek1,474—LMArena2026-10-05
23Gemini 3 FlashGoogle DeepMind1,473$0.500LMArena2026-10-05
24GPT-5.5 InstantOpenAI1,473—LMArena2026-10-05
25grok-4.20-beta-0309-reasoningxAI1,472$1.25LMArena2026-10-05
26grok-4.20-multi-agent-beta-0309xAI1,471$1.25LMArena2026-10-05
27ernie-5.1—1,468$0.563LMArena2026-10-05
28mimo-v2.5-proXiaomi Corp1,468$0.435LMArena2026-10-05

Compare all benchmarks side by side on the live leaderboard →

Among the ten highest scorers, the cheapest listed API price belongs to Gemini 3.8 Flash (high) at $0.750 per million input tokens (score 1,495).

What LMArena Text Elo measures

LMArena (formerly Chatbot Arena) ranks models with Elo-style ratings computed from anonymous head-to-head votes by real users, so it measures human preference rather than accuracy on a fixed test.

How to read these scores

Scores are Elo-style ratings (higher is better). Only differences between models mean anything, and small gaps may sit inside the source's own margin of error. Where a model is listed at several reasoning efforts, only its highest-effort row is shown. See the methodology for how rows are chosen.

Where the data comes from

LMArena (third-party snapshot 2026-10-05) and LMArena (third-party snapshot 2026-10-04). Scores were last captured 2026-10-05; the Captured column gives each row's date.

Related benchmarks