LMArena Code Elo leaderboard: which LLMs score highest

LMArena Code Elo (crowd-voted coding answers) currently has scores for 29 models. The top score is 1,815 by Claude Opus 5.5 (max); the median model scores 1,620 and the lowest scores 1,512.

LMArena Code Elo leaderboard: top 29 of 29 models (highest-effort row per model)
#ModelOrganisationLMArena Code Elo scoreCheapest input $/MSourceCaptured
1Claude Opus 5.5 (max)Anthropic1,815$4.00LMArena2026-10-05
2GPT-6 Astra (max)OpenAI1,788$10.00LMArena2026-10-05
3GPT-6.1 Sol (max)OpenAI1,758$2.00LMArena2026-10-05
4Claude Fable 5.1 (max)Anthropic1,749$10.00LMArena2026-10-05
5Claude Opus 5 (max)Anthropic1,695$5.00LMArena2026-10-05
6GPT-6 Sol (max)OpenAI1,689$2.00LMArena2026-10-05
7gemini-4-argon-highGoogle1,680—LMArena2026-10-05
8Kimi K3 (max)Moonshot1,658$1.29LMArena2026-10-05
9Muse Spark 1.3 (max)Meta AI1,657$1.25LMArena2026-10-05
10Grok 4.7 (xhigh)xAI1,638$2.00LMArena2026-10-05
11Qwen3.8-Flash-NextQwen1,637—LMArena2026-10-05
12Hy4-previewtencent1,633$0.751LMArena2026-10-05
13GLM-5.3 (max)Z.ai (Zhipu AI)1,623$0.070LMArena2026-10-05
14deepseek-v4.1-flash-maxDeepSeek1,620—LMArena2026-10-05
15gpt-5.6-sol-xhigh (codex-harness)OpenAI1,620$2.00LMArena2026-10-05
16mimo-v2.6-pro—1,618$0.435LMArena2026-10-05
17glm-5.2-maxZ.ai1,605—LMArena2026-10-05
18Gemini 3.7 Flash (high)Google DeepMind1,592$0.750LMArena2026-10-05
19deepseek-v4-pro-high-20260813DeepSeek1,583—LMArena2026-10-05
20Gemini 3.8 Flash (high)Google DeepMind1,583$0.750LMArena2026-10-05
21GPT-6 Luna (max)OpenAI1,579$0.100LMArena2026-10-05
22step-5-preview-high—1,570—LMArena2026-10-05
23Gemini 3.6 Flash (High)Google1,536$0.750LMArena2026-10-05
24Muse Spark 1.2 (xhigh)Meta AI1,532$1.25LMArena2026-10-05
25gpt-5.6-terra-xhigh (codex-harness)OpenAI1,519$2.00LMArena2026-10-05
26gpt-5.6-luna-xhigh (codex-harness)OpenAI1,518$0.200LMArena2026-10-05
27seed-2.1-pro-previewByteDance1,517—LMArena2026-10-05
28qwen3.7-max-20260517Alibaba1,515—LMArena2026-10-05
29gpt-5.5-xhigh (codex-harness)OpenAI1,512$5.00LMArena2026-10-05

Compare all benchmarks side by side on the live leaderboard →

Among the ten highest scorers, the cheapest listed API price belongs to Muse Spark 1.3 (max) at $1.25 per million input tokens (score 1,657).

What LMArena Code Elo measures

This is LMArena's coding category: Elo-style ratings from anonymous head-to-head votes on coding answers.

How to read these scores

Scores are Elo-style ratings (higher is better). Only differences between models mean anything, and small gaps may sit inside the source's own margin of error. Where a model is listed at several reasoning efforts, only its highest-effort row is shown. See the methodology for how rows are chosen.

Where the data comes from

LMArena (third-party snapshot 2026-10-05). Scores were last captured 2026-10-05; the Captured column gives each row's date.

Related benchmarks