Model Library / Benchmarks / LMArena (Text Arena)

Knowledge & Reasoning

LMArena (Text Arena)

The crowd-sourced overall leaderboard: blind, side-by-side human votes on model responses, aggregated into an Arena (Elo) score across 6.8M+ votes. Source: LMArena / Chatbot Arena official leaderboard (arena.ai, retrieved Aug 2026).

Features measured

Reasoning Multilingual

Use cases unlocked

general-chat Knowledge Q&A Drafting & summarising

Leaderboard

Higher is better · Arena Elo
1 Claude Mythos 5 1,531
2 Claude Fable 5 1,525
3 Claude Opus 5 1,522
4 GPT-5.6 1,514
5 Claude Opus 4.8 1,512
6 GPT-5.5 Pro 1,510
7 GPT-5.5 1,506
8 Gemini 3.1 Pro Preview 1,505
9 Claude Opus 4.7 1,505
10 Qwen3.8 Max Preview 1,496
11 Grok 4.3 1,496

Bars scaled from 1475 so close scores stay legible — the figure at right is the actual Arena Elo.

Benchmark source ↗

← All benchmarks