Claude Fable 5.1 leads the AI leaderboards
28 leaderboards and 835 models so far (+27 in 30 days), captured daily · 15 language-model boards count, #1 worth 10 points · How it is scored
Latest moves
| Model | |||
|---|---|---|---|
| Oct 6, 2026today | OpenRouter weekly usagenew version published (2026-10-04 to 2026-10-05) · Oct 6, 2026 today | OpenRouter weekly usage | new version published (2026-10-04 to 2026-10-05) |
| Oct 6, 2026today | GPT-5.6 SolOpenRouter weekly usage · score 1324.3 to 1297.3 (still #19) · Oct 6, 2026 today | OpenRouter weekly usage | score 1324.3 to 1297.3 (still #19) |
| Oct 6, 2026today | GLM-5.2OpenRouter weekly usage · score 1331.1 to 1329.2 (still #18) · Oct 6, 2026 today | OpenRouter weekly usage | score 1331.1 to 1329.2 (still #18) |
| Oct 6, 2026today | GPT-6 AstraOpenRouter weekly usage · score 1408.8 to 1388.2 (still #17) · Oct 6, 2026 today | OpenRouter weekly usage | score 1408.8 to 1388.2 (still #17) |
| Oct 6, 2026today | Muse Spark 1.3 ContributorOpenRouter weekly usage · score 1433.4 to 1446.7 (still #16) · Oct 6, 2026 today | OpenRouter weekly usage | score 1433.4 to 1446.7 (still #16) |
| Oct 6, 2026today | Kimi K3OpenRouter weekly usage · score 1619.3 to 1690.2 (still #15) · Oct 6, 2026 today | OpenRouter weekly usage | score 1619.3 to 1690.2 (still #15) |
| Oct 6, 2026today | Hy3OpenRouter weekly usage · score 2071.6 to 1861.3 (still #14) · Oct 6, 2026 today | OpenRouter weekly usage | score 2071.6 to 1861.3 (still #14) |
| Oct 6, 2026today | Gemini 3.8 FlashOpenRouter weekly usage · score 2182.2 to 2191.0 (still #13) · Oct 6, 2026 today | OpenRouter weekly usage | score 2182.2 to 2191.0 (still #13) |
| Oct 6, 2026today | Claude Opus 5.5OpenRouter weekly usage · score 2463 to 2598.9 (still #12) · Oct 6, 2026 today | OpenRouter weekly usage | score 2463 to 2598.9 (still #12) |
| Oct 6, 2026today | GLM-5.3OpenRouter weekly usage · score 2943.8 to 2970.3 (still #11) · Oct 6, 2026 today | OpenRouter weekly usage | score 2943.8 to 2970.3 (still #11) |
| Oct 6, 2026today | Jev 1.13OpenRouter weekly usage · score 3100.1 to 3224.7 (still #10) · Oct 6, 2026 today | OpenRouter weekly usage | score 3100.1 to 3224.7 (still #10) |
| Oct 6, 2026today | GPT-5.6 LunaOpenRouter weekly usage · score 4374.5 to 3366.0 (still #9) · Oct 6, 2026 today | OpenRouter weekly usage | score 4374.5 to 3366.0 (still #9) |
Top 10 of 58 ranked models
| # | Model | Score#1s | ||
|---|---|---|---|---|
| 1 | 50.7 | 50.71 | 10 | |
| 2 | 46.7 | 46.73 | 8 | |
| 3 | 36.0 | 36.03 | 6 | |
| 4 | 32.0 | 32.01 | 8 | |
| 5 | 30.7 | 30.70 | 7 | |
| 6 | 30.0 | 30.02 | 7 | |
| 7 | 23.3 | 23.30 | 5 | |
| 8 | 22.0 | 22.00 | 4 | |
| 9 | 18.7 | 18.71 | 6 | |
| 10 | 18.7 | 18.70 | 5 |
Compare:GPT vs ClaudeClaude vs GeminiAnthropic vs OpenAI vs GoogleOpen-weight vs closed
The podium wall
More views
Questions
Which AI model is best right now?
It depends on the job, which is why Models shows every major leaderboard side by side. The leader of leaders at the top of the home page is the model that finishes highest across all fresh language-model boards (chat, reasoning, coding, agents and vision), scored 10 points for #1 down to 1 point for #10 on each board.
What is the best LLM leaderboard?
There is no single one. LMArena measures what people prefer in blind votes, Artificial Analysis and Epoch AI run independent evaluation suites, Scale SEAL runs private expert benchmarks such as Humanity's Last Exam, SWE-bench and SWE-Bench Pro test real coding, BFCL and MCP Atlas test tool use. Models tracks them all and shows where they agree.
How is the leader of leaders calculated?
Each board is reduced to one entry per model (the best variant), ranked, and the top 10 earn 10, 9, 8 ... 1 points. Points are added across the boards counted, and the podium score is points earned over points available, from 0 to 100. Boards not updated in 180 days are shown but not counted. The full method is on the Method page.
How often is it updated?
Every day at 04:30 UTC, straight from each board's own page, data file or API. Each entry carries the date it was captured, and each board shows the date it last published.
Why do the same model names look different on each board?
Boards list variants: effort levels such as high or max, thinking modes, dated snapshots, and agent harnesses. Models maps them to one model and keeps the exact name each board published next to every entry.
Can I use the data?
Yes. /api/summary returns the computed standings as JSON, /llms-full.txt has every board's top 25 in plain text, and every number links back to the board that published it.