Every AI leaderboard, side by side

Leaderboards28Models ranked835

gpt-transcribe (openai) vs Qwen3.8 Max (0902) vs Claude Fable 5.1

One row per board where at least one is ranked, from each board's current capture.

  • gpt-transcribe (openai)
  • Qwen3.8 Max (0902)
  • Claude Fable 5.1

Head to head on 8 boards: Claude Fable 5.1 places best on 6, Qwen3.8 Max (0902) places best on 2 and gpt-transcribe (openai) places best on 0. All 3 ranked on 0 of 15 boards.

Two at a time on this screen. Tap one to swap it in.

Boardgpt-transcribe (openai)Qwen3.8 Max (0902)Claude Fable 5.1
Chat1
Text ArenaArena ratingyesterday#201,482#61,501best of the picked
Reasoning6
LiveBenchGlobal averageyesterday#1578.5%#183.4%best of the picked
AA IntelligenceIndexyesterday#1045.4#353.4best of the picked
Epoch ECIECIyesterday#22156.4#5164.7best of the picked
HLEAccuracyyesterday#246.5%
ARC-AGI-2Scoreyesterday#690.0%
GPQA DiamondAccuracyyesterday#2192.7%
Coding2
WebDev ArenaArena ratingyesterday#91,671#51,749best of the picked
SWE-Bench ProResolvedyesterday#292.2%
Agents and tools1
MCP AtlasPass rateyesterday#287.2%
Vision2
Vision ArenaArena ratingyesterday#21,301best of the picked#131,288
Document ArenaArena ratingyesterday#21,513
Speech1
Open ASRAverage WERyesterday#204.57 WER
Speed1
AA SpeedOutput speedyesterday#2439 tok/s#1568 tok/sbest of the picked
Price1
AA PriceBlended priceyesterday#14$3.00best of the picked#24$20.0

Green marks the best placing among the picked where two or more are ranked; ties are marked for each. Ranks are each board's own order (lower is better), scores as the board published them. Dash: not on that board.

More comparisons