Every AI leaderboard, side by side

Leaderboards28Models ranked835

gpt-transcribe (openai) vs Qwen3.8 Max (0902) vs Muse Spark

One row per board where at least one is ranked, from each board's current capture.

  • gpt-transcribe (openai)
  • Qwen3.8 Max (0902)
  • Muse Spark

Head to head on 4 boards: Qwen3.8 Max (0902) places best on 3, Muse Spark places best on 1 and gpt-transcribe (openai) places best on 0. All 3 ranked on 0 of 14 boards.

Two at a time on this screen. Tap one to swap it in.

Boardgpt-transcribe (openai)Qwen3.8 Max (0902)Muse Spark
Chat2
Text ArenaArena ratingyesterday#201,482#121,489best of the picked
MultiChallengeScoreyesterday#175.5%
Reasoning5
GPQA DiamondAccuracyyesterday#2192.7%best of the picked#4189.8%
Epoch ECIECIyesterday#22156.4best of the picked#44152.0
HLEAccuracyyesterday#640.6%
AA IntelligenceIndexyesterday#1045.4
LiveBenchGlobal averageyesterday#1578.5%
Coding1
WebDev ArenaArena ratingyesterday#91,671
Agents and tools1
MCP AtlasPass rateyesterday#982.2%
Vision2
Vision ArenaArena ratingyesterday#21,301best of the picked#61,294
Document ArenaArena ratingyesterday#241,444
Speech1
Open ASRAverage WERyesterday#204.57 WER
Speed1
AA SpeedOutput speedyesterday#2439 tok/s
Price1
AA PriceBlended priceyesterday#14$3.00

Green marks the best placing among the picked where two or more are ranked; ties are marked for each. Ranks are each board's own order (lower is better), scores as the board published them. Dash: not on that board.

More comparisons