Skip to content

e5-small-v2 (intfloat) vs Claude Fable 5.1 vs Qwen3.8 Max (0902)

One row per board where at least one is ranked, from each board's current capture.

  • e5-small-v2 (intfloat)
  • Claude Fable 5.1
  • Qwen3.8 Max (0902)

Head to head on 8 boards: Claude Fable 5.1 places best on 6, Qwen3.8 Max (0902) places best on 2 and e5-small-v2 (intfloat) places best on 0. All 3 ranked on 0 of 15 boards.

Two at a time on this screen. Tap one to swap it in.

Boarde5-small-v2 (intfloat)Claude Fable 5.1Qwen3.8 Max (0902)
Chat1
Text ArenaArena ratingtoday#61,501best of the picked#201,482
Reasoning6
LiveBenchGlobal averagetoday#183.4%best of the picked#1578.5%
AA IntelligenceIndextoday#353.4best of the picked#1045.4
Epoch ECIECItoday#4164.8best of the picked#20156.6
HLEAccuracytoday#246.5%
ARC-AGI-2Scoretoday#690.0%
GPQA DiamondAccuracytoday#2192.7%
Coding2
WebDev ArenaArena ratingtoday#51,749best of the picked#91,671
SWE-Bench ProResolvedtoday#292.2%
Agents and tools1
MCP AtlasPass ratetoday#287.2%
Vision2
Vision ArenaArena ratingtoday#131,288#21,301best of the picked
Document ArenaArena ratingtoday#21,513
Embeddings1
MTEBMean task scoretoday#6944.5%
Speed1
AA SpeedOutput speedtoday#1568 tok/sbest of the picked#2439 tok/s
Price1
AA PriceBlended pricetoday#24$20.0#14$3.00best of the picked

Green marks the best placing among the picked where two or more are ranked; ties are marked for each. Ranks are each board's own order (lower is better), scores as the board published them. Dash: not on that board.

More comparisons

New #1s by email

Saturdays, only in weeks when a leaderboard has a new #1.

Double opt-in. Unsubscribe any time.