Claude Sonnet 5.5 vs Muse Spark vs Kimi K3 vs Qwen3.8 Max (0902)
One row per board where at least one is ranked, from each board's current capture.
Claude Sonnet 5.5
Muse Spark
Kimi K3
Qwen3.8 Max (0902)
Head to head on 10 boards: Claude Sonnet 5.5 places best on 5, Kimi K3 places best on 2, Qwen3.8 Max (0902) places best on 2 and Muse Spark places best on 1. All 4 ranked on 3 of 16 boards.
Two at a time on this screen. Tap one to swap it in.
Green marks the best placing among the picked where two or more are ranked; ties are marked for each. Ranks are each board's own order (lower is better), scores as the board published them. Dash: not on that board.
More comparisons
- GPT vs ClaudeGPT-6 Astra and Claude Fable 5.1, the top-ranked of each
- Claude vs GeminiClaude Fable 5.1 and Gemini 3.8 Flash, the top-ranked of each
- Anthropic vs OpenAI vs GoogleEach lab's best-ranked model on every board
- Open-weight vs closedThe best open-weight model on every board against the best closed one