Claude Sonnet 5.5 vs Qwen3.8 Max (0902) vs Kimi K3
One row per board where at least one is ranked, from each board's current capture.
Head to head on 9 boards: Claude Sonnet 5.5 places best on 5, Qwen3.8 Max (0902) places best on 2 and Kimi K3 places best on 2. All 3 ranked on 8 of 13 boards.
Two at a time on this screen. Tap one to swap it in.
| Board | ||||
|---|---|---|---|---|
| Chat1 | ||||
| Oct 5today | #391,471 | #201,482 | #131,488best of the picked | |
| Reasoning5 | ||||
| Oct 5today | #256.0best of the picked | #1045.4 | #1343.6 | |
| Oct 5today | #295.6%best of the picked | #2192.7% | #1893.1% | |
| Oct 5today | #3165.2best of the picked | #20156.6 | #13157.6 | |
| Oct 5today | #2077.8% | #1578.5% | #1379.2%best of the picked | |
| Oct 5today | #2960.4% | |||
| Coding2 | ||||
| Oct 5today | #31,786best of the picked | #91,671 | #101,658 | |
| Oct 5today | #488.2% | |||
| Agents and tools1 | ||||
| Oct 5today | #882.3% | |||
| Vision1 | ||||
| Oct 5today | #291,268 | #21,301best of the picked | ||
| Speed1 | ||||
| Oct 5today | #5139 tok/sbest of the picked | #2439 tok/s | #2536 tok/s | |
| Price1 | ||||
| Oct 5today | #17$4.00 | #14$3.00best of the picked | #22$6.00 | |
| Usage1 | ||||
| Oct 5today | #151.6T tok |
Green marks the best placing among the picked where two or more are ranked; ties are marked for each. Ranks are each board's own order (lower is better), scores as the board published them. Dash: not on that board.
More comparisons
- GPT vs ClaudeGPT-6 Astra and Claude Fable 5.1, the top-ranked of each
- Claude vs GeminiClaude Fable 5.1 and Gemini 3.8 Flash, the top-ranked of each
- Anthropic vs OpenAI vs GoogleEach lab's best-ranked model on every board
- Open-weight vs closedThe best open-weight model on every board against the best closed one