Compare models
Each lab shows its best-ranked model on every board. One row per board where at least one is ranked, from each board's current capture.
Google is ranked on 28 boards. Add another to compare.
Green marks the best placing among the picked where two or more are ranked; ties are marked for each. Ranks are each board's own order (lower is better), scores as the board published them. Dash: not on that board.
Ready-made comparisons
- GPT vs ClaudeGPT-6 Astra and Claude Fable 5.1, the top-ranked of each
- Claude vs GeminiClaude Fable 5.1 and Gemini 3.8 Flash, the top-ranked of each
- Anthropic vs OpenAI vs MetaEach lab's best-ranked model on every board
- Open-weight vs closedThe best open-weight model on every board against the best closed one