AI Battle > Rankings > GDPval-AA v2 Leaderboard
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) by Anthropic leads the GDPval-AA v2 Leaderboard ranking with a score of 1764. The ranking covers 20 evaluated llm index, re-scored on every AI Battle telemetry run.
Last updated: . Source: AI Battle live evaluation telemetry, 20 models ranked.
| # | Model | Creator | Score |
|---|---|---|---|
| 1 | Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) | Anthropic | 1764 |
| 2 | Claude Opus 5 (Adaptive Reasoning, Max Effort) | Anthropic | 1735 |
| 3 | Muse Spark 1.3 (max) | Meta | 1703 |
| 4 | GLM-5.3 (max) | Z AI | 1658 |
| 5 | GLM-5.3-Flash | Z AI | 1655 |
| 6 | Grok 4.6 (high) | SpaceXAI | 1643 |
| 7 | DeepSeek V4.1 Flash (Reasoning, Max Effort) | DeepSeek | 1632 |
| 8 | Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) | Anthropic | 1631 |
| 9 | Qwen3.8 2.4T A95B | Alibaba | 1628 |
| 10 | GPT-5.6 Sol (max) | OpenAI | 1624 |
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) by Anthropic currently leads the GDPval-AA v2 Leaderboard ranking with a score of 1764, based on live AI Battle telemetry across LLM Index.
The GDPval-AA v2 Leaderboard ranking is refreshed automatically from the AI Battle evaluation cluster. Scores are re-evaluated on every telemetry run, so positions reflect the latest published results.
The top 3 are: 1. Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) (1764), 2. Claude Opus 5 (Adaptive Reasoning, Max Effort) (1735), 3. Muse Spark 1.3 (max) (1703).