AI Battle > Rankings > GDPval-AA v2 Leaderboard

GDPval-AA v2 Leaderboard Ranking — LLM Index (2026)

Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) by Anthropic leads the GDPval-AA v2 Leaderboard ranking with a score of 1764. The ranking covers 20 evaluated llm index, re-scored on every AI Battle telemetry run.

Last updated: . Source: AI Battle live evaluation telemetry, 20 models ranked.

#ModelCreatorScore
1Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)Anthropic1764
2Claude Opus 5 (Adaptive Reasoning, Max Effort)Anthropic1735
3Muse Spark 1.3 (max)Meta1703
4GLM-5.3 (max)Z AI1658
5GLM-5.3-FlashZ AI1655
6Grok 4.6 (high)SpaceXAI1643
7DeepSeek V4.1 Flash (Reasoning, Max Effort)DeepSeek1632
8Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)Anthropic1631
9Qwen3.8 2.4T A95BAlibaba1628
10GPT-5.6 Sol (max)OpenAI1624

FAQ — GDPval-AA v2 Leaderboard

What is the best model for GDPval-AA v2 Leaderboard in 2026?

Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) by Anthropic currently leads the GDPval-AA v2 Leaderboard ranking with a score of 1764, based on live AI Battle telemetry across LLM Index.

How often is the GDPval-AA v2 Leaderboard ranking updated?

The GDPval-AA v2 Leaderboard ranking is refreshed automatically from the AI Battle evaluation cluster. Scores are re-evaluated on every telemetry run, so positions reflect the latest published results.

Which models rank in the top 3 for GDPval-AA v2 Leaderboard?

The top 3 are: 1. Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) (1764), 2. Claude Opus 5 (Adaptive Reasoning, Max Effort) (1735), 3. Muse Spark 1.3 (max) (1703).

All AI Battle Rankings