Knowledge work

GDPval-AA v2

economically valuable knowledge work (v2, re-based Elo)

Rankings

Higher is better

GDPval-AA v2 — frequently asked questions

What is GDPval-AA v2?
economically valuable knowledge work (v2, re-based Elo)
Which AI model scores highest on GDPval-AA v2?
Claude Opus 5 by Anthropic holds the best GDPval-AA v2 result among tracked models, at 1861 (released Jul 24 2026). Higher scores are better on this benchmark.
What are the top 5 models on GDPval-AA v2?
1. Claude Opus 5 (Anthropic) — 1861; 2. Claude Fable 5 (Anthropic) — 1760; 3. GPT-5.6 Sol (OpenAI) — 1748; 4. Kimi K3 (Moonshot AI) — 1668; 5. Claude Opus 4.8 (Anthropic) — 1600.
How many models have a published GDPval-AA v2 score?
11 tracked models have a published GDPval-AA v2 score. Scores are the figures reported by each lab at that model's release, so this page is a record of results over time rather than a re-run leaderboard.
← All benchmarks