Knowledge work
GDPval-AA v2
economically valuable knowledge work (v2, re-based Elo)
Rankings
Higher is betterGDPval-AA v2 — frequently asked questions
- What is GDPval-AA v2?
- economically valuable knowledge work (v2, re-based Elo)
- Which AI model scores highest on GDPval-AA v2?
- Claude Opus 5 by Anthropic holds the best GDPval-AA v2 result among tracked models, at 1861 (released Jul 24 2026). Higher scores are better on this benchmark.
- What are the top 5 models on GDPval-AA v2?
- 1. Claude Opus 5 (Anthropic) — 1861; 2. Claude Fable 5 (Anthropic) — 1760; 3. GPT-5.6 Sol (OpenAI) — 1748; 4. Kimi K3 (Moonshot AI) — 1668; 5. Claude Opus 4.8 (Anthropic) — 1600.
- How many models have a published GDPval-AA v2 score?
- 11 tracked models have a published GDPval-AA v2 score. Scores are the figures reported by each lab at that model's release, so this page is a record of results over time rather than a re-run leaderboard.