Advanced math
FrontierMathTier 4 (v2)
The rebuilt edition of FrontierMath's hardest tier — research-level maths of the kind professional mathematicians work on. It is a different question set from the first Tier 4, and scores on it run far higher, so read the two as separate tests rather than progress. Higher is better.
Rankings
Higher is betterFrontierMath — frequently asked questions
- What is FrontierMath?
- The rebuilt edition of FrontierMath's hardest tier — research-level maths of the kind professional mathematicians work on. It is a different question set from the first Tier 4, and scores on it run far higher, so read the two as separate tests rather than progress. Higher is better.
- Which AI model scores highest on FrontierMath?
- GPT-6 Astra by OpenAI holds the best FrontierMath (Tier 4 (v2)) result among tracked models, at 97.6% (released Sep 3 2026). Higher scores are better on this benchmark.
- What are the top 1 models on FrontierMath?
- 1. GPT-6 Astra (OpenAI) — 97.6%.
- How many models have a published FrontierMath score?
- 1 tracked model has a published FrontierMath (Tier 4 (v2)) score. Scores are the figures reported by each lab at that model's release, so this page is a record of results over time rather than a re-run leaderboard.