# Humanity's Last Exam (Verified) — AI model rankings

The re-checked edition of Humanity's Last Exam: the same extremely hard expert questions, minus the ones found to be flawed or wrongly answered. Scores on it run lower than on the original exam, so read the two as separate tests rather than a before-and-after. Higher is better.

2 tracked models have a published Humanity's Last Exam (Verified) score. Higher is better. Scores come from published lab reports and benchmark sources; recorded sources appear beside the scores.

## Ranking

| Rank | Model | Developer | Score | Source | Released |
| --- | --- | --- | --- | --- | --- |
| 1 | Gemini 3.8 Flash | Google | 54.9% | Lab | Sep 2 2026 |
| 2 | Gemini 3.7 Flash | Google | 53.6% | Lab | Aug 13 2026 |


---

Canonical page: https://aireleasetracker.com/benchmark/hle-verified
Site index: https://aireleasetracker.com/llms.txt
Source: AI Release Tracker (https://aireleasetracker.com). Most benchmark scores come from lab launch material; gathered results identify the leaderboard that published them.
