Cheating rate

CheatBench

Measures how often AI agents try to cheat on difficult assignments, such as reading hidden answers, copying work or manipulating grading. The overall score gives equal weight to ten categories; the sycophancy category measures how far an agent shifts its beliefs toward a user's stated views. Scores describe each model in its tested agent setup, and task success is measured separately. Lower is better.

Scores from CheatBench (Center for AI Safety), which runs the benchmark and publishes the full field.

Rankings

Lower is better

Scores marked “via” above are quoted with attribution from CheatBench, retrieved 25 September 2026.

CheatBench — frequently asked questions

Measures how often AI agents try to cheat on difficult assignments, such as reading hidden answers, copying work or manipulating grading. The overall score gives equal weight to ten categories; the sycophancy category measures how far an agent shifts its beliefs toward a user's stated views. Scores describe each model in its tested agent setup, and task success is measured separately. Lower is better.