Codebase understanding

SWEAtlas CodeBase QnA

Questions about how an unfamiliar codebase actually works — where something is handled, what a change would touch — answered by reading the repository rather than editing it. Tests understanding rather than patch-writing. Higher is better.

Rankings

Higher is better

SWEAtlas CodeBase QnA — frequently asked questions

What is SWEAtlas CodeBase QnA?
Questions about how an unfamiliar codebase actually works — where something is handled, what a change would touch — answered by reading the repository rather than editing it. Tests understanding rather than patch-writing. Higher is better.
Which AI model scores highest on SWEAtlas CodeBase QnA?
Muse Spark 1.3 by Meta holds the best SWEAtlas CodeBase QnA result among tracked models, at 59.4% (released Sep 2 2026). Higher scores are better on this benchmark.
What are the top 1 models on SWEAtlas CodeBase QnA?
1. Muse Spark 1.3 (Meta) — 59.4%.
How many models have a published SWEAtlas CodeBase QnA score?
1 tracked model has a published SWEAtlas CodeBase QnA score. Scores are the figures reported by each lab at that model's release, so this page is a record of results over time rather than a re-run leaderboard.