Chart reasoning

CharXiv Reasoning

Can the AI read and reason about complex charts and figures, not just text? Higher is better.

Rankings

Higher is better

CharXiv Reasoning — frequently asked questions

What is CharXiv Reasoning?
Can the AI read and reason about complex charts and figures, not just text? Higher is better.
Which AI model scores highest on CharXiv Reasoning?
Muse Spark by Meta holds the best CharXiv Reasoning result among tracked models, at 88.9% (released Apr 8 2026). Higher scores are better on this benchmark.
What are the top 5 models on CharXiv Reasoning?
1. Muse Spark (Meta) — 88.9%; 2. Muse Spark 1.1 (Meta) — 88.4%; 2. Qwen3.8-Max (Qwen) — 88.4%; 4. Kimi K3 (Moonshot AI) — 84.8%; 5. Gemini 3.5 Flash (Google) — 84.2%.
How many models have a published CharXiv Reasoning score?
12 tracked models have a published CharXiv Reasoning score. Scores are the figures reported by each lab at that model's release, so this page is a record of results over time rather than a re-run leaderboard.
What is the best open model on CharXiv Reasoning?
Qwen3.5 by Qwen is the highest-ranked model with downloadable weights on CharXiv Reasoning, scoring 80.8% at rank 9 overall.
← All benchmarks