Compare AI models

Specifications
Context window
1M—
API pricing
Cheapest input
—$0.25Google
Cheapest output
—$1.50Google
Benchmarks
SWE-Bench Verified
95.5%78%
Terminal-Bench 2.1
88%58%
OSWorld-Verified
85%65.1%
Overview
CompanyAnthropicGoogle
Release dateJun 9 2026Dec 17 2025
AccessClosedClosed
Model detailsView modelView model

Frequently asked questions

Claude Mythos 5 leads Gemini 3.0 Flash on 3 of the 3 benchmarks they both report (SWE-Bench Verified, Terminal-Bench 2.1, OSWorld-Verified). Gemini 3.0 Flash shipped 174 days before Claude Mythos 5, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On SWE-Bench Verified, Claude Mythos 5 leads at 95.5% vs Gemini 3.0 Flash at 78%. On Terminal-Bench 2.1, Claude Mythos 5 leads at 88% vs Gemini 3.0 Flash at 58%. On OSWorld-Verified, Claude Mythos 5 leads at 85% vs Gemini 3.0 Flash at 65.1%.