Compare AI models

Specifications
Parameters
—125B
Context window
1M262k
API pricing
Input price
$0.75—
Output price
$3.75—
Cached input price
$0.075—
Cheapest input
$0.375Google—
Cheapest output
$1.875Google—
Benchmarks
DeepSWE 1.1
65.3%58.7%
OSWorld 2.0
38.1%19.4%
Agent's Last Exam · pass@1
26.3%24.3%
CharXiv Reasoning
84.5%84.6%
LVBench
85.4%76.6%
Overview
CompanyGoogleQwen
Release dateAug 13 2026Aug 26 2026
AccessClosedOpen Weight
Model detailsView modelView model

Frequently asked questions

Gemini 3.7 Flash leads Qwen3.8-Flash-Next on 4 of the 5 benchmarks they both report. Only Gemini 3.7 Flash has a verified first-party API price: $0.75 per million input tokens and $3.75 per million output tokens. No pay-as-you-go API rate is tracked for Qwen3.8-Flash-Next. Gemini 3.7 Flash shipped 13 days before Qwen3.8-Flash-Next, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Gemini 3.7 Flash) vs 262k (Qwen3.8-Flash-Next). Gemini 3.7 Flash is closed, while Qwen3.8-Flash-Next is open weight.

On DeepSWE 1.1, Gemini 3.7 Flash leads at 65.3% vs Qwen3.8-Flash-Next at 58.7%. On OSWorld 2.0, Gemini 3.7 Flash leads at 38.1% vs Qwen3.8-Flash-Next at 19.4%. On Agent's Last Exam · pass@1, Gemini 3.7 Flash leads at 26.3% vs Qwen3.8-Flash-Next at 24.3%. On CharXiv Reasoning, Qwen3.8-Flash-Next leads at 84.6% vs Gemini 3.7 Flash at 84.5%. On LVBench, Gemini 3.7 Flash leads at 85.4% vs Qwen3.8-Flash-Next at 76.6%.