Gemini 3.7 FlashvsQwen3.5

Gemini 3.7 Flash
Qwen3.5
Specifications
Parameters
397B
Context window
1M1M
API pricing
Input price
$0.75$0.60
Output price
$3.75$3.60
Cached input price
$0.075
Cheapest input
$0.08Darkbloom
Cheapest output
$0.13Darkbloom
Benchmarks
CharXiv Reasoning
84.5%80.8%
Benchmarks
BullshitBench v2
35%
ProgramBench
0%
SWE-Bench Verified
76.4%
SWE-Bench Multilingual
69.3%
DeepSWE 1.1
65.3%
FrontierCode v1.1 (Main) · main split
43.6%
Terminal-Bench 4.0
11.21%
Terminal-Bench 3.0
14.9%
Terminal-Bench 2.1
85.8%
Terminal-Bench 2.0
52.5%
BrowseComp
69%
Humanity's Last Exam · no tools
28.7%
Humanity's Last Exam (Verified)
53.6%
BioMysteryBench · hard
43.5%
BioMysteryBench · human solved
87.1%
LAB-Bench 2
82.1%
GPQA Diamond
88.4%
OSWorld 2.0
38.1%
OSWorld-Verified
62.2%
Agent's Last Exam · pass@1
26.3%
AutomationBench
30.4%
Harvey's Legal Agent Benchmark
8.8%
AA Intelligence Index
56
GDPval-AA v2
1525
GDP.PDF
34%
LVBench
85.4%
MMMU-Pro
79%
MMMU
85%
MRCR v2 (8-needle) · 128k average
97%
MRCR v2 (8-needle) · 1M pointwise
62.5%
threejseval
1526
Overview
CompanyGoogleQwen
Release dateAug 13 2026Feb 16 2026
AccessProprietaryOpen Weight

Other comparisons

Gemini 3.7 FlashvsClaude Fable 5.1Qwen3.5vsClaude Fable 5.1Gemini 3.7 FlashvsGPT-6 AstraQwen3.5vsGPT-6 AstraGemini 3.7 FlashvsMuse Spark 1.3Qwen3.5vsMuse Spark 1.3Gemini 3.7 FlashvsGrok 4.6Qwen3.5vsGrok 4.6Gemini 3.7 FlashvsDeepSeek-V4.1-FlashQwen3.5vsDeepSeek-V4.1-FlashGemini 3.7 FlashvsMistral Medium 3.5Qwen3.5vsMistral Medium 3.5

Frequently asked questions

Gemini 3.7 Flash leads Qwen3.5 on 1 of the 1 benchmark they both report (CharXiv Reasoning). Qwen3.5 is cheaper on both input and output: $0.60 vs $0.75 per million input tokens, and $3.60 vs $3.75 per million output tokens. Qwen3.5 shipped 178 days before Gemini 3.7 Flash, so benchmark comparisons should account for the intervening progress.

Context windows are 1M (Gemini 3.7 Flash) vs 1M (Qwen3.5). Gemini 3.7 Flash is proprietary, while Qwen3.5 is open weight.

On CharXiv Reasoning, Gemini 3.7 Flash leads at 84.5% vs Qwen3.5 at 80.8%.