Gemini 3.1 ProvsQwen3.6

Gemini 3.1 Pro
Qwen3.6
Specifications
Parameters
35B
Context window
256k
Benchmarks
Nonsense detection
BullshitBench v2
37%
Prompt injection robustness
Gray Swan IPI · k = 1
14.2%
Prompt injection robustness
Gray Swan IPI · k = 10
45.7%
Prompt injection robustness
Gray Swan IPI · k = 15
49.2%
Agentic coding
SWE-Bench Pro
54.2%Best
49.5%
Coding
SWE-Bench Verified
80.6%Best
73.4%
Multilingual coding
SWE-Bench Multilingual
67.2%
Agentic coding
DeepSWE 1.1
12%
ML engineering
MLE-Bench
42.6%
Next.js coding
Next.js Evals
75%
Agentic terminal coding
Terminal-Bench 2.1
70.3%
Agentic terminal coding
Terminal-Bench 2.0
68.5%Best
51.5%
Multi-step tool use
MCP Atlas
78.2%
General tool use
Toolathlon
48.8%
Web browsing
BrowseComp
85.9%
Multidisciplinary reasoning
Humanity's Last Exam · no tools
44.4%Best
21.4%
Multidisciplinary reasoning
Humanity's Last Exam · with tools
51.4%
Abstract reasoning
ARC-AGI-2
77.1%
Advanced math
FrontierMath · Tier 1–3
36.9%
Advanced math
FrontierMath · Tier 4
16.7%
Science
GPQA Diamond
94.3%Best
86%
Agentic computer use
OSWorld-Verified
76.2%
Agentic financial analysis
Finance Agent v2
43%
Knowledge work
GDPval-AA
1314
Knowledge work
GDPval-AA v2
965
Knowledge work
GDPval (win/tie rate)
67.3%
Chart reasoning
CharXiv Reasoning
83.3%Best
78%
Multimodal reasoning
MMMU-Pro
80.5%Best
75.3%
Multimodal
MMMU
81.7%
Spatial reasoning
Blueprint-Bench 2
26.5%
Long context
MRCR v2 (8-needle) · 128k average
84.9%
Long context
MRCR v2 (8-needle) · 1M pointwise
26.3%
Community preference
Arena Elo (Text)
1485
Community preference (code)
Arena Elo (Code)
1445
Overview
CompanyGoogleQwen
Release dateFeb 19 2026Apr 16 2026
AccessProprietaryOpen Weight

Which is better: Gemini 3.1 Pro or Qwen3.6?

Gemini 3.1 Pro leads Qwen3.6 on 7 of the 7 benchmarks they both report. Gemini 3.1 Pro shipped 56 days before Qwen3.6, so benchmark comparisons should account for the intervening progress.

Gemini 3.1 Pro is proprietary, while Qwen3.6 is open weight.

On SWE-Bench Pro, Gemini 3.1 Pro leads at 54.2% vs Qwen3.6 at 49.5%. On SWE-Bench Verified, Gemini 3.1 Pro leads at 80.6% vs Qwen3.6 at 73.4%. On Terminal-Bench 2.0, Gemini 3.1 Pro leads at 68.5% vs Qwen3.6 at 51.5%. On Humanity's Last Exam · no tools, Gemini 3.1 Pro leads at 44.4% vs Qwen3.6 at 21.4%. On GPQA Diamond, Gemini 3.1 Pro leads at 94.3% vs Qwen3.6 at 86%. On CharXiv Reasoning, Gemini 3.1 Pro leads at 83.3% vs Qwen3.6 at 78%. On MMMU-Pro, Gemini 3.1 Pro leads at 80.5% vs Qwen3.6 at 75.3%.

Frequently asked questions

Gemini 3.1 Pro was released by Google on Feb 19 2026.

Qwen3.6 was released by Qwen on Apr 16 2026.

Gemini 3.1 Pro leads on SWE-Bench Pro — Gemini 3.1 Pro 54.2% vs Qwen3.6 49.5%.

Gemini 3.1 Pro leads on Humanity's Last Exam · no tools — Gemini 3.1 Pro 44.4% vs Qwen3.6 21.4%.

Gemini 3.1 Pro is a proprietary model released by Google. Qwen3.6 is an open weight model released by Qwen.

Other comparisons