Gemini 3.5 FlashvsGPT-6 Sol

Gemini 3.5 Flash
GPT-6 Sol
Specifications
Context window
1.05M
API pricing
Input price
$1.50$2.00
Output price
$9.00$10.00
Cached input price
$0.15$0.20
Cheapest input
$0.75Google
Cheapest output
$4.50Google
Benchmarks
DeepSWE 1.1
37%68.8%
Benchmarks
BullshitBench v2
20%
Gray Swan IPI · k = 1
14.1%
Gray Swan IPI · k = 10
54.2%
Gray Swan IPI · k = 15
60.5%
ProgramBench
0%
SWE-Bench Pro
55.1%
MLE-Bench
49.7%
Terminal-Bench 2.1
76.2%
MCP Atlas
83.6%
Toolathlon
56.5%
BU Bench
58%
Humanity's Last Exam · no tools
40.2%
Humanity's Last Exam · with tools
40.2%
ARC-AGI-2
72.1%
OSWorld 2.0
60.5%
OSWorld-Verified
78.4%
Agent's Last Exam · pass@1
56.4%
AutomationBench
33.2%
Finance Agent v2
57.9%
GDPval-AA
1656
GDPval-AA v2
1349
CharXiv Reasoning
84.2%
MMMU-Pro
83.6%
Blueprint-Bench 2
33.6%
MRCR v2 (8-needle) · 128k average
77.3%
MRCR v2 (8-needle) · 1M pointwise
26.6%
Overview
CompanyGoogleOpenAI
Release dateMay 19 2026Sep 22 2026
AccessProprietaryProprietary

Other comparisons

Gemini 3.5 FlashvsClaude Opus 5.5GPT-6 SolvsClaude Opus 5.5Gemini 3.5 FlashvsMuse Spark 1.3GPT-6 SolvsMuse Spark 1.3Gemini 3.5 FlashvsGrok 4.7GPT-6 SolvsGrok 4.7Gemini 3.5 FlashvsDeepSeek-V4.1-FlashGPT-6 SolvsDeepSeek-V4.1-FlashGemini 3.5 FlashvsMistral Medium 3.5GPT-6 SolvsMistral Medium 3.5Gemini 3.5 FlashvsKimi K3GPT-6 SolvsKimi K3

Frequently asked questions

GPT-6 Sol leads Gemini 3.5 Flash on 1 of the 1 benchmark they both report (DeepSWE 1.1). Gemini 3.5 Flash is cheaper on both input and output: $1.50 vs $2.00 per million input tokens, and $9.00 vs $10.00 per million output tokens. Figures are base-tier rates. Gemini 3.5 Flash shipped 126 days before GPT-6 Sol, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On DeepSWE 1.1, GPT-6 Sol leads at 68.8% vs Gemini 3.5 Flash at 37%.