Gemini 3.7 Flash

Gemini 3.7 Flash is an AI model released by Google on Thursday, Aug 13 2026, 23 days after Gemini 3.5 Flash Cyber. It has a 1M token context window. Benchmark results (shown below) cover Humanity's Last Exam (Verified), LAB-Bench 2, Agent's Last Exam, Harvey's Legal Agent Benchmark, GDP.PDF, LVBench, and 12 more.

Benchmarks

Multidisciplinary reasoning
Humanity's Last Exam (Verified)
53.6%
Biology
LAB-Bench 2
82.1%
Agentic computer use
Agent's Last Exam
26.3%
Agentic legal work
Harvey's Legal Agent Benchmark
90.7%
#1
Document comprehension
GDP.PDF
34%
Video understanding
LVBench
85.4%
Long context
MRCR v2 (8-needle)
97%
128k average
62.5%
1M pointwise
128k average
#1
1M pointwise
#1
Agentic coding
DeepSWE 1.1
65.3%
#6 of 16
Agentic coding
FrontierCode v1.1 (Main)
43.6%
main split
#2 of 2
Agentic terminal coding
Terminal-Bench 3.0
14.9%
#3 of 3
Agentic terminal coding
Terminal-Bench 2.1
85.8%
#6 of 24
Biology
BioMysteryBench
43.5%
hard
87.1%
human solved
hard
#2 of 2
human solved
#2 of 2
Agentic computer use
OSWorld 2.0
38.1%
#2 of 2
Business workflows
AutomationBench
30.4%
#2 of 4
Overall intelligence
AA Intelligence Index
56
#2 of 3
Knowledge work
GDPval-AA v2
1525
#8 of 15
Chart reasoning
CharXiv Reasoning
84.5%
#5 of 14
Community preference (code)
Arena Elo (Code)
1588
#6 of 48

About Gemini 3.7 Flash

Gemini 3.7 Flash, released August 13, 2026, arrived three weeks after Gemini 3.6 Flash and moved the biggest numbers of the whole Flash line in coding. At launch it scored 65.3% on DeepSWE v1.1 for long-horizon software engineering, against 49.0% for 3.6 Flash, and 43.6% on the main split of FrontierCode 1.1 — the best figure in Google's launch comparison, ahead of Claude Sonnet 5 and GPT-5.6 Terra. Web development moved with it, to an Arena.ai WebDev Elo of 1588, and the gains extended past code: 30.4% on AutomationBench for enterprise workflows (up from 17.0%), 34.0% on GDP.PDF document comprehension (up from 22.0%), 90.7% on Harvey's LAB-AA legal benchmark, and 97.0% on MRCR v2 128k long-context recall. On the Artificial Analysis Intelligence Index it landed at 56, above Sonnet 5 at 55 and a point below GPT-5.6 Terra and Muse Spark 1.2.

Price was the other half of the pitch. Google launched it at an introductory $0.75 per million input tokens and $3.75 per million output through the end of 2026 — half what 3.6 Flash had cost at its own launch — with the rate reverting to $1.50/$7.50 in January 2027. It shipped with a 1M-token context window and multimodal input covering images, video, audio, and PDFs, across the Gemini API, AI Studio, Antigravity, Android Studio, and Gemini Enterprise, and reached consumers through Spark for AI Pro and Ultra subscribers. The agentic computer-use scores were the exception to the sweep: 38.1% on OSWorld 2.0 and 14.9% on Terminal-Bench 3.0 both trailed GPT-5.6 Terra at release. Three Flash-tier upgrades between May and August, against a Pro flagship untouched since February, made clear which tier Google saw as the centre of its lineup.

Gemini 3.7 Flash — frequently asked questions

When was Gemini 3.7 Flash released?
Gemini 3.7 Flash was released by Google on Thursday, Aug 13 2026.
Who made Gemini 3.7 Flash?
Gemini 3.7 Flash was built by Google. Builds the Gemini family of models through Google DeepMind. Integrates AI across Google products.
What benchmark scores did Gemini 3.7 Flash get?
Gemini 3.7 Flash reports 20 tracked benchmark scores — DeepSWE 1.1: 65.3%; FrontierCode v1.1 (Main) (main split): 43.6%; Terminal-Bench 3.0: 14.9%; Terminal-Bench 2.1: 85.8%; Humanity's Last Exam (Verified): 53.6%; BioMysteryBench (hard): 43.5%; BioMysteryBench (human solved): 87.1%; LAB-Bench 2: 82.1%; OSWorld 2.0: 38.1%; Agent's Last Exam: 26.3%; AutomationBench: 30.4%; Harvey's Legal Agent Benchmark: 90.7%; AA Intelligence Index: 56; GDPval-AA v2: 1525; CharXiv Reasoning: 84.5%; GDP.PDF: 34%; LVBench: 85.4%; MRCR v2 (8-needle) (128k average): 97%; MRCR v2 (8-needle) (1M pointwise): 62.5%; Arena Elo (Code): 1588. Scores are the figures published at release by Google. It holds the best score among all models tracked here on Humanity's Last Exam (Verified), LAB-Bench 2, Agent's Last Exam, Harvey's Legal Agent Benchmark, GDP.PDF, LVBench, MRCR v2 (8-needle) (128k average) and MRCR v2 (8-needle) (1M pointwise).
What is the context window of Gemini 3.7 Flash?
Gemini 3.7 Flash has a context window of 1M. That is the maximum amount of input plus output the model can hold in a single request.
Is Gemini 3.7 Flash open source?
No. Gemini 3.7 Flash is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
What came before and after Gemini 3.7 Flash?
Google's previous tracked release was Gemini 3.5 Flash Cyber on Jul 21 2026, 23 days earlier. It is the most recent Google model tracked on AI Release Tracker.

Compare Gemini 3.7 Flash with