Gemini 3.6 Flash
Released
Gemini 3.6 Flash is an AI model released by Google on Tuesday, Jul 21 2026, 63 days after Gemini 3.5 Flash. Benchmark results (shown below) cover MLE-Bench, BullshitBench v2, Gray Swan IPI, DeepSWE 1.1, BU Bench, OSWorld-Verified, and 1 more.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $0.75
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $0.075
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $3.75
Benchmarks
Coding
Agentic & tool use
Knowledge work
Robustness
About
Gemini 3.6 Flash, released July 21, 2026 alongside Gemini 3.5 Flash-Lite and the security-focused Gemini 3.5 Flash Cyber, was pitched on efficiency rather than raw scale: Google priced it identically to Gemini 3.5 Flash ($1.50 per million input tokens, $7.50 per million output) while reporting it used roughly 17% fewer output tokens to deliver better results. At launch it posted 49% on DeepSWE 1.1 for long-horizon software engineering — up from 37% for 3.5 Flash — 63.9% on MLE-Bench for machine-learning engineering, a GDPval-AA v2 Elo of 1421 for knowledge work, and 83.0% on OSWorld-Verified computer use.
The release continued Google's pattern of iterating fastest on its workhorse Flash tier, arriving just two months after Gemini 3.5 Flash and jumping the line's numbering to 3.6 while the Pro flagship stayed on 3.1. The pitch — a straight quality upgrade at the exact same cost — targeted the high-volume agentic workloads where Flash had become one of the most heavily used API models. Its turn at the top of the line was the shortest yet: Gemini 3.7 Flash replaced it three weeks later, on August 13, 2026.
Compare Gemini 3.6 Flash with
Suggested comparisons
Frequently asked questions
Gemini 3.6 Flash was released by Google on Tuesday, Jul 21 2026.
Gemini 3.6 Flash was built by Google. Builds the Gemini family of models through Google DeepMind. Integrates AI across Google products.
Gemini 3.6 Flash costs $0.75 per million input tokens and $3.75 per million output tokens through the Google API. Cached input is $0.075 per million tokens. Output prices include reasoning tokens. Rates are pay-as-you-go API prices verified against Google's published pricing on August 18, 2026.
Gemini 3.6 Flash reports 9 tracked benchmark scores — BullshitBench v2: 39%; Gray Swan IPI (k = 1): 7.3%; Gray Swan IPI (k = 10): 32.2%; Gray Swan IPI (k = 15): 37.3%; DeepSWE 1.1: 49%; MLE-Bench: 63.9%; BU Bench: 68%; OSWorld-Verified: 83%; GDPval-AA v2: 1421. Scores are the figures published at release by Google. It holds the best score among all models tracked here on MLE-Bench.
No. Gemini 3.6 Flash is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
Google's previous tracked release was Gemini 3.5 Flash on May 19 2026, 63 days earlier. It was followed by Gemini 3.5 Flash-Lite on Jul 21 2026.