# Gemini 3.6 Flash

Gemini 3.6 Flash is an AI model released by Google on Jul 21 2026. At release it scored 63.9% on MLE-Bench, 39% on BullshitBench v2 and 49% on DeepSWE 1.1.

## Facts

| Field | Value |
| --- | --- |
| Model | Gemini 3.6 Flash |
| Developer | Google |
| Release date | Tuesday, Jul 21 2026 |
| Licensing | Proprietary |

## API pricing

All rates in USD per 1,000,000 tokens, pay-as-you-go.

| Tier | Input | Cached input | Output |
| --- | --- | --- | --- |
| Standard, through 2026-12-31 | $0.75 | $0.075 | $3.75 |

Output prices include reasoning tokens.
Verified August 18, 2026 against the first-party source: https://ai.google.dev/gemini-api/docs/pricing

## Benchmark scores published at release

| Benchmark | Score | Source | What it measures |
| --- | --- | --- | --- |
| BullshitBench v2 | 39% | Lab | Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better. |
| Gray Swan IPI (k = 1) | 7.3% | Lab | Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. Gray Swan's indirect prompt injection benchmark measures how often such an attack succeeds when the attacker gets a single try. Lower is better. |
| Gray Swan IPI (k = 10) | 32.2% | Lab | Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. This variant gives the attacker 10 tries and counts an attack as successful if any of them works. Lower is better. |
| Gray Swan IPI (k = 15) | 37.3% | Lab | Attackers hide malicious instructions inside content the AI reads — a web page, an email, a document — and try to hijack what it does. This variant gives the attacker 15 tries and counts an attack as successful if any of them works. Lower is better. |
| DeepSWE 1.1 | 49% | Lab | Artificial Analysis' independent test of deep, agentic software-engineering work — the AI has to plan and carry out substantial coding tasks end to end. (Version 1.1 of the test.) Higher is better. |
| MLE-Bench | 63.9% | Lab | Can the AI do the work of a machine-learning engineer? It competes in real Kaggle competitions — building, training, and tuning models end to end — and the score reflects how well it places. Higher is better. |
| BU Bench | 68% | Lab | Can the AI drive a real web browser to finish tasks — clicking, filling forms, and navigating sites the way a person would? Run by Browser Use on their BU Bench task set. Higher is better. |
| OSWorld-Verified | 83% | Lab | Can the AI actually operate a computer — clicking, typing, and using real apps — to finish tasks on its own? Higher is better. |
| GDPval-AA v2 | 1421 | Lab | economically valuable knowledge work (v2, re-based Elo) |

## About Gemini 3.6 Flash

Gemini 3.6 Flash, released July 21, 2026 alongside Gemini 3.5 Flash-Lite and the security-focused Gemini 3.5 Flash Cyber, was pitched on efficiency rather than raw scale: Google priced it identically to Gemini 3.5 Flash ($1.50 per million input tokens, $7.50 per million output) while reporting it used roughly 17% fewer output tokens to deliver better results. At launch it posted 49% on DeepSWE 1.1 for long-horizon software engineering — up from 37% for 3.5 Flash — 63.9% on MLE-Bench for machine-learning engineering, a GDPval-AA v2 Elo of 1421 for knowledge work, and 83.0% on OSWorld-Verified computer use.

The release continued Google's pattern of iterating fastest on its workhorse Flash tier, arriving just two months after Gemini 3.5 Flash and jumping the line's numbering to 3.6 while the Pro flagship stayed on 3.1. The pitch — a straight quality upgrade at the exact same cost — targeted the high-volume agentic workloads where Flash had become one of the most heavily used API models. Its turn at the top of the line was the shortest yet: Gemini 3.7 Flash replaced it three weeks later, on August 13, 2026.

## Questions and answers

### When was Gemini 3.6 Flash released?

Gemini 3.6 Flash was released by Google on Tuesday, Jul 21 2026.

### Who made Gemini 3.6 Flash?

Gemini 3.6 Flash was built by Google. Builds the Gemini family of models through Google DeepMind. Integrates AI across Google products.

### How much does Gemini 3.6 Flash cost?

Gemini 3.6 Flash costs $0.75 per million input tokens and $3.75 per million output tokens through the Google API. Cached input is $0.075 per million tokens. Output prices include reasoning tokens. Rates are pay-as-you-go API prices verified against Google's published pricing on August 18, 2026.

### What benchmark scores did Gemini 3.6 Flash get?

Gemini 3.6 Flash reports 9 tracked benchmark scores — BullshitBench v2: 39%; Gray Swan IPI (k = 1): 7.3%; Gray Swan IPI (k = 10): 32.2%; Gray Swan IPI (k = 15): 37.3%; DeepSWE 1.1: 49%; MLE-Bench: 63.9%; BU Bench: 68%; OSWorld-Verified: 83%; GDPval-AA v2: 1421. Scores are the figures published at release by Google. It holds the best score among all models tracked here on MLE-Bench.

### Is Gemini 3.6 Flash open source?

No. Gemini 3.6 Flash is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.

### What came before and after Gemini 3.6 Flash?

Google's previous tracked release was Gemini 3.5 Flash on May 19 2026, 63 days earlier. It was followed by Gemini 3.5 Flash-Lite on Jul 21 2026.


---

Canonical page: https://aireleasetracker.com/model/google/gemini-3.6-flash
Full dataset: https://aireleasetracker.com/llms-full.txt · JSON: https://aireleasetracker.com/models.json
Source: AI Release Tracker (https://aireleasetracker.com). Benchmark scores are the figures published by the releasing lab at launch.
