Gemini 3.5 Flash-Lite
Released
Gemini 3.5 Flash-Lite is an AI model released by Google on Tuesday, Jul 21 2026, the same day as Gemini 3.6 Flash. Benchmark results (shown below) cover BullshitBench v2, SWE-Bench Pro, Terminal-Bench 2.1, BU Bench, OSWorld-Verified, and GDPval-AA v2.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $0.30
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $0.03
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $2.50
Benchmarks
Coding
Terminal & CLI
Agentic & tool use
Knowledge work
Robustness
Compare Gemini 3.5 Flash-Lite with
Suggested comparisons
Frequently asked questions
Gemini 3.5 Flash-Lite was released by Google on Tuesday, Jul 21 2026.
Gemini 3.5 Flash-Lite was built by Google. Builds the Gemini family of models through Google DeepMind. Integrates AI across Google products.
Gemini 3.5 Flash-Lite costs $0.30 per million input tokens and $2.50 per million output tokens through the Google API. Cached input is $0.03 per million tokens. Output prices include reasoning tokens. Rates are pay-as-you-go API prices verified against Google's published pricing on August 18, 2026.
Gemini 3.5 Flash-Lite reports 6 tracked benchmark scores — BullshitBench v2: 65%; SWE-Bench Pro: 54.2%; Terminal-Bench 2.1: 54%; BU Bench: 49%; OSWorld-Verified: 74%; GDPval-AA v2: 1140. Scores are the figures published at release by Google.
No. Gemini 3.5 Flash-Lite is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
Google's previous tracked release was Gemini 3.6 Flash on Jul 21 2026. It was followed by Gemini 3.5 Flash Cyber on Jul 21 2026.