Gemini 3.1 Flash-Lite
Released
Gemini 3.1 Flash-Lite is an AI model released by Google on Tuesday, Mar 3 2026, 12 days after Gemini 3.1 Pro. Benchmark results (shown below) cover BullshitBench v2.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $0.25
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $0.025
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $1.50
Benchmarks
Compare Gemini 3.1 Flash-Lite with
Suggested comparisons
Frequently asked questions
Gemini 3.1 Flash-Lite was released by Google on Tuesday, Mar 3 2026.
Gemini 3.1 Flash-Lite was built by Google. Builds the Gemini family of models through Google DeepMind. Integrates AI across Google products.
Gemini 3.1 Flash-Lite costs $0.25 per million input tokens and $1.50 per million output tokens through the Google API. Cached input is $0.025 per million tokens. Output prices include reasoning tokens. Rates are pay-as-you-go API prices verified against Google's published pricing on August 18, 2026.
Gemini 3.1 Flash-Lite reports 1 tracked benchmark score — BullshitBench v2: 11%. Scores are the figures published at release by Google.
No. Gemini 3.1 Flash-Lite is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
Google's previous tracked release was Gemini 3.1 Pro on Feb 19 2026, 12 days earlier. It was followed by Gemma 4 on Apr 2 2026.