Gemini 2.5 Flash-Lite
Released
Gemini 2.5 Flash-Lite is an AI model released by Google on Tuesday, Jun 17 2025, 61 days after Gemini 2.5 Flash.
API pricing
- InputWhat you pay for everything you send the model — your question, plus any documents or earlier conversation you include with it.
- $0.10
- Cached inputA reduced rate for text you send over and over. If every request starts with the same instructions or the same document, the provider keeps a copy ready and charges less to read it again.
- $0.01
- OutputWhat you pay for the text the model writes back. It is normally the dearer half: producing an answer costs more than reading one.
- $0.40
Compare Gemini 2.5 Flash-Lite with
Suggested comparisons
Frequently asked questions
Gemini 2.5 Flash-Lite was released by Google on Tuesday, Jun 17 2025.
Gemini 2.5 Flash-Lite was built by Google. Builds the Gemini family of models through Google DeepMind. Integrates AI across Google products.
Gemini 2.5 Flash-Lite costs $0.10 per million input tokens and $0.40 per million output tokens through the Google API. Cached input is $0.01 per million tokens. Output prices include reasoning tokens. Rates are pay-as-you-go API prices verified against Google's published pricing on August 18, 2026.
No. Gemini 2.5 Flash-Lite is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
Google's previous tracked release was Gemini 2.5 Flash on Apr 17 2025, 61 days earlier. It was followed by Gemini 3.0 Pro on Nov 18 2025.