Identity
- Status
- active
- Developer ID
gemini-2.5-flash-lite- Released
- Not reported
- Version
- Gemini 2.5 Flash-Lite
- License
- Unknown
Gemini 2.5 Flash-Lite is Google DeepMind's stable small model for high-volume use.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M | Observed | Evidence |
|---|---|---|---|---|---|---|
Google AIAvailable | $0.10 | $0.40 | $0.010 | — | Aug 29, 2026 | Provider-reported |
| Gemini 2.5 Flash-LiteGoogle DeepMind | Berkeley Function Calling Leaderboardv4-ede5081a24bc | 36.87 | — | — | Gemini-2.5-Flash-Lite (FC)unknown | Recomputedwinner eligible | Third-party benchmark |
| Gemini 2.5 Flash-LiteGoogle DeepMind | Berkeley Function Calling Leaderboardv4-ede5081a24bc | 28.03 | — | — | Gemini-2.5-Flash-Lite (Prompt)unknown | Recomputedwinner eligible | Third-party benchmark |
| Gemini 2.5 Flash-LiteGoogle DeepMind | ToneBench2026-08-14-10-task | 68.49 | 2,405 | 50 | Gemini 2.5 Flash-Liteaverage per case | Original sourcewinner eligible | Third-party benchmark |
At a Glance
gemini-2.5-flash-liteThis page represents one immutable developer model ID. Serving endpoints, pricing, regions, quantizations, and service tiers attach separately and never create duplicate model records.
Primary Evidence
Comparable Peers
| A | Pair | B | Context |
|---|---|---|---|
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Gemini 3.1 Flash-LiteGoogle DeepMind | family variantstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Gemini 3.5 Flash-LiteGoogle DeepMind | family variantstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Gemini 3.5 FlashGoogle DeepMind | family variantstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Gemini 3.6 FlashGoogle DeepMind | family variantstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Nova 2 LiteAmazon | cross-developer peerstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Qwen3.8-FlashQwen | cross-developer peerstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | GPT-5.6 TerraOpenAI | cross-developer peerstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | Claude Sonnet 5Anthropic | cross-developer peerstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | GPT-4.1 NanoOpenAI | cross-developer peerstext |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | cross-developer peerstext | |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | cross-developer peerstext | |
Gemini 2.5 Flash-LiteGoogle DeepMind | vs | GLM-5.3Z.ai | cross-developer peerstext |
Questions
Gemini 2.5 Flash-Lite is Google DeepMind's stable small model for high-volume use. It is developed by Google DeepMind and its current sourced lifecycle status is active.
Gemini 2.5 Flash-Lite's current record lists Text, Image, Video, Audio, and Document as input and Text as output.
The current record lists a 1,048,576-token context window and a maximum output of 65,536 tokens.
Gemini 2.5 Flash-Lite is marked as API-available. The directly sourced provider records currently include Google AI.
The lowest directly sourced prices currently attached to Gemini 2.5 Flash-Lite are $0.10 per million input tokens through Google AI and $0.40 per million output tokens through Google AI. Prices are provider-specific and should be checked against each cited observation date.
Gemini 2.5 Flash-Lite's recorded capabilities are chat, generation, reasoning, and tools. Its supported tasks are audio, document, image, text, and video.
Gemini 2.5 Flash-Lite is not marked as an open-weight model. Self-hosting is marked as unsupported, with no license recorded.
No predecessor is recorded. No successor is recorded. Recorded aliases are gemini-2.5-flash-lite.
The current compatible peer set includes Gemini 3.1 Flash-Lite, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash, Gemini 3.6 Flash, and Nova 2 Lite. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons →
Gemini 2.5 Flash-Lite was last verified Aug 29, 2026 from Google DeepMind's primary source, “Gemini 2.5 Flash-Lite | Gemini API.” The evidence is classified as official fact.