Identity
- Status
- active
- Developer ID
gemini-2.5-flash- Released
- Not reported
- Version
- Gemini 2.5 Flash
- License
- Unknown
Gemini 2.5 Flash is Google DeepMind's stable price-performance model for low-latency reasoning workloads.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M | Observed | Evidence |
|---|---|---|---|---|---|---|
Google AIAvailable | $0.30 | $2.50 | $0.030 | — | Aug 29, 2026 | Provider-reported |
| Gemini 2.5 FlashGoogle DeepMind | Berkeley Function Calling Leaderboardv4-ede5081a24bc | 56.24 | — | — | Gemini-2.5-Flash (FC)unknown | Recomputedwinner eligible | Third-party benchmark |
| Gemini 2.5 FlashGoogle DeepMind | Berkeley Function Calling Leaderboardv4-ede5081a24bc | 50.90 | — | — | Gemini-2.5-Flash (Prompt)unknown | Recomputedwinner eligible | Third-party benchmark |
| Gemini 2.5 FlashGoogle DeepMind | LMArena Text Arenatext-2026-09-01-011508720696 | 1,417.31 | — | 122,770 | gemini-2.5-flash; 95% CI [1414.86923843, 1419.74079237]; votes 122770; rank 111unknown | Original sourcewinner eligible | Third-party benchmark |
| Gemini 2.5 FlashGoogle DeepMind | LMArena Vision Arenavision-2026-08-27-011508720696 | 1,235.76 | — | 58,968 | gemini-2.5-flash; 95% CI [1231.05290901, 1240.46508391]; votes 58968; rank 59unknown | Original sourcewinner eligible | Third-party benchmark |
| Gemini 2.5 FlashGoogle DeepMind | ToneBench2026-08-14-10-task | 73.83 | 4,339 | 50 | Gemini 2.5 Flashaverage per case | Original sourcewinner eligible | Third-party benchmark |
At a Glance
gemini-2.5-flashThis page represents one immutable developer model ID. Serving endpoints, pricing, regions, quantizations, and service tiers attach separately and never create duplicate model records.
Primary Evidence
Comparable Peers
| A | Pair | B | Context |
|---|---|---|---|
Gemini 2.5 FlashGoogle DeepMind | vs | Gemini 3.5 FlashGoogle DeepMind | family variantstext |
Gemini 2.5 FlashGoogle DeepMind | vs | Gemini 3.6 FlashGoogle DeepMind | family variantstext |
Gemini 2.5 FlashGoogle DeepMind | vs | Gemini 3.5 Flash-LiteGoogle DeepMind | family variantstext |
Gemini 2.5 FlashGoogle DeepMind | vs | Qwen3.8-FlashQwen | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | Qwen3.7-PlusQwen | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | Gemini Computer UseGoogle DeepMind | family variantstext |
Gemini 2.5 FlashGoogle DeepMind | vs | GPT-5.6 TerraOpenAI | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | GPT-4.1OpenAI | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | GPT-4o MiniOpenAI | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | DeepSeek-V4-FlashDeepSeek | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | DeepSeek-V4-ProDeepSeek | cross-developer peerstext |
Gemini 2.5 FlashGoogle DeepMind | vs | DeepSeek-V4-Flash-BaseDeepSeek | cross-developer peerstext |
Questions
Gemini 2.5 Flash is Google DeepMind's stable price-performance model for low-latency reasoning workloads. It is developed by Google DeepMind and its current sourced lifecycle status is active.
Gemini 2.5 Flash's current record lists Text, Image, Video, and Audio as input and Text as output.
The current record lists a 1,048,576-token context window and a maximum output of 65,536 tokens.
Gemini 2.5 Flash is marked as API-available. The directly sourced provider records currently include Google AI.
The lowest directly sourced prices currently attached to Gemini 2.5 Flash are $0.30 per million input tokens through Google AI and $2.50 per million output tokens through Google AI. Prices are provider-specific and should be checked against each cited observation date.
Gemini 2.5 Flash's recorded capabilities are chat, generation, reasoning, and tools. Its supported tasks are audio, image, text, and video.
Gemini 2.5 Flash is not marked as an open-weight model. Self-hosting is marked as unsupported, with no license recorded.
No predecessor is recorded. No successor is recorded. Recorded aliases are gemini-2.5-flash.
The current compatible peer set includes Gemini 3.5 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Qwen3.8-Flash, and Qwen3.7-Plus. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons →
Gemini 2.5 Flash was last verified Aug 29, 2026 from Google DeepMind's primary source, “Gemini 2.5 Flash | Gemini API.” The evidence is classified as official fact.